BenGER Platform: A Collaborative Web Platform for End-to-End Benchmarking of German Legal…
Explorar
Noticias de IA
30321 elementos — filtrados, clasificados y sin duplicados
Retention Consequence in Lifecycle Memory Control
C-MORAL: Controllable Multi-Objective Molecular Optimization with Reinforcement Alignment…
Data-Efficient On-Policy Distillation for Automatic Speech Recognition
The Forensic Cost of Watermark Removal: From Dedicated Attacks to Image Editing
Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in…
Voice "Cloning" is Style Transfer
Detecting and Mitigating the Correct-Answer Extinction Window in Test-Time Reinforcement …
You Are in Control of Your State: Why Human Outcomes Are Controllable Through Causal Stat…
Hierarchical Prompt-Domain Control and Learning for Resource-Constrained Agentic Language…
SkillGrad: Optimizing Agent Skills Like Gradient Descent
GraD-IBD: Graph Representation Learning from Diagnosis Trajectories for Early Detection o…
C-MIG: Multi-view Information Gain-based Retrieval-Augmented Generation for Clinical Diag…
CubePart: An Open-Vocabulary Part-Controllable 3D Generator
Ligand-Conditioned Discrete Diffusion for Protein Sequence-Structure Co-Design
From Talking to Singing: A New Challenge for Audio-Visual Deepfake Detection
MIRA: A Bilingual Benchmark for Medical Information Response Audit
MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Aud…
Verifiable Benchmarking of Long-Horizon Spatial Biology
Deconstructing Spatial Complexity: Hierarchical Decomposition for LLM Spatial Reasoning
Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning
The Illusion of Opting in AI-Mediated Consequential Decisions
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Questi…
Global Policy-Space Response Oracles for Two-Player Zero-Sum Games
Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR
SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Mode…
DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes
Diffusion Large Language Models for Visual Speech Recognition
Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents
Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design P…