Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation
Explorar
Noticias de IA
29349 elementos — filtrados, clasificados y sin duplicados
A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Aud…
Toward Integrating Adaptive Experience Replay and Online Uncertainty Estimation in Safe A…
PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates
Personalized Federated Sparse Adaptation of Time-Series Foundation Models
Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Cali…
What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend
Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning
VQ-VAD: Vector-quantized Motion Representation Learning for Human-centric Video Anomaly D…
Breadcrumbing Search Agents
When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large L…
Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks
The Order Is the Guarantee: Verifier-Budgeted Code Deletion with Static-First Learned Pro…
Masked diffusion enables coherent beat tracking
EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and …
AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluati…
Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Medi…
Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based…
CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical V…
EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive De…
MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages
Capability-Gated Planning: Cost-to-Goal Discovery and the Limits of Myopic Experiment Sel…
TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction
ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Gene…
Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
Generative Optimization for Incentivized Advertising with Global Level Constraints
DisMix: Order-Aware Mixup for Medical Imaging via Disentangling Ordinal and Non-Ordinal F…
Equitable System-Prompt Selection via Constrained Mixed-Strategy GroupDRO
HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework …