LightEMMA: A Longitudinal Evaluation of Vision-Language Models for Autonomous Driving
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
LDC: Learning to Generate Research Idea with Dynamic Control
AgentRM: Enhancing Agent Generalization with Reward Modeling
Decentralized Vision-Based Autonomous Aerial Wildlife Monitoring
Discovering High Level Patterns from Simulation Traces
CORAL: Towards Autonomous Multi-Agent Evolution for Open-Ended Discovery
PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Lev…
Complete Identification of Deep ReLU Networks through {\L}ukasiewicz Logic
RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constr…
WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquit…
Not All Preferences Deserve Gradients: Understanding Gradient Utility in Offline Reasonin…
Knowledge Acquisition During Pre-training? Large Language Models Learn Better With Auxili…
One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing
A Non-Formulable Theorem: A Fundamental Limit of Finite Syntactic Systems and Its Consequ…
Sequential Beats Joint: On the Interplay between On-Policy Distillation and RLVR
ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize
Grammar-Aligned Decoding
Influence of Extruded Filament Shape on Buildability in 3D Concrete Printing: A Geometry-…
When Models Edit Too Much: On the Fidelity of Minimal Code Edits
Catalogue Photography as a Cold Start: Toward Deployable Carbide Burr Recognition
The Blind Spot in 2D Infants' Pose Estimation:Robust Learning from Noisy Annotations
Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO
Investigating the Ability of Large Language Models to Analyze Recipes for Diabetes
FWBC-VLA: Force-Aware Whole-Body Compensation for Contact-Rich Loco-Manipulation
GraFT: A Training-Free Framework for Spatial Reasoning in Multimodal Large Language Model…
RATL: Learning from Retrieved Residuals for Robust Multivariate Time-Series Forecasting
TAP-Path: Task-Adaptive Structural and Token Pruning for Efficient and Trustworthy Pathol…
Witnesses Explain Anomalies
The impact of phase information for few-shot fine-grained image classification
Beyond BLEU: A Case for Redefining Sign Language Translation Benchmarks