Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models
Explorar
Noticias de IA
21813 elementos — filtrados, clasificados y sin duplicados
PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates
A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination
Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent
Towards a satellite image manipulation and deepfake localization benchmark dataset
Toward Integrating Adaptive Experience Replay and Online Uncertainty Estimation in Safe A…
RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists
What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend
SimMOF: AI agent for Automated MOF Simulations
Personalized Federated Sparse Adaptation of Time-Series Foundation Models
AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in…
Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Cali…
The Order Is the Guarantee: Verifier-Budgeted Code Deletion with Static-First Learned Pro…
Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks
Masked diffusion enables coherent beat tracking
RAG-Stack: Co-Optimizing RAG Serving Performance and Quality
EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive De…
Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark
Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Questio…
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large L…
Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based…
Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Medi…
EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and …
CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical V…
TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction
AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluati…
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Dis…
Generative Optimization for Incentivized Advertising with Global Level Constraints