Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
Explorar
Noticias de IA
29647 elementos — filtrados, clasificados y sin duplicados
When Do Attention Circuits Form? Developmental Trajectories of Capability and Attention-S…
Evolutionary Discovery of Bivariate Bicycle Codes with LLM-Guided Search
Policy and World Modeling Co-Training for Language Agents
AutoForest: Automatically Generating Forest Plots from Biomedical Studies with End-to-End…
PaSBench-Video: A Streaming Video Benchmark for Proactive Safety Warning
MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence
Ghost Tool Calls: Issue-Time Privacy for Speculative Agent Tools
Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events
SimSD: Simple Speculative Decoding in Diffusion Language Models
From Layers to Submodules: Rethinking Granularity in Replacement-Based LLM Compression
AdaCodec: A Predictive Visual Code for Video MLLMs
Algebraic anti-unification
Explainable AI Through a Democratic Lens: DhondtXAI for D'Hondt-Projected Feature Attribu…
Safety Must Precede the Deployment of Open-Ended AI
Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Sca…
Agent Guide: A Simple Agent Behavioral Watermarking Framework
Language Model Networks: Supervision-Efficient Learning through Dense Communication
Formally Solving Answer-Construction Problems in Lean
On the Theoretical Limitations of Embedding-based Link Prediction
Query Circuits: Explaining How Language Models Answer User Prompts
REBot: From RAG to CatRAG with Semantic Enrichment and Graph Routing
Addressing Longstanding Challenges in Cognitive Science with Language Models
LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services
On the Collapse of Generative Paths: A Criterion and Correction for Diffusion Steering
Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention
Unplugging a Seemingly Sentient Machine Is the Rational Choice -- A Metaphysical Perspect…
PolarMem: A Training-Free Polarized Latent Graph Memory for Verifiable Vision-Language Mo…
The Refusal--Compliance Tradeoff: A Large-Scale Safety Behavior Audit of Large Language M…
Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge