When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for…
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0
AI-accelerated End-to-End Framework for Rapid Professional Upskilling
Policy of Thoughts: Scaling Test-Time Training for LLM Reasoning via Online Policy Evolut…
Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models
Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level A…
SemaDiff: Identifying Semantic-Changing Commits with Generated Code and Tests
Tabular Foundation Models for Discrete Choice Estimation
CAS I: A Geometric Coding Theorem
FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents
Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychi…
Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate
From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of …
A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Mul…
Early Adoption of Agentic Coding Tools by GitHub Projects
Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth
The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Ne…
Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Pract…
LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents
Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open…
Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for D…
Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streami…
Autonomous UAV Route Planning for Coverage Maximization in Environmental Monitoring: A Sy…
Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation
The Hitchhiker's Guide to Monoculture
Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative E…
Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs
HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Larg…
A Hybrid Mamba for Audio-Visual Navigation
Removable Defects: The Economics and Limits of Deliberate Deficiency