Explorar

Noticias de IA

22318 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for…
arXiv cs.AI Research & Papers
Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0
arXiv cs.AI Research & Papers
AI-accelerated End-to-End Framework for Rapid Professional Upskilling
arXiv cs.AI Research & Papers
Policy of Thoughts: Scaling Test-Time Training for LLM Reasoning via Online Policy Evolut…
arXiv cs.AI Research & Papers
Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models
arXiv cs.AI Research & Papers
Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level A…
arXiv cs.AI Research & Papers
SemaDiff: Identifying Semantic-Changing Commits with Generated Code and Tests
arXiv cs.AI Research & Papers
Tabular Foundation Models for Discrete Choice Estimation
arXiv cs.AI Research & Papers
CAS I: A Geometric Coding Theorem
arXiv cs.AI Research & Papers
FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents
arXiv cs.AI Research & Papers
Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychi…
arXiv cs.AI Research & Papers
Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate
arXiv cs.AI Research & Papers
From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of …
arXiv cs.AI Research & Papers
A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Mul…
arXiv cs.AI Research & Papers
Early Adoption of Agentic Coding Tools by GitHub Projects
arXiv cs.AI Research & Papers
Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth
arXiv cs.AI Research & Papers
The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Ne…
arXiv cs.AI Research & Papers
Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Pract…
arXiv cs.AI Research & Papers
LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents
arXiv cs.AI Research & Papers
Federated Explainable Artificial Intelligence: Roles, Architectures, Evaluation, and Open…
arXiv cs.AI Research & Papers
Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for D…
arXiv cs.AI Research & Papers
Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streami…
arXiv cs.AI Research & Papers
Autonomous UAV Route Planning for Coverage Maximization in Environmental Monitoring: A Sy…
arXiv cs.AI Research & Papers
Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation
arXiv cs.AI Research & Papers
The Hitchhiker's Guide to Monoculture
arXiv cs.AI Research & Papers
Discovering Ordinary Differential Equations with LLM-Based Qualitative and Quantitative E…
arXiv cs.AI Research & Papers
Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs
arXiv cs.AI Research & Papers
HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Larg…
arXiv cs.AI Research & Papers
A Hybrid Mamba for Audio-Visual Navigation
arXiv cs.AI Research & Papers
Removable Defects: The Economics and Limits of Deliberate Deficiency