Explorar

Noticias de IA

30934 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
PriorProof: A Point-in-Time Measure of Technique Novelty for Formal Proofs
arXiv cs.AI Research & Papers
Expected Free Energy as Belief-Dependent Utility for rho-POMDPs
arXiv cs.AI Research & Papers
Lomekwi: Resource-Bounded Tool Discovery in LLM Agents
arXiv cs.AI Research & Papers
Environment-free Synthetic Data Generation for API-Calling Agents
arXiv cs.AI Research & Papers
From Overload to Insights: How AI Agents Can Support Scientists in Analyzing Complex Data
arXiv cs.AI Research & Papers
RELIC: Revealed Principles for Learning Interpretable Composable Skills in Multi-Agent Pl…
arXiv cs.AI Research & Papers
Constraint-Anchored Reasoning Traces
arXiv cs.AI Research & Papers
RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts
arXiv cs.AI Research & Papers
DS@GT ARC at eRisk 2026: Hybrid Multi-Agent LLM System with Structured Algorithmic Guidan…
arXiv cs.AI Research & Papers
Diversity-Oriented Fine-Tuning for Uncertainty-Based Hallucination Detection
arXiv cs.AI Research & Papers
LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Mode…
arXiv cs.AI Research & Papers
Berkeley and Heiserman as an Unexhausted Architecture for Embodied Machine Intelligence
arXiv cs.AI Research & Papers
Generalist AI Control: Towards Multi-purpose Adaptive Algorithms
arXiv cs.AI Research & Papers
SelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation
arXiv cs.AI Research & Papers
Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers
arXiv cs.AI Research & Papers
Accurate and Efficient Long-Term Memory for LLM Agents
arXiv cs.AI Research & Papers
A Survey on the Verification of Reinforcement Learning Policies
arXiv cs.AI Research & Papers
Shapley Context Pruning: A Cooperative Game Perspective for Context Reranking and Pruning
arXiv cs.AI Research & Papers
ColGraphRAG: Late-Interaction Evidence Retrieval for Multimodal GraphRAG
arXiv cs.AI Research & Papers
PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Covera…
arXiv cs.AI Research & Papers
It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches
arXiv cs.AI Research & Papers
Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Age…
arXiv cs.AI Research & Papers
Show Me How You Reason and I'll Tell You Who You Are: Reasoning Graphs for Robust LLM Aut…
arXiv cs.AI Security & Safety
Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security
arXiv cs.AI Research & Papers
Just A Rather Very Intelligent Spoken Agent
arXiv cs.AI Research & Papers
A Research Prototype for Closed-Loop Generative Design of Customized Foot Orthoses via Se…
arXiv cs.AI Research & Papers
Exact Network Surgery: Functional Invariance and Gradient Plasticity in Reactive Computat…
arXiv cs.AI Research & Papers
Training Continuous Chain of Thought Models: A Tale of Two Regimes
arXiv cs.AI Research & Papers
DeLIVeR: Decomposed Learning for Information-grounded Veracity Recognition via Reinforced…
arXiv cs.AI Research & Papers
CommitLLM: A Fine-Tuned Pipeline for Git Commit Message Generation