PriorProof: A Point-in-Time Measure of Technique Novelty for Formal Proofs
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Expected Free Energy as Belief-Dependent Utility for rho-POMDPs
Lomekwi: Resource-Bounded Tool Discovery in LLM Agents
Environment-free Synthetic Data Generation for API-Calling Agents
From Overload to Insights: How AI Agents Can Support Scientists in Analyzing Complex Data
RELIC: Revealed Principles for Learning Interpretable Composable Skills in Multi-Agent Pl…
Constraint-Anchored Reasoning Traces
RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts
DS@GT ARC at eRisk 2026: Hybrid Multi-Agent LLM System with Structured Algorithmic Guidan…
Diversity-Oriented Fine-Tuning for Uncertainty-Based Hallucination Detection
LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Mode…
Berkeley and Heiserman as an Unexhausted Architecture for Embodied Machine Intelligence
Generalist AI Control: Towards Multi-purpose Adaptive Algorithms
SelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation
Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers
Accurate and Efficient Long-Term Memory for LLM Agents
A Survey on the Verification of Reinforcement Learning Policies
Shapley Context Pruning: A Cooperative Game Perspective for Context Reranking and Pruning
ColGraphRAG: Late-Interaction Evidence Retrieval for Multimodal GraphRAG
PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Covera…
It Takes 8 Tokens: Weak-to-Strong Off-Policy RL via Auxiliary Branches
Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Age…
Show Me How You Reason and I'll Tell You Who You Are: Reasoning Graphs for Robust LLM Aut…
Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security
Just A Rather Very Intelligent Spoken Agent
A Research Prototype for Closed-Loop Generative Design of Customized Foot Orthoses via Se…
Exact Network Surgery: Functional Invariance and Gradient Plasticity in Reactive Computat…
Training Continuous Chain of Thought Models: A Tale of Two Regimes
DeLIVeR: Decomposed Learning for Information-grounded Veracity Recognition via Reinforced…
CommitLLM: A Fine-Tuned Pipeline for Git Commit Message Generation