Accurate and Efficient Long-Term Memory for LLM Agents
Explorar
Noticias de IA
22116 elementos — filtrados, clasificados y sin duplicados
CIGPO: Contextual Information-Gain Policy Optimization for Multi-Turn Evidence-Reading LL…
Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers
SelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation
Generalist AI Control: Towards Multi-purpose Adaptive Algorithms
RobustMAD: Evaluating Real-World Robustness of Multimodal Small Language Models for Deplo…
Berkeley and Heiserman as an Unexhausted Architecture for Embodied Machine Intelligence
LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Mode…
TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment
Diversity-Oriented Fine-Tuning for Uncertainty-Based Hallucination Detection
DS@GT ARC at eRisk 2026: Hybrid Multi-Agent LLM System with Structured Algorithmic Guidan…
RECON: Benchmarking Agent Memory for Compositional Reasoning over Long Contexts
Constraint-Anchored Reasoning Traces
Fully-sensorized smart-eyewear platform for on-device Machine Learning
RELIC: Revealed Principles for Learning Interpretable Composable Skills in Multi-Agent Pl…
From Overload to Insights: How AI Agents Can Support Scientists in Analyzing Complex Data
What Makes Linguistic Representations Good Models of High-Level Visual Perception in the …
Environment-free Synthetic Data Generation for API-Calling Agents
Lomekwi: Resource-Bounded Tool Discovery in LLM Agents
Expected Free Energy as Belief-Dependent Utility for rho-POMDPs
PriorProof: A Point-in-Time Measure of Technique Novelty for Formal Proofs
Reward-Driven LLM Agent Workflows: Synthesizing POMDP Routing and Self-Correction for Aut…
When LLMs Over-Answer: Measuring and Mitigating Quality Issues in LLM-Based Hardware Desc…
Bridging the Information Gap: Semantic Densification and Hindsight Distillation for Cold-…
Fourier Geometric Wind Power Forecasting with Numerical Weather Prediction
A Diagnostic Framework for AI Agent Behavior
Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase…
Toward Anthropomorphic Dialogue: A Closed-Loop Framework for Human-Like Chat Generation, …
Logical Judgments Under Pressure: Diagnosing Syllogistic Stability with Learned Soft Pref…
Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost