$\varepsilon$-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence
DiG-bench: Discovery in Games
Dead text or binding clause? Measuring and restoring constraint influence in black-box LL…
@skills: Attention is all you have
Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence
General Probabilities of Causation with Causal Knowledge
Designing AI Pipelines for Decision-Ready ITSM Intelligence
On the Expressive Power of Transformers
The Role of Natural Language Understanding in Multimodal Video-Based Dengue Diagnosis
Spatial Memory Agent: Experience-Grounded Procedure Memory for Spatial Intelligence
PROVE-RT: Generating Mechanized Theorem Prover Scripts for Real-Time Systems using LLMs
ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs
Beyond Retrieval: Query-Conditioned Reuse of Long-Horizon Agent Trajectories
Practice Makes Unsafe: Skill Misevolution in Self-Improving LLM Agents
ReflectFact: Self-Reflective Agents for Improving Comprehension and Reasoning in Multi-Ho…
Predictive Memory Localization: Forecasting Selective Intervention Paths from Internal Si…
Agent Behavioral Contracts II: Certifying Compositional Reliability Without Assuming Inde…
FlashDrive: Flash Vision-Language-Action Inference for Autonomous Driving
Decomposition of Evidence, Contradiction, and Fragility in Perturbation Responses
Moose: Latent concept learning with reasoning-shortcut awareness in $\mathcal{EL}^{++}$
BoardroomAI: Dependency-Aware Human-Steerable Multi-Agent Deliberation through Evolving D…
DMDIntel: Interpreting Large Language Models via Dynamic Mode Decomposition
Uniform Herding: Exemplar Replay with Representation Refresh
EEG-PRIME: Prototype-Aligned Representation Learning with Multi-Level Conditioning for EE…
Multi-Layer Context Camouflaging: A Semantic Superposition and Contextual Lamination Fram…
Robust Dempster-Shafer Evidence Fusion with Chaos-Conflict Measurement and Historical-Exp…
Numeracy in Large Language Models: Fundamental Limitations and Paths to Improvement
Teach the Magnitude, Not the Direction: Verifier-Bounded Credit Assignment for Multi-Turn…
Capability Sheaves for Compositional Agent-Harness Repair: Controlled Quotients and a Rea…