A2RAG: Adaptive Agentic Graph Retrieval for Cost-Aware and Reliable Reasoning
Explorar
Noticias de IA
29349 elementos — filtrados, clasificados y sin duplicados
Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding
CTIConnect: A Benchmark for Retrieval-Augmented LLMs over Heterogeneous Cyber Threat Inte…
Beyond Vector Similarity: A Structural Analysis of Graph-Augmented Retrieval for Industri…
VASO: Formally Verifiable Self-Evolving Skills for Physical AI Agents
Escaping the Verifier: Learning to Reason via Demonstrations
RAG Security and Privacy: Formalizing the Threat Model and Attack Surface
Agents' Last Exam
Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff
A Survey on Diffusion Language Models
EGTR-Review: Efficient Evidence-Grounded Scientific Peer Review Generation via Multi-Agen…
Emergent Language as an Approach to Conscious AI
In-Training Defenses against Emergent Misalignment in Language Models
Scaling Laws and Spectra of Shallow Neural Networks in the Feature Learning Regime
Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Mod…
Residual Modeling for High-Fidelity Learned Compression of Scientific Data
Explainable AI-Driven Cyber Risk Analytics and Model Reliability Assessment for Intellige…
Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchm…
Reformulating Neural Operators in $d+1$ Dimensions for Embedding Evolution
CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspecti…
ReasoningFlow: Discourse Structures for Understanding LLM Reasoning Traces
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Halluc…
ProSarc: Prosody-Aware Sarcasm Recognition Framework via Temporal Prosodic Incongruity
Multilingual Coreference Resolution via Cycle-Consistent Machine Translation
Channel-Wise Mixed-Precision Quantization for Large Language Models
SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horiz…
A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice
A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing
A Taxonomy of Runtime Faults in Model Context Protocol Servers
Entropy-Based Evaluation of AI Agents: A Lightweight Framework for Measuring Behavioral P…