Trust and Its Betrayal under Three Representational Strategies
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
CrystalMem: Elastic Memory for Self-Evolving LLM Agents via Knowledge Crystallization
WM-Cov: Test Adequacy for Interactive World-Model-Style Autonomous Driving Simulation
RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals
Geometric Self-Supervised Pre-training for Neural Combinatorial Optimization
More Debate, Same Evidence: Structural Limits of Homogeneous Multi-Agent Groundedness
Personalizing Large Language Model Agents with Small Policy Models
TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity…
AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?
Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale
H+ Embedding: Harmonizing Global and Token-Level Retrieval with Context-Dependent Phrases
Evolutionary Curriculum Learning Improves Biological Sequence Modeling
Linguistic Context Recodes Visual Representations in Vision-Language Models
SIRIN: A Unified Toolkit for Detecting Contextual Hallucinations in Retrieval-Augmented a…
Nova: An End-to-End MLIR Compiler for Deep Learning
Motif-Mamba: network motif improved mamba for long-range sequence modeling
Memory Reward Inflation in Self-Improving LLM Agents
Optimization and Constraint Modeling using LLMs with a Retrieval Augmented Generation Pro…
Energy Efficiency of Locally Deployed LLMs: A Preliminary Quantitative GPU Power Benchmar…
DGA$_2$D: Directed Graph-Guided Automated Algorithm Design with Large Language Models
Behavioral Grammar: Detecting Adaptive Malware via Tiny Language Model Priors and Second-…
FinDeepIndicator: Benchmarking Deep Research Agents in End-to-End Financial Indicator Con…
Large language models improve physician accuracy but lead to false reliance
Learning What to Remember and What to Internalize in LLM Self-Evolution via Adaptive Memo…
402Pilot: An x402 Decision Layer for Autonomous Agent Micropayments
High-Stakes Decisions with Language Models: Insights from Emergency Triage
Key-Value Means: Transformers with Expandable Block-Recurrent Compressed Memory
No Single Neuron of Failure: Distributed Safety Alignment Against White-Box Attacks
ELECTRIC: Evidential Learning-Enhanced CT Reconstruction via Iterative Correction
Enhancing LLMs with Context-Specific Knowledge for Mitigating Misinformation in SMEs: A R…