What Drives LLM Self-Reflection? A Controlled Ablation of Uncertainty Routing in Armed Co…
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
Unlearning at Scale: State-Exact Trace-Preserving Deletion in Billion-Parameter Language …
The AI Accountability Ecosystem in the Era of Language Models
Jagged Judges: Epistemic Stability Under Silence, Pressure, and Persistence
AnchorSIPS: A Synthetic Dataset and Evaluation Resource for Evidence-Supported Psychosis-…
StreamReason-Bench: Can Large Language Models Reason about Event-Time Stream-Processing S…
Governed Persistent Memory: Source-Bound State Semantics and Fail-Closed Release for Long…
When AI Is Your Pastor: A Benchmark for Theological Triage and Pastoral Guidance in Large…
Research Assistant: AstraZeneca's Agentic System for R&D
Learning to Adapt Cross-Domain Preferences via Meta-LoRA for LLM Personalization
DiG-bench: Discovery in Games
CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence
Assessment Design in the GenAI Era: The X1-X2-X3 Assessment Pattern for Testing Students'…
Humans are Missing from AI Coding Agent Research
From Observation to Intervention: Memory in Brains and Large Language Models
Unified Multi-Dimensional Benchmark for Complex Graph Reasoning in Large Language Models
Multi-Agent Scheduling with LLM-Assisted Contract Net Negotiation for Stream Processing i…
Personalized Scorer Modeling: A Learning-Based Framework for Deriving Robust Sleep Stage …
SSPO: Structure-Aware Similarity-Weighted Preference Optimization for Neural Combinatoria…
Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces
Predicting When Random Low-Dimensional Reparameterizations Train Neural Networks
Not All Nudges Land: Behavioral Controllability and Elaboration Quality in AI-Supported J…
SchemaLink: An Intelligent Web Editor for LinkML Schema Curation
LLMs Are Not Good Strategists, Yet Memory-Enhanced Agency Boosts Reasoning
Interpretable Causal Discovery via Causal-Effect Constraints
Tracing Provenance and Detecting Tampering with Complementary LLM Watermarks
HybridSB-MoE: Dual-Domain Schr\"odinger Bridges with Scene-Adaptive Expert Routing for Sp…
Error-Aware Reverse Auction Mechanism for Large Language Model Routing
Beyond Outcome Rewards: Step-Level Self-Distilled Policy Optimization for Deep Search Age…
SynWeaver: Website-Prior Task and Trajectory Co-Synthesis for Web Agents