When and Why LLM Causal Priors Help: Closed-Loop Prior Selection for Amortized Causal Inf…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Exposing Weaknesses in Emotion Recognition in Conversations
Quantile-Led Feature Extraction for Multi-Horizon Predictive Maintenance in Industrial Ma…
Beyond Final Decisions: A Process-Centric Benchmark for Transparent AI-Assisted Peer Revi…
Learning Counterfactual World Models for Embodied Reasoning under Partial Observability
AgentBrew: Offline Tool-Use Agent Learning from Raw Real-World Trajectories
Agentic Pressure: The Endogenous Entropy of Reliable Autonomy
The Normalization of Deviance in AI Development
Modus Tollens and Counterfactuals and Counterfactual Reasoning Based on Three Types of Ne…
Beyond Prompts: Measuring and Optimizing LLM Tool-Agent Harnesses
From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational As…
Reducing Hallucinations in LLM-based Scientific Literature Analysis Using Peer Context Ou…
Learning to Focus: CSI-Free Hierarchical MARL for Reconfigurable Reflectors
CUSP: Decomposable Collective Uncertainty for Multi-Agent Multimodal Reasoning
Recovering Temporal and Geographic Signals from Language Model Embeddings
Distilling Vision-Language Models for On-Device Fire Understanding
More Than Mimicking Reviewers: Evaluating LLMs for Pre-Submission Peer Review
Generator-Independent Runtime Assurance under Partial Observation
Towards Embodied Air-Ground Cooperative Object Search: Benchmark, Dataset and Agentic Met…
DAREBench: Deployment-Aware and Reliable Evaluation of Models as Agents
Deep belief networks are exact
EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent
The convergent laboratory: when AI reasoning, autonomous experiments, high performance an…
The Failure Happens Before the Drift: The Social Dynamics of Values in LLM Agent Societies
Reasoning-Aware Compression: Identifying and Protecting Vulnerable Reasoning Circuits for…
Beyond "AI Helps Humans": Decision-Targeted Evaluation Design for Human-Agent Teams in th…
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record f…
Generating Instance Generators in PDDL Planning
SynthRCT: Scalable Conditional Deformation Synthesis for Synthetic Repeat CT Generation
From Where to How: Continuous 4D Interaction Forecasting from Egocentric Video