Explorar

Noticias de IA

29629 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics
arXiv cs.AI Research & Papers
A Theory of Conditional Collapse under Low-Rank Weight-Space Ablations: I. The Single-Blo…
arXiv cs.AI Research & Papers
Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algo…
arXiv cs.AI Research & Papers
Policy Fragmentation or Institutional Alignment? Institutional Governance of AI in Univer…
arXiv cs.AI Research & Papers
Large language models for partial differential equation workflows
arXiv cs.AI Research & Papers
When Teachers Mislead: Spurious-Signal-Aware On-Policy Distillation
arXiv cs.AI Research & Papers
Don't Peek at the Answer: Outcome-Masked Group Relative Policy Optimization for Label-Fre…
arXiv cs.AI Research & Papers
Shaping Wind-Tunnel Airflow for Unmanned Aerial Vehicles using Online Learning
arXiv cs.AI Research & Papers
UniGD: A Unified Generative-Discriminative Framework for Industrial Retrieval
arXiv cs.AI Research & Papers
Compound and Parallel Modes of Tropical Convolutional Neural Networks
arXiv cs.AI Research & Papers
When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Rea…
arXiv cs.AI Research & Papers
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents
arXiv cs.AI Research & Papers
GPTKB 2.0: Direct Construction of Disambiguated Knowledge Bases from Large Language Models
arXiv cs.AI Research & Papers
UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low H…
arXiv cs.AI Research & Papers
AI Assistance Reduces Persistence and Hurts Independent Performance
arXiv cs.AI Research & Papers
HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents
arXiv cs.AI Research & Papers
Diversity is Not Ambiguity: Toward Accurate and Efficient Ambiguity Detection for Open-Do…
arXiv cs.AI Research & Papers
State Propagation Also Satisfies: A Complex-Valued State-Space Model for Deterministic St…
arXiv cs.AI Research & Papers
DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces
arXiv cs.AI Research & Papers
Does Forgetting Transfer Across Modalities? A Real-World Benchmark for Cross-Modal Knowle…
arXiv cs.AI Research & Papers
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks
arXiv cs.AI Security & Safety
Permission Denied: Policy-Graded Evaluation of Coding Agents in Hardened Environments
arXiv cs.AI Research & Papers
Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems
arXiv cs.AI Research & Papers
LeanMem: Simple and Efficient Long-Term Memory for LLM Agents
arXiv cs.AI Research & Papers
GROW: Group-Relative Advantage-Weighted On-Policy Reinforcement Learning of Autoregressiv…
arXiv cs.AI Research & Papers
ChartAnno: Evaluating MLLMs for Chart Annotation Generation
arXiv cs.AI Research & Papers
Dr. AGENTONOMICS: A Didactic Experiment of AGENTONOMICS
arXiv cs.AI Research & Papers
AI World Cup 2026: Benchmarking Large Language Models for End-to-End Football Tournament …
arXiv cs.AI Research & Papers
WCM: World-Cognition Model for Generalizable Human-Robot Interaction
arXiv cs.AI Research & Papers
IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autofor…