Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
The Planetary Cost of AI Acceleration, Part II: The 10th Planetary Boundary and the 6.5-Y…
arXiv cs.AI Research & Papers
Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Po…
arXiv cs.AI Research & Papers
MediHive: A Decentralized Agent Collective for Medical Reasoning
arXiv cs.AI Research & Papers
When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-La…
arXiv cs.AI Research & Papers
FormalEvolve: Neuro-Symbolic Evolutionary Search for Diverse Autoformalization
arXiv cs.AI Research & Papers
DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories
arXiv cs.AI Research & Papers
Recurrent Structural Policy Gradient for Partially Observable Mean Field Games
arXiv cs.AI Research & Papers
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approxi…
arXiv cs.AI Research & Papers
Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangle…
arXiv cs.AI Research & Papers
Benchmarking at the Edge of Comprehension
arXiv cs.AI Research & Papers
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Da…
arXiv cs.AI Research & Papers
A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Langua…
arXiv cs.AI Research & Papers
PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data
arXiv cs.AI Research & Papers
VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
arXiv cs.AI Research & Papers
LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning
arXiv cs.AI Research & Papers
PTCG-Bench: Can LLM Agents Master Pok\'emon Trading Card Game?
arXiv cs.AI Research & Papers
Toward User Preference Alignment in LLM Recommendation via Explicit Context Feedback
arXiv cs.AI Research & Papers
On Language Generation in the Limit with Bounded Memory
arXiv cs.AI Research & Papers
In-Context Reward Adaptation for Robust Preference Modeling
arXiv cs.AI Research & Papers
Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion
arXiv cs.AI Research & Papers
MedCase-Structured: A Text-to-FHIR Dataset for Benchmarking Diagnostic Reasoning in Clini…
arXiv cs.AI Research & Papers
NICE: A Theory-Grounded Diagnostic Benchmark for Social Intelligence of LLMs
arXiv cs.AI Research & Papers
Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context…
arXiv cs.AI Research & Papers
BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices
arXiv cs.AI Research & Papers
Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction
arXiv cs.AI Research & Papers
Review Arcade: On the Human Alignment and Gameability of LLM Reviews
arXiv cs.AI Research & Papers
Frontier LLM-based agents can overcome the ontology curation bottleneck for natural pheno…
arXiv cs.AI Research & Papers
VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis
arXiv cs.AI Research & Papers
When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis
arXiv cs.AI Research & Papers
The Hamilton-Jacobi Theory of Deep Learning