Explorar

Noticias de IA

30329 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
MedCase-Structured: A Text-to-FHIR Dataset for Benchmarking Diagnostic Reasoning in Clini…
arXiv cs.AI Research & Papers
Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion
arXiv cs.AI Research & Papers
In-Context Reward Adaptation for Robust Preference Modeling
arXiv cs.AI Research & Papers
On Language Generation in the Limit with Bounded Memory
arXiv cs.AI Research & Papers
Toward User Preference Alignment in LLM Recommendation via Explicit Context Feedback
arXiv cs.AI Research & Papers
PTCG-Bench: Can LLM Agents Master Pok\'emon Trading Card Game?
arXiv cs.AI Research & Papers
LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning
arXiv cs.AI Research & Papers
VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
arXiv cs.AI Research & Papers
PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data
arXiv cs.AI Research & Papers
A Matter of Interest: Understanding Interestingness of Math Problems in Humans and Langua…
arXiv cs.AI Research & Papers
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Da…
arXiv cs.AI Research & Papers
Benchmarking at the Edge of Comprehension
arXiv cs.AI Research & Papers
Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangle…
arXiv cs.AI Research & Papers
ReSpinQuant: Efficient Layer-Wise LLM Quantization via Subspace Residual Rotation Approxi…
arXiv cs.AI Research & Papers
Recurrent Structural Policy Gradient for Partially Observable Mean Field Games
arXiv cs.AI Research & Papers
DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories
arXiv cs.AI Research & Papers
FormalEvolve: Neuro-Symbolic Evolutionary Search for Diverse Autoformalization
arXiv cs.AI Research & Papers
When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-La…
arXiv cs.AI Research & Papers
MediHive: A Decentralized Agent Collective for Medical Reasoning
arXiv cs.AI Research & Papers
Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Po…
arXiv cs.AI Research & Papers
The Planetary Cost of AI Acceleration, Part II: The 10th Planetary Boundary and the 6.5-Y…
arXiv cs.AI Research & Papers
Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling
arXiv cs.AI Research & Papers
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With N…
arXiv cs.AI Research & Papers
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodi…
arXiv cs.AI Research & Papers
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing
arXiv cs.AI Research & Papers
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
arXiv cs.AI Research & Papers
Beyond Normalization: Rethinking the Partition Function as a Difficulty Scheduler for RLVR
arXiv cs.AI Research & Papers
GPS-Enhanced Tourist Mobility Modeling with Seasonal Spatial Priors and LLM-Based Activit…
arXiv cs.AI Research & Papers
ParaTool: Shifting Tool Representations from Context to Parameters
arXiv cs.AI Research & Papers
A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search