Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language M…
arXiv cs.AI Research & Papers
No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
arXiv cs.AI Research & Papers
LoRe: Adaptive Interaction-Evaluation Routing with Per-Step Interaction Budgets for Itera…
arXiv cs.AI Research & Papers
CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Langu…
arXiv cs.AI Research & Papers
First head-to-head comparison of agentic AI applied to the analysis of simulated data of …
arXiv cs.AI Research & Papers
Hallucination Detection-Guided Preference Optimization for Clinical Summarization
arXiv cs.AI Research & Papers
MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models
arXiv cs.AI Research & Papers
GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human
arXiv cs.AI Research & Papers
Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strate…
arXiv cs.AI Research & Papers
Predicting Causal Effects from Natural Language Queries using Structured Representations
arXiv cs.AI Research & Papers
When 2D Tasks Meet 1D Serialization: On Serialization Friction in Structured Tasks
arXiv cs.AI Research & Papers
The Little Book of Generative AI Foundations: An Intuitive Mathematical Primer
arXiv cs.AI Research & Papers
Learning Context-Conditioned Predicate Semantics via Prototype Feedback
arXiv cs.AI Research & Papers
Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across P…
arXiv cs.AI Research & Papers
Architecture-Sensitive Supervised Fine-Tuning for Screen-Conditioned Action Prediction: A…
arXiv cs.AI Research & Papers
VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring ov…
arXiv cs.AI Research & Papers
Quantifying and Optimizing Simplicity via Polynomial Representations
arXiv cs.AI Research & Papers
SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents
arXiv cs.AI Research & Papers
Moment-KV: Momentum-Based Decode-Time KV Cache Compression for Long Generation
arXiv cs.AI Research & Papers
Toward AI Systems That Understand Self and Others: A Multi-Phase Inference Framework for …
arXiv cs.AI Research & Papers
Formalizing Mathematics at Scale
arXiv cs.AI Research & Papers
KairosAgent: Agentic Time Series Forecasting with Fused Semantic Reasoning
arXiv cs.AI Research & Papers
Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication …
arXiv cs.AI Research & Papers
VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior T…
arXiv cs.AI Research & Papers
When Should Models Change Their Minds? Contextual Belief Management in Large Language Mod…
arXiv cs.AI Research & Papers
ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive In…
arXiv cs.AI Research & Papers
Locally Coherent, Globally Incoherent: Bounding Compositional Incoherence in Multi-Compon…
arXiv cs.AI Research & Papers
SchGen: PCB Schematic Generation with Semantic-Grounded Code Representations
arXiv cs.AI Research & Papers
Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation
arXiv cs.AI Research & Papers
Orthogonal Concept Erasure for Diffusion Models