Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Benchmarking LLM Judges for Mobile Agent Evaluation
arXiv cs.AI Research & Papers
EnterpriseRAG: Benchmarking LLM Instruction Adherence and Robustness under Non-Ideal Ente…
arXiv cs.AI Research & Papers
JieZi: A Large-Scale Expert-Audited Dataset and Benchmark for Ancient Chinese Character E…
arXiv cs.AI Research & Papers
Uncertainty-Aware Probabilistic Constrained Clustering from Entangled Pairwise Supervision
arXiv cs.AI Research & Papers
CLEAR: Class-wise Expert Aggregation with Structured Sampling for Long-Tailed Classificat…
arXiv cs.AI Research & Papers
Do You See What You Draw? A Semantic Closed-Loop Framework for Holistic Evaluation of Uni…
arXiv cs.AI Research & Papers
Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended …
arXiv cs.AI Research & Papers
Governing Agentic AI in FinTech
arXiv cs.AI Research & Papers
Domain-Aware Lightweight Spectral-Grouped Convolutions for Hyperspectral Fish Freshness C…
arXiv cs.AI Research & Papers
Information Abundance Paradox: Long-Context Training Undermines Parametric Knowledge
arXiv cs.AI Research & Papers
Dion3: Full-Stack Orthogonal Updates
arXiv cs.AI Research & Papers
Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes
arXiv cs.AI Research & Papers
Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperativ…
arXiv cs.AI Research & Papers
RoadWeaver: Large-Scale Lane-Level HD Map Generation from Scratch for Autonomous Driving …
arXiv cs.AI Research & Papers
Confidence Calibration of Deep Learning Systems
arXiv cs.AI Research & Papers
Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World …
arXiv cs.AI Research & Papers
Towards the Harness of Embodied Agents
Hugging Face Daily Papers Research & Papers
SCOPE: Subspace Clustering with Online Per-Head Top-K Estimation for Sparse Video Attenti…
Hugging Face Daily Papers Research & Papers
CRAFT: LLM-Based Iterative Refinement for Temporal Reasoning over Clinical Narratives
Hugging Face Daily Papers Research & Papers
ViTOED: A Dataset for Target-Oriented Emotion Detection on Vietnamese Social Media Texts
Hugging Face Daily Papers Research & Papers
Dual-Stream Cross-Anchor Correction Grounding Long-Form Captions and the Domain Limits of…
Hugging Face Daily Papers Research & Papers
A Cloud-Edge System for Multimodal Clinical Screening in Resource-Constrained Rural Setti…
Hugging Face Daily Papers Research & Papers
Error-Aware Reverse Auction Mechanism for Large Language Model Routing
Hugging Face Blog Research & Papers
What We Learned by Reproducing 2,200 papers from ICML
Hugging Face Daily Papers Research & Papers
Drive-to-Music: Context-Aware Generative Audio for In-Vehicle Experiences
Hugging Face Daily Papers Research & Papers
Represent, Then Generate: Multimodal-Conditioned Time-Series Generation under Irregular M…
Hugging Face Daily Papers Research & Papers
Attribute-Conditioned Multimodal Slot Factorization for Controllable Fashion Retrieval
Hugging Face Daily Papers Research & Papers
Can Vision-Language Models Assess Proxemic Risk from Egocentric Robot Images?
Hugging Face Daily Papers Research & Papers
DreamFly: Causal Memory and Receding-Horizon Diffusion Planning for Aerial Vision-Languag…
Hugging Face Daily Papers Research & Papers
AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses