Explorar

Noticias de IA

21272 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
HEART: Achieving Timely Multi-Model Training for Vehicle-Edge-Cloud-Integrated Hierarchic…
arXiv cs.AI Research & Papers
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation
arXiv cs.AI Research & Papers
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
arXiv cs.AI Research & Papers
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
arXiv cs.AI Research & Papers
LiDDA: Data Driven Attribution at LinkedIn
arXiv cs.AI Research & Papers
Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Polic…
arXiv cs.AI Research & Papers
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
arXiv cs.AI Research & Papers
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
arXiv cs.AI Research & Papers
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
arXiv cs.AI Research & Papers
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
arXiv cs.AI Research & Papers
Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting
arXiv cs.AI Research & Papers
LLM Watermark Evasion via Bias Inversion
arXiv cs.AI Research & Papers
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Tra…
arXiv cs.AI Research & Papers
Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration…
arXiv cs.AI Research & Papers
Cross-Entropy Games and Frost Training
arXiv cs.AI Research & Papers
Voluntary Collusion with Secret Tools in Competing LLM Agents
arXiv cs.AI Research & Papers
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
arXiv cs.AI Research & Papers
BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Servi…
arXiv cs.AI Research & Papers
Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection
arXiv cs.AI Research & Papers
When prompt perturbations break your A/B test: A valid statistical test for generative su…
arXiv cs.AI Research & Papers
A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space O…
arXiv cs.AI Research & Papers
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Repro…
arXiv cs.AI Research & Papers
When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference
arXiv cs.AI Research & Papers
FD-RAG: Federated Dual-System Retrieval-Augmented Generation
arXiv cs.AI Research & Papers
RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge
arXiv cs.AI Research & Papers
Calibrating Conservatism for Scalable Oversight
arXiv cs.AI Research & Papers
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
arXiv cs.AI Research & Papers
Multi-Adapter Representation Interventions via Energy Calibration
arXiv cs.AI Research & Papers
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
arXiv cs.AI Research & Papers
Continual Model Routing in Evolving Model Hubs