Explorar

Noticias de IA

21272 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Online Irregular Multivariate Time Series Forecasting via Uncertainty-Driven Dual-Expert …
arXiv cs.AI Research & Papers
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
arXiv cs.AI Research & Papers
Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images
arXiv cs.AI Research & Papers
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
arXiv cs.AI Research & Papers
Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of…
arXiv cs.AI Research & Papers
Multi-Adapter Representation Interventions via Energy Calibration
arXiv cs.AI Research & Papers
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
arXiv cs.AI Research & Papers
Calibrating Conservatism for Scalable Oversight
arXiv cs.AI Research & Papers
Architecture-driven Shift: towards a lightweight selector for capturing the trends of log…
arXiv cs.AI Research & Papers
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
arXiv cs.AI Research & Papers
RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge
arXiv cs.AI Research & Papers
FD-RAG: Federated Dual-System Retrieval-Augmented Generation
arXiv cs.AI Research & Papers
When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference
arXiv cs.AI Research & Papers
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Repro…
arXiv cs.AI Research & Papers
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
arXiv cs.AI Research & Papers
A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space O…
arXiv cs.AI Research & Papers
When prompt perturbations break your A/B test: A valid statistical test for generative su…
arXiv cs.AI Research & Papers
Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection
arXiv cs.AI Research & Papers
BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Servi…
arXiv cs.AI Research & Papers
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
arXiv cs.AI Research & Papers
Voluntary Collusion with Secret Tools in Competing LLM Agents
arXiv cs.AI Research & Papers
Cross-Entropy Games and Frost Training
arXiv cs.AI Research & Papers
Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration…
arXiv cs.AI Research & Papers
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
arXiv cs.AI Research & Papers
LiDDA: Data Driven Attribution at LinkedIn
arXiv cs.AI Research & Papers
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
arXiv cs.AI Research & Papers
Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines
arXiv cs.AI Research & Papers
Hallucination Behavior in Multimodal LLMs Across Agricultural Image Interpretation and Ge…
arXiv cs.AI Research & Papers
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
arXiv cs.AI Research & Papers
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation