Online Irregular Multivariate Time Series Forecasting via Uncertainty-Driven Dual-Expert …
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of…
Multi-Adapter Representation Interventions via Energy Calibration
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Calibrating Conservatism for Scalable Oversight
Architecture-driven Shift: towards a lightweight selector for capturing the trends of log…
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge
FD-RAG: Federated Dual-System Retrieval-Augmented Generation
When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Repro…
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space O…
When prompt perturbations break your A/B test: A valid statistical test for generative su…
Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection
BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Servi…
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
Voluntary Collusion with Secret Tools in Competing LLM Agents
Cross-Entropy Games and Frost Training
Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration…
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
LiDDA: Data Driven Attribution at LinkedIn
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines
Hallucination Behavior in Multimodal LLMs Across Agricultural Image Interpretation and Ge…
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation