HEART: Achieving Timely Multi-Model Training for Vehicle-Edge-Cloud-Integrated Hierarchic…
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
LiDDA: Data Driven Attribution at LinkedIn
Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Polic…
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting
LLM Watermark Evasion via Bias Inversion
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Tra…
Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration…
Cross-Entropy Games and Frost Training
Voluntary Collusion with Secret Tools in Competing LLM Agents
LaneRoPE: Positional Encoding for Collaborative Parallel Reasoning and Generation
BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Servi…
Detect by Yourself: Self-Designing Agentic Workflows for Few-Shot Graph Anomaly Detection
When prompt perturbations break your A/B test: A valid statistical test for generative su…
A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space O…
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Repro…
When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference
FD-RAG: Federated Dual-System Retrieval-Augmented Generation
RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge
Calibrating Conservatism for Scalable Oversight
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Multi-Adapter Representation Interventions via Energy Calibration
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
Continual Model Routing in Evolving Model Hubs