LLM Watermark Evasion via Bias Inversion
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting
Online Irregular Multivariate Time Series Forecasting via Uncertainty-Driven Dual-Expert …
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
Multi-Adapter Representation Interventions via Energy Calibration
CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning
Do Clinical Models Change Treatment Decisions?
LiDDA: Data Driven Attribution at LinkedIn
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
Calibrating Conservatism for Scalable Oversight
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge
MCTS-Judge: Test-Time Scaling in LLM-as-a-Judge for Code Correctness Evaluation
HEART: Achieving Timely Multi-Model Training for Vehicle-Edge-Cloud-Integrated Hierarchic…
Isometry pursuit
FD-RAG: Federated Dual-System Retrieval-Augmented Generation
Revisiting Graph Autoencoders as Implicit Contrastive Learners
When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference
Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Repro…
FLUID: From Ephemeral IDs to Multimodal Semantic Codes for Industrial-Scale Livestreaming…
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
Delay-Aware Reinforcement Learning for Highway On-Ramp Merging under Stochastic Communica…
Verifiable Process Rewards for Agentic Reasoning
A Systematic Evaluation of Retrieval-Augmented Generation and Language Models for Space O…
Escher-Loop: Mutual Evolution by Closed-Loop Self-Referential Optimization