What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration…
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-S…
Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimiz…
When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models
Deep Learning Method for Stationary Distribution of Reflected Brownian Motion
ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical D…
LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity
Out of Sight: Compression-Aware Content Protection against Agentic Crawlers
LEEVLA: Seeing What Matters in Latent Environment Evolution for Vision-Language-Action
Leveraging Color Naming for Image Enhancement
Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning
TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Ta…
Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment
Multi-Agent Firewall Architecture for Privacy Protection of Sensitive Data in Interaction…
From Legacy Documentation to OSCAL: An MCP-Based Agent Pipeline for Threat-Informed Conti…
FSD-VLN: Fast-Slow Dual-System Modeling for Aerial Long-Horizon Vision-Language Navigation
On the Role of Conversational Timing in Synthetic Training Data for ASR
Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connect…
Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototy…
Swapping Faces, Saving Features: A Dual-Purpose Pipeline for Pedestrian Privacy in ITS
DrugGen 2: A disease-aware language model for enhancing drug discovery
Track2Map: Online Deformable SLAM with Motion-Aware Pose Optimization in Robotic Surgery
WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autono…
ADORN: Adaptive Drift handling for Open RAN using Reinforcement Learning
Two Axes of LLM Abstention: Answer Correctness and Question Answerability
VEGAS: Human-Aligned Video Caption Evaluation via Gaze
The Context Access Divide: Interaction-Level Architecture as a Complementary Dimension of…
Cognitive-structured Multimodal Agent for Multimodal Understanding, Generation, and Editi…
DocMaster: A Hierarchical Structure-Aware System for Document Analysis
SMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric Scheduli…