Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clin…
Explorar
Noticias de IA
29673 elementos — filtrados, clasificados y sin duplicados
Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, …
Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reve…
Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Gen…
Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence
Chatterbox-Flash: Prior-Calibrated Block Diffusion for Streaming Zero-Shot TTS
GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
On the impact of retrieved content representations in RAG Pipelines
OpenSTBench: Beyond Semantic Evaluation for Speech Translation
MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing U…
Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution
Differentially Private Preference Data Synthesis for Large Language Model Alignment
Unlearning in Diffusion Models: A Unified Framework with KL Divergence and Likelihood Con…
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
Fine-Tuning Improves Information Conveyance in Language Models
Safe Equilibrium Policy Optimization for Strategic Agent Policies
Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set App…
BlueFin: Benchmarking LLM Agents on Financial Spreadsheets
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucinati…
Do Large Language Models Encode Institutional Experience? Evidence from Cross-Linguistic …
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and …
DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological S…
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in …
Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Dom…
Linear Ordering Problem: Time for a Change
Fighting Numerical Hallucinations via Data-centric Compilation for Online Financial QA