Explorar

Noticias de IA

29673 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clin…
arXiv cs.AI Research & Papers
Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty
arXiv cs.AI Security & Safety
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, …
arXiv cs.AI Security & Safety
Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reve…
arXiv cs.AI Research & Papers
Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Gen…
arXiv cs.AI Research & Papers
Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence
arXiv cs.AI Research & Papers
Chatterbox-Flash: Prior-Calibrated Block Diffusion for Streaming Zero-Shot TTS
arXiv cs.AI Research & Papers
GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation
arXiv cs.AI Research & Papers
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
arXiv cs.AI Research & Papers
On the impact of retrieved content representations in RAG Pipelines
arXiv cs.AI Research & Papers
OpenSTBench: Beyond Semantic Evaluation for Speech Translation
arXiv cs.AI Research & Papers
MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing U…
arXiv cs.AI Research & Papers
Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution
arXiv cs.AI Research & Papers
Differentially Private Preference Data Synthesis for Large Language Model Alignment
arXiv cs.AI Research & Papers
Unlearning in Diffusion Models: A Unified Framework with KL Divergence and Likelihood Con…
arXiv cs.AI Research & Papers
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring
arXiv cs.AI Research & Papers
Fine-Tuning Improves Information Conveyance in Language Models
arXiv cs.AI Research & Papers
Safe Equilibrium Policy Optimization for Strategic Agent Policies
arXiv cs.AI Research & Papers
Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation
arXiv cs.AI Research & Papers
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set App…
arXiv cs.AI Research & Papers
BlueFin: Benchmarking LLM Agents on Financial Spreadsheets
arXiv cs.AI Research & Papers
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucinati…
arXiv cs.AI Research & Papers
Do Large Language Models Encode Institutional Experience? Evidence from Cross-Linguistic …
arXiv cs.AI Research & Papers
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and …
arXiv cs.AI Research & Papers
DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological S…
arXiv cs.AI Research & Papers
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
arXiv cs.AI Research & Papers
Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in …
arXiv cs.AI Research & Papers
Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Dom…
arXiv cs.AI Research & Papers
Linear Ordering Problem: Time for a Change
arXiv cs.AI Research & Papers
Fighting Numerical Hallucinations via Data-centric Compilation for Online Financial QA