Explorar

Noticias de IA

22065 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents
arXiv cs.AI Research & Papers
Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessme…
arXiv cs.AI Research & Papers
From Cellular Responses to Pharmacological Domains: Multimodal Zero-Shot Drug Representat…
arXiv cs.AI Research & Papers
Entangled by Design: Spurious Intra-Variable Signal Routing in Tabular In-Context Learners
arXiv cs.AI Research & Papers
Instruction-Tuned Language Models Cannot Sample from Distributions They Can Describe
arXiv cs.AI Research & Papers
ScalableRAG: High-Quality RAG at Zero Ingestion Cost
arXiv cs.AI Research & Papers
PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention
arXiv cs.AI Research & Papers
ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning
arXiv cs.AI Research & Papers
Toward an Organizational Science of Multi-Agent LLM Systems: Decoupling Who, How, and Whi…
arXiv cs.AI Research & Papers
Are the High-weight Neurons the Important Ones in Image Classification Neural Networks?
arXiv cs.AI Research & Papers
Salient Knowledge Pathways: Sparse Cross-Modal Routing for Efficient Knowledge-Intensive …
arXiv cs.AI Research & Papers
Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical …
arXiv cs.AI Research & Papers
Does Runtime Topology Context Improve LLM-Generated Kubernetes Security Patches?
arXiv cs.AI Research & Papers
Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT
arXiv cs.AI Research & Papers
Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA
arXiv cs.AI Research & Papers
From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios
arXiv cs.AI Research & Papers
Everyone is unique: Towards Behaviorally Heterogeneous Negotiation Dialogue Systems for D…
arXiv cs.AI Research & Papers
scMIR: a vision-language foundation model for single-cell light microscopy image represen…
arXiv cs.AI Research & Papers
RankFormer: A Propose-then-Select Transformer for Multi-Agent Multimodal Trajectory Predi…
arXiv cs.AI Research & Papers
The Disruptive Impact of Large Language Models on Capture the Flag Competitions and the P…
arXiv cs.AI Research & Papers
EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff
arXiv cs.AI Research & Papers
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
arXiv cs.AI Research & Papers
Deep Delta Learning
arXiv cs.AI Research & Papers
Localizing Persona Representations in LLMs
arXiv cs.AI Research & Papers
Do Models Fake Alignment Without Clear Consequences?
arXiv cs.AI Research & Papers
GrocLM: Grocery Category Recommendation in E-Commerce with Large Language Models
arXiv cs.AI Research & Papers
OmniQEC: discovering practical quantum error-correcting codes by an AI scientist
arXiv cs.AI Research & Papers
Fairness Is Not Enough: Auditing Competence and Intersectional Bias in AI-powered Resume …
arXiv cs.AI Research & Papers
A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks
arXiv cs.AI Research & Papers
Measuring the State of Open Science in Transportation Using Large Language Models