Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
AI Can Be Easily Persuaded in Clinical Decision Making
arXiv cs.AI Research & Papers
SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents
arXiv cs.AI Research & Papers
MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative…
arXiv cs.AI Research & Papers
Towards Cognitive Process-Aware Proactive Writing Support
arXiv cs.AI Research & Papers
Subtraction-Based Tumor Segmentation and Lesion-Centered pCR Prediction for the MAMA-MIA …
arXiv cs.AI Research & Papers
An Open-Source, Event-Driven Pipeline for Cryptocurrency Market Data: Ingestion, Forecast…
arXiv cs.AI Research & Papers
AutoCRAT: Within-trajectory Joint Control of Stochasticity and Compute for LLM Reasoning
arXiv cs.AI Research & Papers
DS-Lighting: Making Agent Harnesses Explicit for Data-Science Automation
arXiv cs.AI Research & Papers
HEAR Who Said What: Unlocking Speaker-Attributed Reasoning via Counterfactual Voice Groun…
arXiv cs.AI Research & Papers
Reverse N-Wise Output-Oriented Testing for AI/ML and Quantum Computing Systems
arXiv cs.AI Research & Papers
PokaiTrainer: Scaling Belief-State Search to Competitive Pok\'emon VGC
arXiv cs.AI Research & Papers
CoLa-ICD: A Knowledge-Enhanced Framework for Long-Tail Automated Medical Coding
arXiv cs.AI Research & Papers
A rigor-matched audit of periodic-step layer skipping for efficient llm inference: confla…
arXiv cs.AI Research & Papers
Curvature Cryptanalysis of Smooth Transformer Feed-Forward Networks
arXiv cs.AI Research & Papers
mRNA Design and Optimization with Deep Knowledge-Infused Approach
arXiv cs.AI Research & Papers
Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Ur…
arXiv cs.AI Research & Papers
Hybrid Offline-Online Multi-Agent Decision Transformers for Wireless Resource Management
arXiv cs.AI Research & Papers
Beyond Dense States: Sparse Transcoders as Causally Testable Operators for LLM Latent Rea…
arXiv cs.AI Research & Papers
FAA Framework: A Large Language Model-Based Approach for Credit Card Fraud Investigations
arXiv cs.AI Research & Papers
An Agentic Retrobiosynthesis Framework with Learned Frontier Selection
arXiv cs.AI Research & Papers
The reach of a verification tool decides its value: A controlled study of verification su…
arXiv cs.AI Research & Papers
AgentLogs: A Dataset for Opening the Black Box of GitHub's Cloud Agent
arXiv cs.AI Research & Papers
Learning to Evaluate Before Improving: Automatic Rubric Induction for Automatic Research …
arXiv cs.AI Research & Papers
Learning Simple Test-Time Environments for LLM Web Agents
arXiv cs.AI Research & Papers
Masked Distillation: Internalizing the Chain-of-Thought in Language Models
arXiv cs.AI Research & Papers
post-graph-rag: A PostgreSQL-Native Bi-Temporal Graph RAG Engine with Temporal Grounding …
arXiv cs.AI Research & Papers
Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems
arXiv cs.AI Research & Papers
On the Prospects of Dynamic LLM Conversations in Software Development
arXiv cs.AI Research & Papers
Automatic Conversion of NICE Guidelines to an Executable Computational Model Using Large …
arXiv cs.AI Research & Papers
The Race between Agentic AI Capabilities and Data Quality Control in Online Surveys