Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
ClawProBench: Trace-Aware Evaluation of AI Agents with Runtime Coverage and Frozen Workpl…
arXiv cs.AI Research & Papers
Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
arXiv cs.AI Research & Papers
Proxy reliance in large language model decisions is uncalibrated to predictive evidence
arXiv cs.AI Research & Papers
CDEG: Learning Decision-Critical Evidence for Long-Horizon Diagnostic Agents
arXiv cs.AI Research & Papers
Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN
arXiv cs.AI Research & Papers
ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Eva…
arXiv cs.AI Research & Papers
FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts
arXiv cs.AI Research & Papers
Deep-Learning-Based Pixelated Microwave Filter Design and Characterization using Electro-…
arXiv cs.AI Research & Papers
Decision-Support and Modeling with Large Language Models for Geothermal Well Arrays
arXiv cs.AI Research & Papers
Language Chain in Alignment: Cross-Lingual Ranking Preference Optimization
arXiv cs.AI Research & Papers
On the Role of Citations in Preference Data
arXiv cs.AI Research & Papers
When Does AI for PDEs Yield Scientific Evidence?
arXiv cs.AI Research & Papers
More Accurate or More Efficient? Evaluating Locally Deployed Compact Open-Weight Language…
arXiv cs.AI Research & Papers
SchemaRouter: Field-Aware Tool Routing for Efficient Heterogeneous Agentic RAG
arXiv cs.AI Research & Papers
Sycophants in the Courtroom: Are LLMs Fragile to Juridical Authority and Evolving Legal S…
arXiv cs.AI Research & Papers
Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning
arXiv cs.AI Research & Papers
Evaluating Multimodal Narrative Understanding of Popular Hollywood Films
arXiv cs.AI Research & Papers
ST-EVO: Towards Generative Spatio-Temporal Evolution of Multi-Agent Communication Topolog…
arXiv cs.AI Research & Papers
Performance of a domain-specific large language model in answering patient questions in p…
arXiv cs.AI Research & Papers
MegaMem: A Retrieval Solution for Ultra-Large Context Windows
arXiv cs.AI Research & Papers
Artificial Empathy: Towards a Framework for Unsupervised Agency Detection and Policy Reco…
arXiv cs.AI Research & Papers
KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Effici…
arXiv cs.AI Research & Papers
From SQL Generation to Tool Selection: A Domain-Oriented Pattern for MCP Servers
arXiv cs.AI Research & Papers
Improving Energy Efficiency of Oil Platforms Through Optimal Loading of Diesel Generators…
arXiv cs.AI Research & Papers
StrategyBench: Evaluating Explicit Strategy Induction in Large Language Models
arXiv cs.AI Research & Papers
SAFE-G: Structure-aware Faithful Evidence-guided Generation for Knowledge-based Visual Qu…
arXiv cs.AI Research & Papers
Let the Bullets Fly: Multimodal Fake News Detection with Temporal-Aligned Generative Danm…
arXiv cs.AI Research & Papers
Towards a Densing Law for User Representation Learning at Billion-Scale Capacity
arXiv cs.AI Research & Papers
From Recognition to Reasoning: Advancing Multimodal Harmful Meme Detection via Chain-of-T…
arXiv cs.AI Research & Papers
Radial Compensation: The Inverse Base-Distribution Problem for Chart-Based Generative Mod…