Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
RIACT: A Responsible AI System for Personalized Study Habit Tracking and Early Burnout Si…
arXiv cs.AI Research & Papers
Is Next-Chunk Reasoning RL Really Better than SFT? Revisiting Training Strategies under n…
arXiv cs.AI Research & Papers
CONTRAMEM: Learning Self-Evolving Procedural Memory from Contrasting Multi-Model Trajecto…
arXiv cs.AI Research & Papers
Towards Comprehensive Basketball Understanding
arXiv cs.AI Research & Papers
Procedural Knowledge Extraction from Industrial Troubleshooting Guides Using Vision Langu…
arXiv cs.AI Research & Papers
Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
arXiv cs.AI Research & Papers
Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN
arXiv cs.AI Research & Papers
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-…
arXiv cs.AI Research & Papers
ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Eva…
arXiv cs.AI Research & Papers
FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts
arXiv cs.AI Research & Papers
On the Role of Citations in Preference Data
arXiv cs.AI Research & Papers
Language Chain in Alignment: Cross-Lingual Ranking Preference Optimization
arXiv cs.AI Research & Papers
Decision-Support and Modeling with Large Language Models for Geothermal Well Arrays
arXiv cs.AI Research & Papers
When Does AI for PDEs Yield Scientific Evidence?
arXiv cs.AI Research & Papers
More Accurate or More Efficient? Evaluating Locally Deployed Compact Open-Weight Language…
arXiv cs.AI Research & Papers
Sycophants in the Courtroom: Are LLMs Fragile to Juridical Authority and Evolving Legal S…
arXiv cs.AI Research & Papers
Evaluating Multimodal Narrative Understanding of Popular Hollywood Films
arXiv cs.AI Research & Papers
Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning
arXiv cs.AI Research & Papers
SchemaRouter: Field-Aware Tool Routing for Efficient Heterogeneous Agentic RAG
arXiv cs.AI Research & Papers
KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Effici…
arXiv cs.AI Research & Papers
Performance of a domain-specific large language model in answering patient questions in p…
arXiv cs.AI Research & Papers
ST-EVO: Towards Generative Spatio-Temporal Evolution of Multi-Agent Communication Topolog…
arXiv cs.AI Research & Papers
MegaMem: A Retrieval Solution for Ultra-Large Context Windows
arXiv cs.AI Research & Papers
StrategyBench: Evaluating Explicit Strategy Induction in Large Language Models
arXiv cs.AI Research & Papers
SAFE-G: Structure-aware Faithful Evidence-guided Generation for Knowledge-based Visual Qu…
arXiv cs.AI Research & Papers
Radial Compensation: The Inverse Base-Distribution Problem for Chart-Based Generative Mod…
arXiv cs.AI Research & Papers
From Recognition to Reasoning: Advancing Multimodal Harmful Meme Detection via Chain-of-T…
arXiv cs.AI Research & Papers
LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platfo…
arXiv cs.AI Research & Papers
LLM4LLM: Bridging Kernel Benchmarks and Real Deployment via Closed-Loop Agentic Optimizat…
arXiv cs.AI Research & Papers
EDGE: Experience-Distillation for Guided Exploration in Agentic Reinforcement Learning