Explorar

Noticias de IA

22318 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulat…
arXiv cs.AI Research & Papers
PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation
arXiv cs.AI Research & Papers
Uncovering Latent Depression Severity for Binary Depression Detection via Advantage-weigh…
arXiv cs.AI Research & Papers
StateFuse: Deterministic Conflict-Preserving Memory for Multi-Agent Systems
arXiv cs.AI Research & Papers
LLM-Guided Measurement Credibility Correction for Trustworthy Industrial Process Inference
arXiv cs.AI Research & Papers
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Traini…
arXiv cs.AI Research & Papers
From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structure…
arXiv cs.AI Research & Papers
Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Larg…
arXiv cs.AI Research & Papers
Synthetic Consumer Insight Generation with Large Language Models
arXiv cs.AI Research & Papers
ArtisanCAD: An Industrial-Level CAD Agent with Expert-Grounded Knowledge Distillation
arXiv cs.AI Research & Papers
Memory in the Loop: In-Process Retrieval as ExtendedWorking Memory for Language Agents
arXiv cs.AI Research & Papers
FirstResearch: Auditable Question Formation for LLM Scientific Discovery Agents
arXiv cs.AI Research & Papers
CSTutorBench: Benchmarking Small Language Models as Tutors for Block-Based Programming
arXiv cs.AI Research & Papers
Reduced NEXI protocol for the quantification of human gray matter microstructure on the C…
arXiv cs.AI Research & Papers
Toward AI standardization: A triadic human-ai collaboration framework for multi-level aut…
arXiv cs.AI Research & Papers
Detoxify: A framework for abusive text transformation using LLMs
arXiv cs.AI Research & Papers
LLM4Delay: Flight Delay Prediction via Cross-Modality Adaptation of Large Language Models…
arXiv cs.AI Research & Papers
Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance
arXiv cs.AI Research & Papers
BaFCo: A Document Understanding Benchmark for Complex Bangla Form Comprehension
arXiv cs.AI Research & Papers
EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems
arXiv cs.AI Research & Papers
Do It Right! A Methodology for Successful NLP System Development
arXiv cs.AI Research & Papers
FORGE: Towards Functional Tool-Use Generalization via Keypoint Trajectory Reasoning
arXiv cs.AI Research & Papers
IMR: Iterative Mode-World Weighted Regression for Multi-Agent Trajectory Prediction
arXiv cs.AI Research & Papers
Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An…
arXiv cs.AI Research & Papers
Data-dependent Evaluations for Budgeted Submodular Maximization
arXiv cs.AI Research & Papers
LEGATO 2: Toward Multimodal Sheet Music Recognition and Understanding
arXiv cs.AI Research & Papers
When Should LLMs Search? Counterfactual Supervision for Search Routing
arXiv cs.AI Research & Papers
SCOReD: Student-Aware CoT Optimization for Recommendation Distillation
arXiv cs.AI Research & Papers
Complementary Roles of Image Classification and Vessel Segmentation in AI-Based Screening…
arXiv cs.AI Research & Papers
Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability An…