Explorar

Noticias de IA

30239 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
CoT-Core: Accelerating LLM Evaluation via CoT-Aware Coreset Selection
arXiv cs.AI Research & Papers
402Pilot: An x402 Decision Layer for Autonomous Agent Micropayments
arXiv cs.AI Research & Papers
High-Stakes Decisions with Language Models: Insights from Emergency Triage
arXiv cs.AI Research & Papers
Characterizing Readability Issue Patterns and the Role of Prompt Design in LLM-Generated …
arXiv cs.AI Research & Papers
Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning
arXiv cs.AI Research & Papers
LakeMLB: Data Lake Machine Learning Benchmark
arXiv cs.AI Research & Papers
Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tut…
arXiv cs.AI Research & Papers
RAG-TESTER: Automated End-to-End Testing of Retrieval-Augmented Large Language Models
arXiv cs.AI Research & Papers
Computational Approaches to Understanding Large Language Model Impact on Writing and Info…
arXiv cs.AI Research & Papers
TCPO: Turn-Level Credit Policy Optimization
arXiv cs.AI Research & Papers
From We to Me: Theory Informed Narrative Shift with Abductive Reasoning
arXiv cs.AI Research & Papers
A Contractualist Argumentation Framework for Moral Decision-Making
arXiv cs.AI Research & Papers
Where Reasoning Diverges: Localized Multi-Agent Debate
arXiv cs.AI Research & Papers
From AI Technical Debt to Agentic Technical Debt: A Systematic Mapping of Root Causes and…
arXiv cs.AI Research & Papers
Role-Decoupled Attention Residuals: Separating Matching and Content Retrieval Across Depth
arXiv cs.AI Research & Papers
Oscillatory Hierarchical Reservoirs for Human-like Rhythm Perception and Anticipation
arXiv cs.AI Research & Papers
Cross-Benchmark Generalization in Long-Horizon Agents
arXiv cs.AI Research & Papers
Evolutionary Curriculum Learning Improves Biological Sequence Modeling
arXiv cs.AI Research & Papers
Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models
arXiv cs.AI Research & Papers
Interpretable Recognition of Cognitive Distortions in Natural Language Texts
arXiv cs.AI Research & Papers
Judging Is Not Enumerating: Silent Omissions in LLM-Authored Acceptable Sets
arXiv cs.AI Research & Papers
Escaping Confidence Trap: Evolutionary Decoding for Mathematical Reasoning in Diffusion L…
arXiv cs.AI Research & Papers
Auditing Discovery Claims: A Two-Sided Criterion for Agentic Science, with the Negative S…
arXiv cs.AI Research & Papers
SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents
arXiv cs.AI Research & Papers
RAP: KV-Cache Compression via RoPE-Aligned Pruning
arXiv cs.AI Research & Papers
Large language models improve physician accuracy but lead to false reliance
arXiv cs.AI Research & Papers
Passing Coarse Marginal Checks Can Be Cheap: Persona Mixtures and Imprecise Treatment-Res…
arXiv cs.AI Research & Papers
When LLM Essays Outscore Student Essays: What a Korean Writing Rubric Rewards and Where R…
arXiv cs.AI Research & Papers
Request-Level Energy Attribution for Batched LLM Serving
arXiv cs.AI Research & Papers
Auditing Semantic Gains in Sequential Recommendation: A Lightweight Recovery Test