Explorar

Noticias de IA

21863 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
UR$^2$: Unify RAG and Reasoning through Reinforcement Learning
arXiv cs.AI Research & Papers
Lean-GAP: A Dataset of Formalized Graduate Algebra Problems
arXiv cs.AI Research & Papers
ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models
arXiv cs.AI Research & Papers
Learning the Neighborhood: Contrast-Free Multimodal Self-Supervised Molecular Graph Pretr…
arXiv cs.AI Research & Papers
The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs
arXiv cs.AI Research & Papers
When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning
arXiv cs.AI Research & Papers
Solipsistic Superintelligence is Unlikely to be Cooperative
arXiv cs.AI Research & Papers
Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning
arXiv cs.AI Research & Papers
Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Pers…
arXiv cs.AI Research & Papers
ChatHealthAI: Aligning Electronic Health Record Representations with Large Language Model…
arXiv cs.AI Research & Papers
Evaluating Transformer and LSTM Frameworks for Prediction in Ungauged Basins
arXiv cs.AI Research & Papers
Safety Measurements for Fine-tuned LLMs Should be Grounded in Capability
arXiv cs.AI Research & Papers
SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural …
arXiv cs.AI Research & Papers
Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Hu…
arXiv cs.AI Research & Papers
Toward Training Superintelligent Software Agents through Self-Play SWE-RL
arXiv cs.AI Research & Papers
Edge-Aware and Content-Adaptive Infrared Gas Leak Detection for Industrial Safety Monitor…
arXiv cs.AI Research & Papers
Physics-Guided Policy Optimization with Self-Distillation
arXiv cs.AI Research & Papers
FLIPS: Instance-Fingerprinting for LLMs via Pseudo-random Sequences
arXiv cs.AI Research & Papers
Relational Linearity is a Predictor of Hallucinations
arXiv cs.AI Research & Papers
Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Mod…
arXiv cs.AI Research & Papers
RobotValues: Evaluating Household Robots When Human Values Conflict
arXiv cs.AI Research & Papers
AI-Generated Traces for Novice Programmers: Learning Effects and Learner Differences in a…
arXiv cs.AI Research & Papers
Coupled Local and Global World Models for Efficient First Order RL
arXiv cs.AI Research & Papers
LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
arXiv cs.AI Research & Papers
PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classifi…
arXiv cs.AI Research & Papers
InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning
arXiv cs.AI Research & Papers
WebRISE: Requirement-Induced State Evaluation for MLLM-Generated Web Artifacts
arXiv cs.AI Research & Papers
AI Rater Discrimination Depends on Scoring Protocol in Complex Clinical Decision-Making
arXiv cs.AI Research & Papers
Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path…
arXiv cs.AI Research & Papers
Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapte…