Explorar

Noticias de IA

22318 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Safety from Honesty in a Disinterested AI Predictor
arXiv cs.AI Research & Papers
Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-G…
arXiv cs.AI Research & Papers
SPARK: Susceptibility-Guided Profiling and Steering of Latent Reasoning States in Large L…
arXiv cs.AI Research & Papers
Measure the Sim-to-Real Gap: Designing an Affordable Real-World Benchmark Platform for Re…
arXiv cs.AI Research & Papers
ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory
arXiv cs.AI Research & Papers
Toward Inclusive Avatar Design with Limb Differences Through Artificial Intelligence
arXiv cs.AI Research & Papers
Interpreting Latent CoT Reasoning as Dynamical Systems
arXiv cs.AI Research & Papers
From ML Predictions to Informed Diagnostic Assistance Using the Toulmin Model of Argument…
arXiv cs.AI Research & Papers
CMSL: Constructive Multi-Sequence Learning for Recommendation Systems
arXiv cs.AI Research & Papers
RankGraph-2: Lifecycle Co-Design for Billion-Node Graph Learning in Recommendation
arXiv cs.AI Research & Papers
Prompt Compression in Diffusion Large Language Models: Evaluating LLMLingua-2 on LLaDA
arXiv cs.AI Research & Papers
Graph Construction and Matching for Imperative Programs using Neural and Structural Metho…
arXiv cs.AI Research & Papers
Stable On-Policy Distillation through Adaptive Target Reformulation
arXiv cs.AI Research & Papers
Enhancing Adversarial Transferability through Block Stretch and Shrink
arXiv cs.AI Research & Papers
Graph Optimization Foundation Model: Tokenizing Graph via A Language-Model Paradigm
arXiv cs.AI Research & Papers
The Verifier is the Curriculum: Execution-Gated Self-Distillation for Cross-Family Game G…
arXiv cs.AI Research & Papers
GES-TSP: Graph Edge Sparsification for TSP
arXiv cs.AI Research & Papers
Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces
arXiv cs.AI Research & Papers
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness…
arXiv cs.AI Research & Papers
ECoLAD: Selecting Anomaly Detectors for Automotive Deployment via Compute-Reduction Evalu…
arXiv cs.AI Research & Papers
AutoMatBench: An Automatic Optimization Toolkit for the Acceleration of Material Properti…
arXiv cs.AI Research & Papers
Heuristic Learning for Active Flow Control Using Coding Agents
arXiv cs.AI Research & Papers
Extending LLM Context via Associative Recurrent Memory
arXiv cs.AI Research & Papers
PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour
arXiv cs.AI Research & Papers
Constrained Reinforcement Learning for Safe Heat Pump Control
arXiv cs.AI Research & Papers
Asynchronous Perception Machine For Efficient Test-Time-Training
arXiv cs.AI Research & Papers
Efficient Q-Learning and Actor-Critic Methods for Robust Average-Reward Reinforcement Lea…
arXiv cs.AI Research & Papers
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software …
arXiv cs.AI Research & Papers
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Of…
arXiv cs.AI Research & Papers
PiCSAR: Probabilistic Confidence Selection And Ranking for Reasoning Chains