Failing Gracefully: Mitigating Impact of Inevitable Robot Failures
Explorar
Noticias de IA
29347 elementos — filtrados, clasificados y sin duplicados
HoloCount: A Holistic Visual Counting Benchmark for MLLMs
The Ignition Index: Measuring Global Workspace Dynamics in Language Models
Fast Rates for Inverse Reinforcement Learning
HarnessOpt-Bench: Evaluating LLMs at Harness Optimization
Bayesian Expected Uncertainty Reduction (B-EUR) Model: A Computational Account of What Ma…
RA-CAD: Learning Post-Execution Critique for State-Aware Text-to-CAD Generation
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
Ge$^\text{2}$mS-T: Multi-Dimensional Grouping for Ultra-High Energy Efficiency in Spiking…
Estimating time spent on work tasks
Challenges for Musical Education in the Age of AI and Digital Transformation
Autonomous Research Agents: A Survey of AI Scientists and the Verification Gap
ECHO: A Locally-Deployable Agentic Health Assistant with Temporal Memory, Safety Guardrai…
Trust-Based Incentive Mechanisms in Semi-Decentralized Federated Learning Systems
An Optimal Agnostic PAC Algorithm
SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploit…
Beyond Information Retrieval: Generative AI as an Epistemic Arbiter to Enhance Collaborat…
Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Applic…
CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Pe…
IMMENSE: Inductive Multi-perspective User Classification in Social Networks
BlockPython: A Process-Aware Agent-Supported Platform for the Transition from Block-Based…
Text Steganography with Dynamic Codebook and Multimodal Large Language Model
Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New S…
Otter: A Time-Aware, History-Conditioned Human Chess AI
LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs
Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforceme…
One Qubit Can Beat One Bit: Quantum Advantage for Post-Training Quantization
SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse
PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis
Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to …