RA-DCA: A Randomized Active-Set DCA for Directional Stationarity in Max-Structured DC Pro…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Goal-Conditioned Agents that Learn Everything All at Once
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide …
Understanding Goal Generalisation in Sequential Reinforcement Learning
HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retriev…
OnePred: Next-Query Prediction via Recursive Intent Memory in Multi-Turn Conversations
PhotoFlow: Agentic 3D Virtual Photography Missions
CHRONOS: Temporally-Aware Multi-Agent Coordination for Evolving Data Marketplaces
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transform…
ETCHR: Editing To Clarify and Harness Reasoning
LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws
Interactive Query Answering on Knowledge Graphs with Soft Entity Constraints
Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quan…
MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchest…
ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation
NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG…
Agentivism: a learning theory for the age of artificial intelligence
Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure
Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning
ConjNorm: Tractable Density Estimation for Out-of-Distribution Detection
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
Naturalistic Computational Cognitive Science: Towards generalizable models and theories t…
GlyTwin: Digital Twin for Glucose Control in Type 1 Diabetes Through Optimal Behavioral M…
GeoMAE: Masking Representation Learning for Spatio-Temporal Graph Forecasting with Missin…
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
GILT: An LLM-Free, Tuning-Free Graph Foundational Model for In-Context Learning
Sparser Block-Sparse Attention via Token Permutation
Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures
DocVAL: Validated Chain-of-Thought Distillation for Grounded Document VQA
Operator-Based Generalization Bound for Deep Learning: Insights on Multi-Task Learning