R$^3$L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivot…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Bridging Data and Physics: A Graph Neural Network-Based Hybrid Twin Framework
Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning
IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents
Bridging AI and Clinical Reasoning: Abductive Explanations for Alignment on Critical Symp…
MedSAE: Dissecting MedCLIP Representations with Sparse Autoencoders
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Sys…
Robust Counterfactual Inference in Markov Decision Processes
PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs
It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amp…
Any2Any: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking
Learning Through Noise: Why Subliminal Learning Works and When It Fails
Does Your Wildfire Prediction Model Actually Work, or Just Score Well?
WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native …
How Mobile World Model Guides GUI Agents?
Towards Trustworthy and Explainable AI for Perception Models: From Concept to Prototype V…
ProtDBench: A Unified Benchmark of Protein Binder Design and Evaluation
A Comparative Analysis on the Performance of Upper Confidence Bound Algorithms in Adaptiv…
Decompose, Structure, and Repair: A Neuro-Symbolic Framework for Autoformalization via Op…
MemReward: Graph-Based Experience Memory for LLM Reward Prediction with Limited Labels
Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling
V-VLAPS: Value-Guided Planning for Vision-Language-Action Models
Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Ap…
STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction
Socially fluent AI decouples conversational signals from source identity in online intera…
Empowering 9-1-1 Calltaking Training with Generative AI: Experiences and Lessons Learned
XAttnMark: Learning Robust Audio Watermarking with Cross-Attention
Is Capability a Liability? More Capable Language Models Make Worse Forecasts When It Matt…
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning
Variance Reduction for Expectations with Diffusion Teachers