Ground-Truth Neighborhood Regularization for Reinforcement Learning Post-Training of Time…
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
Context Is Not Authority: Structured Runtime Governance for Financial Market Agents
MoRSE: Task-Oriented Multi-Agent System with Mixture of Role-Subtask Experts
See Me, Believe Me: Causality, Intersectionality, and Interventions Improving the Appeara…
Ethical Framework for Responsible Foundational Models in Medical Imaging
Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and …
Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adap…
Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Infere…
HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model …
PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling
Epistemic Transfer in AI-Assisted Verification: A Framework and Evaluation Protocol
FedTVD: Balancing Data Quality and Quantity for Robust Federated Learning
Integrated Multimodal AI System for Retrieval-Augmented Reasoning, Object Sensing, and Da…
Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution
SuperNeuroMAT: An Efficient Matrix-based Simulator for Spiking Neural Networks
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interacti…
VeinCast: Physics-Guided Dynamic Field Graphs with Graph-Conditioned Fusion for Global Me…
Fourier Self-Supervision for Fine-Grained Generalized Category Discovery
From Alignment to Synthesis Contrastive Volumetric Grounding for Text-to-CT Generation
Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discove…
Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent G…
$\texttt{DisMorph}$: learning to disentangle technical distortions from true biological c…
RAVEN-Eval: Rubric-Guided Automatic Evaluation for AI Video Generation Models Based on LM…
When Latents Forget Pixels: Restoring Fidelity in Diffusion Transformer Super-Resolution
Enhanced Real-Time 6-DOF Extended Reality Catheter Tracking for Evaluating Potential Impr…
Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Pol…
Temporal Sepsis Modeling: a Relational and Explainable-by-Design Framework
LGNNIC: Acceleration of Large-Scale GNN Training using SmartNICs
CoRE: Consensus Rewards via Equilibrium for Test-Time Reinforcement Learning