ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without B…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
ROBUST-WT: Robust Uncertainty-aware Segmentation Transform via Whitening and Training Enh…
Constitutional On-Policy Safe Distillation
Regret Pre-training: Bridging Prior and Posterior Views for Enhanced Knowledge Grounding
Libra: Efficient Resource Management for Agentic RL Post-Training
NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Si…
Decoupled Smart Contract Audits: Lightweight LLM Framework via Distillation and Aggregati…
AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Foll…
PhotoCraft: Agentic Reasoning with Hierarchical Self-Evolving Memory for Deep Image Search
Reinforcement Learning from Cross-domain Videos with Video Prediction Model
Fully Automated Identification of Lexical Alignment and Preference-Stage Shifts in Large …
AirDreamer: Generalist Drone Navigation with World Models
GFFMERGE: Efficient Merging of Graph Neural Force Fields and Beyond
Are Common Substructures Transferable? Riemannian Graph Foundation Model with Neural Vect…
Learning Multi-Scale Hypergraph for High-Order Brain Connectivity Analysis
Message Tuning Outshines Graph Prompt Tuning: A Prismatic Space Perspective
dstack-capsule: Pod-Level Remote Attestation for Confidential Workloads on Kubernetes
Multi-Modal Graph Neural Network with Transformer-Guided Adaptive Diffusion for Preclinic…
The Unsampled Truth: Psychometrics in SLMs Measure Prompt Artifacts, Not Psychological Co…
AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation…
Causal Evidence of Stack Representations in Modeling Counter Languages Using Transformers
When Model Merging Breaks Routing: Training-Free Calibration for MoE
Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Ge…
PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy…
SPADE: Sketch-guided Path Planning Augmented with Diffusion Experts
Efficient Transformer-Based Localized Patch Sampling for Choroid Plexus Segmentation in M…
When Should the Teacher Move? Temporal Coupling and Stability in Self On-Policy Distillat…
Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks
CauTion: Knowing When to Trust LLMs for Ensemble Causal Discovery
PHASER: Phase-Aware and Semantic Experience Replay for Vision-Language-Action Models