Post-Training Pruning for Diffusion Transformers
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL
MalariAI: A Label-Resilient Decoupled Framework for Universal Cell Segmentation and Expla…
Learning to Compose: Revisiting Proxy Task Design for Zero-Shot Composed Image Retrieval
MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixtur…
A Category Theory Account of AI Identity
AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation
Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech…
DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoni…
K-Inverse-RFM: A Modified RFM that Bridges the Gap to Neural Networks for Data-Corrupted …
RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail
Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stabili…
Learning When to Listen: Gated Affect Fusion for Human Motion Prediction
What's Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Languag…
Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds
Entropy-Regularized Probabilistic Gates for Sparse Model Discovery in Scarce-Data Federat…
Human-Machine Collaboration on Generative Meta-Learning: Model and Algorithm
ASPIRE: Agentic /Skills Discovery for Robotics
Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification
Leveraging Phase Information to Boost Unrolled Network Learning for Image Deblurring
Adaptive Perturbation Selection for Contrastive Audio Decoding
REALM: An RGB- and Event-Aligned Latent Manifold for Cross-Modal Perception
Scaling Up Thermodynamic AI Models
EgoSafetyBench: A Diagnostic Egocentric Video Benchmark for Evaluating Embodied VLMs as R…
SLIM-RL: Risk-Budgeted Random-Masking RL for Diffusion LLMs Without Trajectory Slicing
Play Like Champions: Counterfactual Feedback Generation in Latent Space
From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narr…
LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection
GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation…
A Mechanism-Driven Theory of Phase Transitions in Active Learning