When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for O…
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Mod…
From Paper to Program: Knowledge Externalization for AI-Assisted Quantum Many-Body Code G…
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic…
Evaluating Interactive 2D Visualization as a Sample Selection Strategy for Biomedical Tim…
Beyond MACs: Hardware Efficient Architecture Design for Vision Backbones
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embeddi…
Decidable By Construction: Design-Time Verification for Trustworthy AI
CogGen: Cognitive-Load-Inspired Fully Unsupervised Deep Generative Modeling for Compressi…
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
Position: Modular Memory is the Key to Continual Learning Agents
GOT-JEPA: Generic Object Tracking with Model Adaptation and Occlusion Handling using Join…
Brep2Shape: Boundary and Shape Representation Alignment via Self-Supervised Transformers
Optimism Stabilizes Thompson Sampling for Adaptive Inference
PLATE: Plasticity-Tunable Efficient Adapters for Geometry-Aware Continual Learning
R1-SyntheticVL: Is Synthetic Data from Generative Models Ready for Multimodal Large Langu…
LVLMs and Humans Ground Differently in Referential Communication
m2sv: A Scalable Benchmark for Map-to-Street-View Spatial Reasoning
Vulcan: Instance-specialized, Verifiable Systems Heuristics Through LLM-driven Search
Prototype-Based Semantic Consistency Alignment for Domain Adaptive Retrieval
First, do NOHARM: towards clinically safe large language models
EngTrace: A Symbolic Benchmark for Verifiable Process Supervision of Engineering Reasoning
A geometric and deep learning reproducible pipeline for monitoring floating anthropogenic…
Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization
RooseBERT: A New Deal For Political Language Modelling
Blueprint First, Model Second: A Framework for Deterministic LLM Workflow
Detail++: Training-Free Detail Enhancer for Text-to-Image Diffusion Models
AnalogFed: Privacy-Preserving Discovery of Analog Circuits at Scale with Federated Genera…
A Gradient-based Causal Discovery Framework with Applications to Complex Industrial Proce…
Gaussian DP for Reporting Differential Privacy Guarantees in Machine Learning