Enabling KV Caching of Shared Prefix for Diffusion Language Models
Explorar
Noticias de IA
29629 elementos — filtrados, clasificados y sin duplicados
SurfDesign: Effective Protein Design on Molecular Surfaces
Considerations for an Integrated Detector Design at FCC-ee: A Human-AI Exploration
The Montparnasse Algorithm for RNA Design
Semantic Cache Distillation: Efficient State Transfer via Reuse and Selective Patching
SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?
Can You Trust What You See? Human and AI Detection of Synthetic Legal Evidence
LEAF: Growing Trees Without Branching for Speech-Aware Large Language Model Post-Training
Contribution Weights: A Geometrical Analysis of Self-Attention Transformers
Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning
Reachability and asymptotics of Gaussian Transformer dynamics
Repetition Mismatch: Why Data Mixture Experiments Don't Scale and How to Fix Them
VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Age…
MedicalRec: Medical recommender system for image classification without retraining
Concerns and Strategic Responses of Older Workers Navigating Generative AI in Bridge Empl…
Self-Explainability in Self-Adaptive and Self-Organising Systems: Status and Research Dir…
From Rigid to Dynamic: Entropy-Guided Adaptive Inference for Long-Context LLMs
LLM-Orchestrated Conformance Checking in Stroke Care Without Computer-Interpretable Guide…
TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics
Bayesian Selective Latent Inference for Wastewater-First Influenza Monitoring
From Coarse to Fine: Managing Temporal Granularity in Spatio-Temporal Data for Fine-Grain…
Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs
Leveraging Structural Constraints for Diffusion-based Neural TSP Solvers
Anything2Skill: Compiling External Knowledge into Reusable Skills for Agents
FF-JEPA: Long-Horizon Planning in World Models with Latent Planners
IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Gener…
A Regret Minimization Framework on Preference Learning in Large Language Models
Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts
Can the Environment Speak for Itself? $T^{2}$-GRPO: A Turn-Trajectory Group Relative Poli…
Instrumental convergence and power-seeking