Topology-Guided Modular Actor-Critic Learning for Continuous Systems under Temporal Objec…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Preference Data Selection for Mitigating the Alignment Tax in Large Language Models
Ollivier-Ricci Curvature of Riemannian Manifolds and Directed Graphs with Applications to…
StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments
ShardMeter: Sharded and Geo-Distributed Training Without the Guesswork
ICA: Information-Aware Credit Assignment for Visually Grounded Long-Horizon Information-S…
More Rejective, Not More Discriminative: The Unit of Verification in Pre-Execution LLM Ov…
STAIN-FL: Stealthy Targeted Attack Injection with Contextual Triggers in Federated Learni…
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
FLARE: A Systematic, Uncertainty-Aware Framework for Evidence-Based Adoption of Artificia…
A tale of perfect fit and phantom optima: how data-driven models can fail in real-time op…
VideoHarness-RSI: Recursive Harness Self-Improvement for Long-Video Understanding with Fr…
Elastic KV Cache for LLM Serving:A Working Reclamation Mechanism, and Why Chunked Prefill…
Efficient LLM Collaboration via Planning
Taming foundation model with invariance-oriented pre-training for broad-spectrum EEG anal…
The Handoff Tax: Continuing Non-Native Trajectories in LLM Agents
Constrained Hyperparameter Optimization for Streaming Data
AI Finds A Way
Right Diagnoses, Decorative Reasoning:A Perturbation Audit of Medical Chain-of-Thought
A Literate Programming Environment for Human and Machine Agents
FARCA: Fact-Aligned Reliability-Aware Credit Assignment for Reinforcement Learning with F…
Memory Is Not Always Needed: Characterizing Conditional Memory in Scientific Reasoning
Beyond Accuracy: A Dual-Judge Evaluation Protocol for Vision-Language Models in Legally G…
Tlow: Flow-based Item Tokenizer for Recommendation
Deep Learning Super Resolution for Satellite Cloud Mask Downscaling
PROOF-Gen: From Optimized Data to Better Distillation
Simthesizer: An Agent-Driven Simulation Framework for LLM Serving Systems
LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding
Diverse by Reasoning: Harnessing the Wisdom of LLM Crowds for Future Prediction
Mahalanobis-Based Multi-Head Attention for Complex State Propagation