Making Models Unmergeable via Scaling-Sensitive Loss Landscape
Explorar
Noticias de IA
22116 elementos — filtrados, clasificados y sin duplicados
Calibration Drift Under Reasoning: How Chain-of-Thought Budgets Induce Overconfidence in …
Substrate Asymmetry in User-Side Memory: A Diagnostic Framework
Physics-Distilled Neural Network enabled by Large Language Models for Manufacturing Proce…
Unifying Learning Dynamics and Generalization in Transformers Scaling Law
Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning
ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generati…
FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse
RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways
Semantic search for 100M+ galaxy images using AI-generated captions
The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independen…
MA-DLE: Speech-based Automatic Depression Level Estimation via Memory Augmentation
Grounding Computer Use Agents on Human Demonstrations
Federated continual learning: A comprehensive survey on lifelong and privacy-preserving l…
Quantized Stochastic Primal-Dual Methods for Distributed Optimization under Relaxed Globa…
Causal Emotion Recognition in Conversation: Context Saturation and Discourse-Marker Evide…
To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending
Nonslop: A Gamified Experiment in Human-AI Collaborative Writing
Dual-Stance Evaluation of Sycophancy: The Structure of Agreement and the Limits of Interv…
BioDivergence: A Benchmark and Evaluation Framework for Hidden Contextual Contradictions …
Interpretable Factor Decomposition for Decision Intelligence in Large-Scale Financial Mar…
CLARITree: Cholesky and Lookahead Accelerations for Regression with Interpretable Piecewi…
Graph Reinforcement Learning for Calibration-Aware Quantum Circuit Routing
SymQNet: Amortized Acquisition for Low-Latency Adaptive Hamiltonian Learning
Agentic MPC for Semantic Control System Resynthesis
How an astrophysicist uses Codex to help simulate black holes
Context-Driven Incremental Compression for Multi-Turn Dialogue Generation
System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5
Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling
Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics