The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven…
GONDOR to the Rescue: Satisficing Planning with Low Memory
You Live More Than Once: Towards Hierarchical Skill Meta-Evolving
Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs
Prompt Codebooks: Discrete Compositional Optimization for Language Model Instruction Refi…
Picid: A Modular Evaluation Infrastructure for Reproducible PHM Across Tasks and Domains
On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective
PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting
Visualizing Latent Phase Structures in Locomotion Policies: A Multi-Environment Study wit…
EAGer: Entropy-Aware GEneRation for Adaptive Inference-Time Scaling
The Principles of Diffusion Models
Multi-Agent LLM-based Metamorphic Testing for REST APIs
The Cases LJP Never Sees: Prosecution Decision Prediction for More Complete Criminal Liab…
Stochastic Gradient Descent with Momentum is Algorithmically Stable
A Multi-dimensional Framework for Evaluating Generalization in EEG Foundation Models
Models That Know How Evaluations Are Designed Score Safer
Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification
Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning
Towards Reliable Multilingual LLMs-as-a-Judge: An Empirical Study
Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Mod…
Heterogeneous Causal Discovery of Repeated Undesirable Health Outcomes
DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision…
Text2Model: Modeling Copilots for Text-to-Model Translation
Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM A…
GradientStabilizer:Fix the Norm, Not the Gradient
STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation
Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Soko…
Probability-Entropy Calibration: An Elastic Indicator for Adaptive Fine-tuning
ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process…