Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation ac…
Explorar
Noticias de IA
21270 elementos — filtrados, clasificados y sin duplicados
ReCA: Multi-Shot Long Video Extrapolation via Recursive Context Allocation
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
DSA-Tokenizer: Disentangled Semantic-Acoustic Tokenization via Flow Matching-based Hierar…
Can LLMs Introspect? A Reality Check
Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning
Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interac…
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
Adaptive Multi-prompt Contrastive Network for Few-shot Out-of-distribution Detection
A Physics-Informed Hierarchical Neural Network for Microwave Scattering Analysis of 3D PE…
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
Monte Carlo Permutation Search
SWAP: Towards Copyright Auditing of Soft Prompts via Sequential Watermarking
Spend Your Rollouts Where It Counts: Rollout Allocation for Group-Based RL Post-Training
Trust Region Q Adjoint Matching
More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation …
Beyond Trajectory-Level Attribution: Graph-Based Credit Assignment for Agentic Reinforcem…
Measuring Prediction Uncertainty in Neural Cellular Automata
ICCU: In-Context Continual Unlearning via Pattern-Induced Refusal Rules
DEI: Diversity in Evolutionary Inference for Quality-Diversity Search
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
Innovation: An Almost Characterization of Hallucination
The ATOM Report: Measuring the Open Language Model Ecosystem
Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study
The Strongest Teacher Is Not Always the Best Teacher: Student-Centric Answer Selection
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentat…