A Model Can Help Itself: Reward-Free Self-Training for LLM Reasoning
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of A…
SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natura…
TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Releva…
Verifier-free Test-Time Sampling for Vision-Language-Action Models
OrthoReg: Orthogonal Regularization for Hybrid Symbolic-Neural Dynamical Systems
TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech
Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observat…
Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs fo…
TransitNet: A Compact Attention-Augmented Deep Learning Framework for Low-SNR Transit Bli…
Clin-JEPA: A Multi-Phase Co-Training Framework for Joint-Embedding Predictive Pretraining…
Evolutionary Ensemble of Agents
Automating the Design of Embodied Agent Architectures
The Unverifiability of Artificial General Intelligence (AGI) Alignment, Static and Dynami…
A Stochastic--Geometric Theory of Scaling Laws in Grokking
Curvature-Guided Sheaf Diffusion for Unsupervised Community Detection on Heterophilic Gra…
BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Mul…
IRC-Bench: Recognizing Entities from Contextual Cues in First-Person Reminiscences
A Deep Multiscale Neural Network for Accurate Neurological Disorder Detection from MRI Sc…
ARISE: A Repository-level Graph Representation and Toolset for Agentic Program Repair and…
Fitting Horn DL Ontologies to ABox and Query Examples: A Tale of Simulation Quantifiers a…
Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Ca…
Kwai Summary Attention Technical Report
Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidenc…
CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors
The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier M…
Towards Generalizable Deepfake Image Detection with Vision Transformers
Embodied Operators and Benchmarking: Toward Reusable and Deployable Embodied Intelligence…
MentalThink: Shaping Thoughts in Mental SVG World
How to Avoid Debate: Scalable AI Safety via Doubly-Efficient Interactive Proofs