Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning
Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched C…
From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs
Quantization Degradation in Large Language Models: A Signal-Noise Perspective
Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent G…
Three Necessary Principles for Self-Supervised Visual Representation Learning
A Fair Objective for Human-Empowerment-Preserving AI: Desiderata, Design, and Likely Beha…
Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Sce…
Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset…
Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guide…
AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts
Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks
Can Coding Agents Solve Repository-Level Issues with Rendered Code? An Exploratory Study …
GRASP: Granularity-Aware Region Alignment and Semantic Prototype Learning for Fine-Graine…
How Simple Can It Get? From Interpretable Equations to Readable Rules for Financial Decis…
Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Disti…
FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning
FeedbackTrack: Visual-Cortex-Inspired Cross-Frame Feedback for Transformer Tracking
Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Sca…
Concept-Guided Spatial Regularization for World Models in Atari Pong
PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering
Second Order Drifting Models
Ground-Truth Neighborhood Regularization for Reinforcement Learning Post-Training of Time…
EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference
Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questi…
See Me, Believe Me: Causality, Intersectionality, and Interventions Improving the Appeara…
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interacti…