Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training
arXiv cs.AI Research & Papers
MedCalc-R1: Knowledge-Guided Reward Framework for Medical Mathematical Reasoning
arXiv cs.AI Research & Papers
Test-Time Augmentation for LLMs: When Input Diversity Beats Output Diversity at Matched C…
arXiv cs.AI Research & Papers
From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs
arXiv cs.AI Research & Papers
Quantization Degradation in Large Language Models: A Signal-Noise Perspective
arXiv cs.AI Research & Papers
Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent G…
arXiv cs.AI Research & Papers
Three Necessary Principles for Self-Supervised Visual Representation Learning
arXiv cs.AI Research & Papers
A Fair Objective for Human-Empowerment-Preserving AI: Desiderata, Design, and Likely Beha…
arXiv cs.AI Security & Safety
Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Sce…
arXiv cs.AI Research & Papers
Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset…
arXiv cs.AI Research & Papers
Long SKILL Compliance as Logical Reasoning: Closure-Grounded Detection with Scaling-Guide…
arXiv cs.AI Research & Papers
AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts
arXiv cs.AI Research & Papers
Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks
arXiv cs.AI Research & Papers
Can Coding Agents Solve Repository-Level Issues with Rendered Code? An Exploratory Study …
arXiv cs.AI Research & Papers
GRASP: Granularity-Aware Region Alignment and Semantic Prototype Learning for Fine-Graine…
arXiv cs.AI Research & Papers
How Simple Can It Get? From Interpretable Equations to Readable Rules for Financial Decis…
arXiv cs.AI Research & Papers
Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Disti…
arXiv cs.AI Research & Papers
FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning
arXiv cs.AI Research & Papers
FeedbackTrack: Visual-Cortex-Inspired Cross-Frame Feedback for Transformer Tracking
arXiv cs.AI Research & Papers
Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Sca…
arXiv cs.AI Research & Papers
Concept-Guided Spatial Regularization for World Models in Atari Pong
arXiv cs.AI Research & Papers
PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering
arXiv cs.AI Research & Papers
Second Order Drifting Models
arXiv cs.AI Research & Papers
Ground-Truth Neighborhood Regularization for Reinforcement Learning Post-Training of Time…
arXiv cs.AI Research & Papers
EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference
arXiv cs.AI Research & Papers
Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questi…
arXiv cs.AI Research & Papers
See Me, Believe Me: Causality, Intersectionality, and Interventions Improving the Appeara…
arXiv cs.AI Research & Papers
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
arXiv cs.AI Research & Papers
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
arXiv cs.AI Research & Papers
Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interacti…