DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
From AR to Diffusion: Efficiently Adapting Large Language Models with Strictly Causal and…
From Instructor to Collaborator: What a 90-Participant Study Reveals about Human-Agent Co…
The Alignment Floor: When Persona Customization Is Safe
SafeMed-R1: Clinician-Audited Safety and Ethics Alignment for Medical Large Language Mode…
BioELX: Cross-lingual Biomedical Entity Linking via Alias-based Retrieval and LLM Ranking
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-t…
Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR
SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Netwo…
Utility-Aware Multimodal Contrastive Learning for Product Image Generation
LiveBrowseComp: Are Search Agents Searching, or Just Verifying What They Already Know?
The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic
Global Policy-Space Response Oracles for Two-Player Zero-Sum Games
Bandwidth-Efficient and Privacy-Preserving Edge-Cloud Many-to-Many Speech Translation
Adaptive Multimodal Agents-Based Framework for Automatic Workflow Execution
Modeling Vehicle-Type-Specific Pedestrian Crash Avoidance Behavior in Safety-Critical Int…
Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Questi…
Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning
The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level En…
Eliot: Interactively $\underline{E}$xploring Fast-Changing Scientific $\underline{Li}$ter…
How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks
Worker Disagreement Reveals Sharp Directions in Local SGD
Mahalanobis PatchCore: Covariance-Aware and Streaming-Compatible Industrial Anomaly Detec…
Restoring the Sweet Spot: Pass-Rate Weighted Self-Distillation for LLM Reasoning
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
Turning Video Models into Generalist Robot Policies
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
Score Based Error Correcting Code Decoder
Fine-Tuned LLM as a Complementary Predictor Improving Ads System