LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
You Don't Need Attention: Gated Convolutional Modeling for Watch-Based Fall Detection
Matérn Noise for Triangulation-Agnostic Flow Matching on Meshes
Cross-Paradigm Knowledge Distillation: A Comprehensive Study of Bidirectional Transfer Be…
FPED: A Functional-Network Prior-Guided Mixture-of-Experts Framework for Interpretable Br…
CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering
Fast-tracking genetic leads to reverse cellular aging
PIXLRelight: Controllable Relighting via Intrinsic Conditioning
General Preference Reinforcement Learning
EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL
SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents
Import AI 457: AI stuxnet; cursed Muon optimizer; and positive alignment
Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention
Finding the molecular switches behind new infectious diseases
Opening new paths in aging research
Accelerating discovery of liver disease mechanisms
Uniting biological toolkits for a new approach to ALS
Uncovering repurposed medicines to fight liver fibrosis
How LLMs Compute the Right Answer, Then Match the Swarm’s Wrong One, and How to Wire Arou…
Researchers Just Counted 146,932 Hallucinated Citations. This Repo Is the First Installab…
Unlocking asynchronicity in continuous batching
When AI agents learn to engineer themselves
Co-Scientist: A multi-agent AI partner to accelerate research
not much happened today
What Parameter Golf taught us about AI-assisted research
+29k Stars, No Vectors: How PageIndex Replaces Embeddings With LLM Reasoning
How HeavySkill Turns Agentic Harness Tricks Into a One-File Inner Skill
Adding Benchmaxxer Repellant to the Open ASR Leaderboard
How frontier firms are pulling ahead
🔬Doing Vibe Physics — Alex Lupsasca, OpenAI