Introducing the OpenAI Economic Research Exchange
Explorar
Noticias de IA
22065 elementos — filtrados, clasificados y sin duplicados
Five labs, five minds: building a multi-model finance drama on small models
LLM Research Papers: The 2026 List (January to May)
Thousand Token Wood: shipping a multi-agent economy on a 3B model
How to Stop Shipping Low-Quality RL Environments (with Examples)
The University of Cambridge says it successfully tested a vaccine with an AI-designed ant…
Fine-tuning an LLM to write docs like it's 1995
The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Halluc…
Comprehensive and Reliable Feature Attribution for Diverse Modalities and Models via Freq…
Separation Power of Equivariant Neural Networks
Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss
Evaluating the Utility of Personal Health Records in Personalized Health AI
Inverse Entropic Optimal Transport Solves Semi-supervised Learning via Data Likelihood Ma…
HomeWorld: A Unified Floorplan-to-Furnished Framework for Generating Controllable, Densel…
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
OneReason Technical Report
Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent Reasoning
GIPO: Gaussian Importance Sampling Policy Optimization
Reducing Hallucinations in Complex Question Answering using Simple Graph-based Retrieval-…
RAT: RunAnyThing via Fully Automated Environment Configuration
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memoriz…
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
When Attention Beats Fourier: Multi-Scale Transformers for PDE Solving on Irregular Domai…
CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspecti…
Trust, but Don't Verify: Epistemic Blind Spots in LLM Source Evaluation
Channel-Wise Mixed-Precision Quantization for Large Language Models
Towards World Models in Biomedical Research
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
Semantic Partial Grounding via LLMs