OpenAI Fellows Summer 2018: Final projects
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
How AI training scales
Quantifying generalization in reinforcement learning
Spinning Up in Deep RL
Learning concepts with energy functions
Plan online, learn offline: Efficient learning and exploration via model-based control
Reinforcement learning with prediction-based rewards
Learning complex goals with iterated amplification
FFJORD: Free-form continuous dynamics for scalable reversible generative models
The International 2018: Results
Large-scale study of curiosity-driven learning
OpenAI Five Benchmark: Results
Learning dexterity
Variational option discovery algorithms
OpenAI Five Benchmark
Glow: Better reversible generative models
Learning Montezuma’s Revenge from a single demonstration
OpenAI Five
Retro Contest: Results
Learning policy representations in multiagent systems
Improving language understanding with unsupervised learning
GamePad: A learning environment for theorem proving
AI and compute
AI safety via debate
Evolved Policy Gradients
Gotta Learn Fast: A new benchmark for generalization in RL
Retro Contest
Variance reduction for policy gradient with action-dependent factorized baselines
Improving GANs using optimal transport
On first-order meta-learning algorithms