Interpretable and pedagogical examples
Explorar
Noticias de IA
22065 elementos — filtrados, clasificados y sin duplicados
Learning a hierarchy
Generalizing from simulation
Sim-to-real transfer of robotic control with dynamics randomization
Asymmetric actor critic for image-based robot learning
Domain randomization and generative models for robotic grasping
Meta-learning for wrestling
Competitive self-play
Nonlinear computation in deep linear networks
Learning to model other minds
Learning with opponent-learning awareness
More on Dota 2
Dota 2
Better exploration with parameter noise
Hindsight Experience Replay
Teacher–student curriculum learning
Learning from human preferences
Learning to cooperate, compete, and communicate
UCB exploration via Q-ensembles
Robots that learn
Equivalence between policy gradients and soft Q-learning
Stochastic Neural Networks for hierarchical reinforcement learning
Unsupervised sentiment neuron
Spam detection in the physical world
Evolution strategies as a scalable alternative to reinforcement learning
One-shot imitation learning
Learning to communicate
Emergence of grounded compositional language in multi-agent populations
Prediction and control with temporal segment models
Third-person imitation learning