Gotta Learn Fast: A new benchmark for generalization in RL
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Retro Contest
Variance reduction for policy gradient with action-dependent factorized baselines
Improving GANs using optimal transport
On first-order meta-learning algorithms
Reptile: A scalable meta-learning algorithm
Some considerations on learning to explore via meta-reinforcement learning
Ingredients for robotics research
Multi-Goal Reinforcement Learning: Challenging robotics environments and request for rese…
Interpretable machine learning through teaching
Discovering types for entity disambiguation
Requests for Research 2.0
Learning sparse neural networks through L₀ regularization
Interpretable and pedagogical examples
Learning a hierarchy
Generalizing from simulation
Asymmetric actor critic for image-based robot learning
Sim-to-real transfer of robotic control with dynamics randomization
Domain randomization and generative models for robotic grasping
Competitive self-play
Meta-learning for wrestling
Nonlinear computation in deep linear networks
Learning to model other minds
Learning with opponent-learning awareness
More on Dota 2
Dota 2
Better exploration with parameter noise
Hindsight Experience Replay
Teacher–student curriculum learning
Learning from human preferences