Multi-Goal Reinforcement Learning: Challenging Robotics Environments and\n Request for Research
Matthias Plappert(OpenAI (United States)), Wojciech Zaremba(OpenAI (United States)), Bob McGrew(OpenAI (United States)), Bowen Baker(OpenAI (United States)), Joshua W.D. Tobin(Mater Health Services), Alex Ray, Marcin Andrychowicz(University of Warsaw), Vikash Kumar, Glenn Powell, Peter Welinder(California Institute of Technology), Jonas Schneider(OpenAI (United States)), Maciek Chociej(OpenAI (United States))
Cited by 165
Related Papers
Training language models to follow instructions with human feedback
|arXiv (Cornell University)|2022|4.3k
Learning dexterous in-hand manipulation
|The International Journal of Robotics Research|2019|1.6k
Evaluating Large Language Models Trained on Code
|arXiv (Cornell University)|2021|1.5k
Training Language Models to Follow Instructions with Human Feedback
|Unknown|2022|760
The Multidimensional Wisdom of Crowds
|CaltechAUTHORS (California Institute of Technology)|2010|735