Multi-Goal Reinforcement Learning: Challenging Robotics Environments and Request for Research
Matthias Plappert(OpenAI (United States)), Wojciech Zaremba(OpenAI (United States)), Maciek Chociej(OpenAI (United States)), Bob McGrew(OpenAI (United States)), Alex Ray(OpenAI (United States)), Bowen Baker(OpenAI (United States)), Josh Tobin(OpenAI (United States)), Marcin Andrychowicz(University of Warsaw), Vikash Kumar(Google (United States)), Glenn Powell(OpenAI (United States)), Jonas Schneider(OpenAI (United States)), Peter Welinder(California Institute of Technology)
Cited by 193
Related Papers
Training language models to follow instructions with human feedback
|arXiv (Cornell University)|2022|4.3k
Learning dexterous in-hand manipulation
|The International Journal of Robotics Research|2019|1.6k
Evaluating Large Language Models Trained on Code
|arXiv (Cornell University)|2021|1.5k
Training Language Models to Follow Instructions with Human Feedback
|Unknown|2022|760
The Multidimensional Wisdom of Crowds
|CaltechAUTHORS (California Institute of Technology)|2010|735