Hindsight Experience Replay
Marcin Andrychowicz(University of Warsaw), Wojciech Zaremba(OpenAI (United States)), Alex Ray(OpenAI (United States)), Pieter Abbeel(OpenAI (United States)), Rachel Fong, Filip Wolski, Josh Tobin(OpenAI (United States)), Peter Welinder(California Institute of Technology), Jonas Schneider(OpenAI (United States)), Bob McGrew(OpenAI (United States))
Cited by 350
Related Papers
Training language models to follow instructions with human feedback
|arXiv (Cornell University)|2022|4.3k
Learning dexterous in-hand manipulation
|The International Journal of Robotics Research|2019|1.6k
Evaluating Large Language Models Trained on Code
|arXiv (Cornell University)|2021|1.5k
Training Language Models to Follow Instructions with Human Feedback
|Unknown|2022|760
The Multidimensional Wisdom of Crowds
|CaltechAUTHORS (California Institute of Technology)|2010|735