> ML_LITERATURE // ANDRYCHOWICZ-2017-HINDSIGHT-EXPERIENCE-REPLAY-HER_v1.0
Hindsight Experience Replay (HER)
Marcin Andrychowicz, Filip Wolski, Alex Ray, Jonas Schneider, Rachel Fong, Peter Welinder, Bob McGrew, Josh Tobin, Pieter Abbeel, Wojciech Zaremba · Advances in Neural Information Processing Systems (NeurIPS) (2017)
algorithm2017foundationalthirdPartyReproduced
Principal Contribution
Re-examined failed trajectories in multi-goal RL by relabeling the achieved final state as the intended target goal, learning successfully from sparse binary failure rewards.
Operational Relevance
Serves as qualified theoretical and systems foundation for task-reinforcement-learning, task-robotics.
Assumptions
- Markovian state dynamics and stationary reward functions hold in target evaluation environments
Limitations
- Sample efficiency, exploration stability, and real-world sim-to-real transfer gaps require specialized tuning
Connected Algorithms, Architectures & Tools
Related Algorithms:
Related Architectures:
Implementing Libraries:
