Skip to main content

> ML_LITERATURE // ANDRYCHOWICZ-2017-HINDSIGHT-EXPERIENCE-REPLAY-HER_v1.0

Hindsight Experience Replay (HER)

Marcin Andrychowicz, Filip Wolski, Alex Ray, Jonas Schneider, Rachel Fong, Peter Welinder, Bob McGrew, Josh Tobin, Pieter Abbeel, Wojciech Zaremba · Advances in Neural Information Processing Systems (NeurIPS) (2017)

algorithm2017foundationalthirdPartyReproduced

Principal Contribution

Re-examined failed trajectories in multi-goal RL by relabeling the achieved final state as the intended target goal, learning successfully from sparse binary failure rewards.

Operational Relevance

Serves as qualified theoretical and systems foundation for task-reinforcement-learning, task-robotics.

Assumptions

  • Markovian state dynamics and stationary reward functions hold in target evaluation environments

Limitations

  • Sample efficiency, exploration stability, and real-world sim-to-real transfer gaps require specialized tuning

Connected Algorithms, Architectures & Tools

Related Algorithms:
Related Architectures:
Implementing Libraries: