Hindsight Experience Replay

Explore this paper's citation graph

Summary

A novel technique is presented which allows sample-efficient learning from rewards which are sparse and binary and therefore avoid the need for complicated reward engineering and may be seen as a form of implicit curriculum.

Type
preprint
Published
2017-07-05
Cited by
2,855
References
47
Access
Open access

Keywords

Hindsight bias, Task (project management), Computer science, Reinforcement learning, Artificial intelligence

References

Cited by

Related papers