Evolving Curricula with Regret-Based Environment Design(accelagent.github.io)
accelagent.github.io
Evolving Curricula with Regret-Based Environment Design
https://accelagent.github.io/
1 comments
Thanks! The demo just shows the final agents after training (30K gradient updates). Interesting work re the reward maximizing curricula. I have not seen this before, so thanks for the pointer.
(ps: I co-wrote a short elementary paper on auto curricula design for RL in 2017 [0])
[0]: https://arxiv.org/abs/1703.07853