mirror of
https://github.com/gryf/coach.git
synced 2026-04-23 17:11:29 +02:00
RL in Large Discrete Action Spaces - Wolpertinger Agent (#394)
* Currently this is specific to the case of discretizing a continuous action space. Can easily be adapted to other case by feeding the kNN otherwise, and removing the usage of a discretizing output action filter
This commit is contained in:
Binary file not shown.
Reference in New Issue
Block a user