This paper explores whether supervising both agent actions and environment observations during reinforcement learning improves agent exploration and performance. Practitioners might care because it could lead to better initialization for reinforcement learning tasks.
Firehose
Filtered to Papers, tagged “environment modeling” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives