This paper develops a new method to improve the stability of asynchronous reinforcement learning by adapting the trust region to account for staleness, which is a common problem in this field. Practitioners might care about this because stable reinforcement learning can lead to better performance and more efficient training.
Firehose
Filtered to Papers, tagged “staleness” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News