arXiv · 2402.12479
In value-based deep reinforcement learning, a pruned network is a good network
Abstract
Recent work has shown that deep reinforcement learning agents have difficulty in effectively using their network parameters. We leverage prior insights into the advantages of sparse training techniques and demonstrate that gradual magnitude pruning enables value-based agents to maximize parameter effectiveness. This results in networks that yield dramatic performance improvements over traditional networks, using only a small fraction of the full network parameters.
Explore related subjects
Keep this discovery
Johan Obando-Ceron, Aaron Courville, Pablo Samuel Castro. 2024-02-19. In value-based deep reinforcement learning, a pruned network is a good network. https://arxiv.org/abs/2402.12479
Cite the original work for its findings. Save a collection to share your selection of sources.