arXiv · 1909.13360
Library network, a possible path to explainable neural networks
Abstract
Deep neural networks (DNNs) may outperform human brains in complex tasks, but the lack of transparency in their decision-making processes makes us question whether we could fully trust DNNs with high stakes problems. As DNNs' operations rely on a massive number of both parallel and sequential linear/nonlinear computations, predicting their mistakes is nearly impossible. Also, a line of studies suggests that DNNs can be easily deceived by adversarial attacks, indicating that their decisions can easily be corrupted by unexpected factors. Such vulnerability must be overcome if we intend to take advantage of DNNs' efficiency in high stakes problems. Here, we propose an algorithm that can help us better understand DNNs' decision-making processes. Our empirical evaluations suggest that this algorithm can effectively trace DNNs' decision processes from one layer to another and detect adversarial attacks.
Explore related subjects
Keep this discovery
Jung Hoon Lee. 2019-09-29. Library network, a possible path to explainable neural networks. https://arxiv.org/abs/1909.13360
Cite the original work for its findings. Save a collection to share your selection of sources.