A Law of Data Separation in Deep Learning

October 31, 2022 ยท Declared Dead ยท ๐Ÿ› Proceedings of the National Academy of Sciences of the United States of America

๐Ÿ‘ป CAUSE OF DEATH: Ghosted
No code link whatsoever

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Hangfeng He, Weijie J. Su arXiv ID 2210.17020 Category cs.LG: Machine Learning Cross-listed cs.AI, cs.CV, cs.IT, stat.ML Citations 49 Venue Proceedings of the National Academy of Sciences of the United States of America Last Checked 5 months ago
Abstract
While deep learning has enabled significant advances in many areas of science, its black-box nature hinders architecture design for future artificial intelligence applications and interpretation for high-stakes decision makings. We addressed this issue by studying the fundamental question of how deep neural networks process data in the intermediate layers. Our finding is a simple and quantitative law that governs how deep neural networks separate data according to class membership throughout all layers for classification. This law shows that each layer improves data separation at a constant geometric rate, and its emergence is observed in a collection of network architectures and datasets during training. This law offers practical guidelines for designing architectures, improving model robustness and out-of-sample performance, as well as interpreting the predictions.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

๐Ÿ“œ Similar Papers

In the same crypt โ€” Machine Learning

Died the same way โ€” ๐Ÿ‘ป Ghosted