Parallel Trust-Region Approaches in Neural Network Training: Beyond Traditional Methods

December 21, 2023 Β· Declared Dead Β· πŸ› arXiv.org

πŸ‘» CAUSE OF DEATH: Ghosted
No code link whatsoever

"No code URL or promise found in abstract"

Evidence collected by the PWNC Scanner

Authors Ken Trotti, Samuel A. Cruz Alegría, Alena KopaničÑkovÑ, Rolf Krause arXiv ID 2312.13677 Category math.NA: Numerical Analysis Cross-listed cs.LG Citations 2 Venue arXiv.org Last Checked 2 months ago
Abstract
We propose to train neural networks (NNs) using a novel variant of the ``Additively Preconditioned Trust-region Strategy'' (APTS). The proposed method is based on a parallelizable additive domain decomposition approach applied to the neural network's parameters. Built upon the TR framework, the APTS method ensures global convergence towards a minimizer. Moreover, it eliminates the need for computationally expensive hyper-parameter tuning, as the TR algorithm automatically determines the step size in each iteration. We demonstrate the capabilities, strengths, and limitations of the proposed APTS training method by performing a series of numerical experiments. The presented numerical study includes a comparison with widely used training methods such as SGD, Adam, LBFGS, and the standard TR method.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

πŸ“œ Similar Papers

In the same crypt β€” Numerical Analysis

R.I.P. πŸ‘» Ghosted

Tensor Ring Decomposition

Qibin Zhao, Guoxu Zhou, ... (+3 more)

math.NA πŸ› arXiv πŸ“š 427 cites 9 years ago

Died the same way β€” πŸ‘» Ghosted