arrow
返回

Efficient gradient computation for optimization of hyperparameters

delete2022-02-16
delete1
PRE
AI
J
Jingyan Xu *
F
Frédéric Noo
DOI:10.1088/1361-6560/ac4442delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
We are interested in learning the hyperparameters in a convex objective function in a supervised setting. The complex relationship between the input data to the convex problem and the desirable hyperparameters can be modeled by a neural network; the hyperparameters and the data then drive the convex minimization problem, whose solution is then compared to training labels. In our previous work (Xu and Noo 2021 Phys. Med. Biol. 66 19NT01), we evaluated a prototype of this learning strategy in an optimization-based sinogram smoothing plus FBP reconstruction framework. A question arising in this setting is how to efficiently compute (backpropagate) the gradient from the solution of the optimization problem, to the hyperparameters to enable end-to-end training. In this work, we first develop general formulas for gradient backpropagation for a subset of convex problems, namely the proximal mapping. To illustrate the value of the general formulas and to demonstrate how to use them, we consider the specific instance of 1D quadratic smoothing (denoising) whose solution admits a dynamic programming (DP) algorithm. The general formulas lead to another DP algorithm for exact computation of the gradient of the hyperparameters. Our numerical studies demonstrate a 55%-65% computation time savings by providing a custom gradient instead of relying on automatic differentiation in deep learning libraries. While our discussion focuses on 1D quadratic smoothing, our initial results (not presented) support the statement that the general formulas and the computational strategy apply equally well to TV or Huber smoothing problems on simple graphs whose solutions can be computed exactly via DP.
Keyword:
dynamic programming
gradient backpropagation
hyperparameter learning
proximal mapping
automatic differentiation
implicit differentiation

期刊

Physics in Medicine and Biology 封面图
Physics in Medicine and Biology
IF:
3.4
论文数:
1.4W
被引数:
3.1W

机构

J
Johns Hopkins University
学者数:
10.2W
论文数: 8.8W
被引数: 13.0W
U
Utah System of Higher Education
学者数:
4.6W
论文数: 4.0W
被引数: 161
引用论文

引用论文

Browplasty as an adjunct to rhinoplasty
err2006-07-18
err0
PREAI
errRichard C. Webster; Terence M. Davidson; Richard C. Smith
err分享
err收藏
err分享
err收藏
Learning Convex Optimization Models学习凸优化模型
err2021-08-01
err29
errOAAI
errAgrawal, Akshay; Barratt, Shane; Boyd, Stephen
err分享
err收藏
An overview of bilevel optimization
err2007-04-20
err1.2K
PREAI
errColson, Benoit; Marcotte, Patrice; Savard, Gilles
err分享
err收藏
学者 查看更多内容