arrow
返回

Mini-batch algorithms with online step size

delete2019-02-01
delete25
PRE
AI
杨壮 封面图
杨壮 (Zhuang Yang)
C
Cheng Wang
Z
Zhemin Zhang *
J
Jonathan Li
DOI:10.1016/j.knosys.2018.11.031delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Mini-batch algorithms have been proposed as a way to speed-up stochastic optimization methods and good results for mini-batch algorithms have been reported previously. A major issue with mini-batch algorithms is how to timely and readily acquire step size while running the algorithm. Usually, mini-batch algorithms employ a diminishing step size, or a best-tuned step size by mentor, which, in practice, are time consuming. To solve this problem, we propose using a hypergradient to compute an online step size (OSS) for mini-batch algorithms. Specifically, we incorporate online step size into advanced mini-batch algorithms, mini-batch nonconvex stochastic variance reduced gradient (MSVRG), thereby generating a new method, MSVRG-OSS. When computing step size in MSVRG-OSS, mini-batch samples are used. In addition, MSVRG-OSS, which needs little additional computation, requires only one extra copy of the original gradient to be stored in memory. We prove that MSVRG-OSS converges linearly in expectation and analyze its complexity. We present numerical results on problems arising with machine learning that indicate the proposed method shows great promise. We also show that, with slightly large batch samples, MSVRG-OSS is insensitive to the initial parameters, which are the key factor for controlling the performance of the algorithm. (C) 2018 Elsevier B.V. All rights reserved.
Keyword:
Stochastic optimization
Convex optimization
Mini-batch of samples
Online step size
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

K
Knowledge-Based Systems
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

X
xiamen university
学者数:
5.9W
论文数: 3.8W
被引数: 67
引用论文

引用论文

Neutron diffraction investigation of the evolution of the crystal structure of oxygen-conducting solid solutions (Yb1 − x Ca x )2Ti2O7 (x = 0, 0.05, 0.10)
err2009-02-08
err0
PREAI
errA. V. Shlyakhtina; A. E. Sokolov; V. A. Ul’yanov; V. A. Trunov; M. V. Boguslavskiĭ; A. V. Levchenko; L. G. Shcherbakova
err分享
err收藏
err分享
err收藏
Explicit and implicit reinforcement learning across the psychosis spectrum.
err2017-07-01
err0
errOAAI
errDeanna M. Barch; Cameron S. Carter; James M. Gold; Sheri L. Johnson; Ann M. Kring; Angus W. MacDonald; Diego A. Pizzagalli; J. Daniel Ragland; Steven M. Silverstein; Milton E. Strauss
err分享
err收藏
Reinforcing effects of carbon nanotubes in structural aluminum matrix nanocomposites
err2011-01-31
err0
PREAI
errHyunjoo Choi; Jaehyuck Shin; Byungho Min; Junsik Park; Donghyun Bae
err分享
err收藏
学者 查看更多内容