• Home
  • Search
  • Dynamic learning rate decay for stochastic variational inference
  • https://doi.org/10.1088/2632-2153/ae19ccCopy DOI Icon

Dynamic learning rate decay for stochastic variational inference

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Abstract Like many optimization algorithms, stochastic variational inference is sensitive to the choice of the learning rate. If the learning rate is too small, the optimization process may be slow, and the algorithm might get stuck in local optima. On the other hand, if the learning rate is too large, the algorithm may oscillate or diverge, failing to converge to a solution. Adaptive learning rate methods such as Adam, AdaMax, Adagrad, or root mean square propagation automatically adjust the learning rate based on the history of gradients. Nevertheless, if the base learning rate is too large, the variational parameters might still oscillate around the optimal solution. With learning rate schedules, the learning rate can be reduced gradually to mitigate this problem. However, the amount at which the learning rate should be decreased in each iteration is not known a priori , which can significantly impact the performance of the optimization. In this work, we propose a method to decay the learning rate based on the history of the variational parameters. We use an empirical measure to quantify the amount of oscillations against the progress of the variational parameters to adapt the learning rate. The approach requires little memory and is computationally efficient. We demonstrate in various numerical examples that our method reduces the sensitivity of the optimization performance to the learning rate and that it can also be used in combination with other adaptive learning rate methods.

Similar Papers
  • Research Article
  • Citations2

Using adaptive learning rate to generate adversarial images

  • Sep 01, 2023
  • Current Directions in Biomedical Engineering
  • Ning Ding +1
  • Conference Article
  • Citations2

Shocking: an approach to stabilize backprop training with greedy adaptive learning rates

  • May 04, 1998
  • J.A Janet +5
  • Research Article
  • Citations2

Performance analysis of convolutional neural networks for image classification with appropriate optimizers

  • Jan 01, 2023
  • i-manager’s Journal on Mathematics
  • Danish Sana +2
  • Research Article
  • Citations17

DropCyclic: Snapshot Ensemble Convolutional Neural Network Based on a New Learning Rate Schedule for Land Use Classification

  • Jan 01, 2022
  • IEEE Access
  • Sangdaow Noppitak +1
  • Research Article
  • Citations16

An Optimized Back Propagation Learning Algorithm with Adaptive Learning Rate

  • Oct 30, 2017
  • International Journal on Advanced Science, Engineering and Information Technology
  • Nazri Mohd Nawi +5
  • Research Article
  • Citations65

Adaptive learning rate of SpikeProp based on weight convergence analysis

  • Dec 10, 2014
  • Neural Networks
  • Sumit Bam Shrestha +1
  • Research Article
  • Citations60

Autistic traits are related to worse performance in a volatile reward learning task despite adaptive learning rates.

  • Oct 08, 2020
  • Autism
  • Judith Goris +5
  • Book Chapter
  • Citations2

MaxVA: Fast Adaptation of Step Sizes by Maximizing Observed Variance of Gradients

  • Jan 01, 2021
  • Chen Zhu +5
  • PDF
  • Research Article
  • Citations48

Aerial Scene Classification through Fine-Tuning with Adaptive Learning Rates and Label Smoothing

  • Aug 21, 2020
  • Applied Sciences
  • Biserka Petrovska +5
  • Research Article
  • Citations119

A New Reinforcement Learning Based Learning Rate Scheduler for Convolutional Neural Network in Fault Classification

  • Dec 21, 2020
  • IEEE Transactions on Industrial Electronics
  • Long Wen +2
  • Research Article
  • Citations1

Research on fast self-learning improvement of ADRC control algorithm for film thickness control system

  • Apr 01, 2019
  • Journal of Physics: Conference Series
  • Liao Xue-Chao +2
  • Research Article
  • Citations8

Real-time adaptive impedance compensator using simultaneous perturbation stochastic approximation for enhanced physical human–robot interaction transparency

  • Oct 26, 2021
  • Robotics and Autonomous Systems
  • Kyeong Ha Lee +4
  • Book Chapter
  • Citations2

Comparison of Electricity Usage Forecasting Model Evaluation Based on Historical Load Dataset Duration Using Long Short-Term Memory Architecture

  • Jan 01, 2022
  • Nur Shakirah Md Salleh +2
  • Research Article

A Deep Dive into Training Algorithms for Deep Belief Networks

  • Feb 25, 2025
  • Journal of Information Systems Engineering and Management
  • S Manoha
  • Research Article
  • Citations4

Application neural network controller and active mass damper in structural vibration suppression

  • Jan 01, 2014
  • Journal of Intelligent & Fuzzy Systems
  • Chuen-Jyh Chen +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.