Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

S. Fathi Hafshejani

New logarithmic step size for stochastic gradient descent

Apr 01, 2024

M. Soheil Shamaee, S. Fathi Hafshejani, Z. Saeidian

Abstract:In this paper, we propose a novel warm restart technique using a new logarithmic step size for the stochastic gradient descent (SGD) approach. For smooth and non-convex functions, we establish an $O(\frac{1}{\sqrt{T}})$ convergence rate for the SGD. We conduct a comprehensive implementation to demonstrate the efficiency of the newly proposed step size on the ~FashionMinst,~ CIFAR10, and CIFAR100 datasets. Moreover, we compare our results with nine other existing approaches and demonstrate that the new logarithmic step size improves test accuracy by $0.9\%$ for the CIFAR100 dataset when we utilize a convolutional neural network (CNN) model.

* Frontiers of Computer Science, 2025

Via

Access Paper or Ask Questions

Modified Step Size for Enhanced Stochastic Gradient Descent: Convergence and Experiments

Sep 03, 2023

M. Soheil Shamaee, S. Fathi Hafshejani

Abstract:This paper introduces a novel approach to enhance the performance of the stochastic gradient descent (SGD) algorithm by incorporating a modified decay step size based on $\frac{1}{\sqrt{t}}$. The proposed step size integrates a logarithmic term, leading to the selection of smaller values in the final iterations. Our analysis establishes a convergence rate of $O(\frac{\ln T}{\sqrt{T}})$ for smooth non-convex functions without the Polyak-{\L}ojasiewicz condition. To evaluate the effectiveness of our approach, we conducted numerical experiments on image classification tasks using the FashionMNIST, and CIFAR10 datasets, and the results demonstrate significant improvements in accuracy, with enhancements of $0.5\%$ and $1.4\%$ observed, respectively, compared to the traditional $\frac{1}{\sqrt{t}}$ step size. The source code can be found at \\\url{https://github.com/Shamaeem/LNSQRTStepSize}.

* Mathematics Interdisciplinary Research 2023

Via

Access Paper or Ask Questions

Binary Orthogonal Non-negative Matrix Factorization

Oct 19, 2022

S. Fathi Hafshejani, D. Gaur, S. Hossain, R. Benkoczi

Figure 1 for Binary Orthogonal Non-negative Matrix Factorization

Figure 2 for Binary Orthogonal Non-negative Matrix Factorization

Abstract:We propose a method for computing binary orthogonal non-negative matrix factorization (BONMF) for clustering and classification. The method is tested on several representative real-world data sets. The numerical results confirm that the method has improved accuracy compared to the related techniques. The proposed method is fast for training and classification and space efficient.

Via

Access Paper or Ask Questions