Machine Learning & AI
Intermediate
SGD with Momentum
Accelerates gradient descent by accumulating a velocity of past gradients.
Formula
Variables
vVelocity
\betaMomentum (~0.9)
\alphaLearning rate
Example
beta=0.9 keeps 90% of prior velocity
Did You Know?
Momentum helps the optimizer roll through small local dips and ravines like a ball gaining speed.
Share this formula
More in Machine Learning & AI
View allLinear Regression Model
BasicPredicts a continuous value as a weighted sum of input features plus a bias.
Gradient Descent Update
BasicIteratively moves parameters in the direction that most reduces the loss.
ReLU Activation
BasicRectified Linear Unit: outputs the input if positive, else zero.
Leaky ReLU
BasicA ReLU variant that lets a small gradient flow for negative inputs.
Tanh Activation
BasicSquashes input to the range (-1, 1); zero-centred activation.
Softmax
BasicTurns a vector of scores into a probability distribution that sums to 1.