Machine Learning & AI
Advanced
KL Divergence
Measures how one probability distribution differs from a reference distribution.
Formula
Variables
PTrue distribution
QApproximating distribution
Example
Zero only when P equals Q
Did You Know?
KL divergence is asymmetric — the distance from P to Q is not the same as Q to P.
Share this formula
More in Machine Learning & AI
View allLinear Regression Model
BasicPredicts a continuous value as a weighted sum of input features plus a bias.
Gradient Descent Update
BasicIteratively moves parameters in the direction that most reduces the loss.
ReLU Activation
BasicRectified Linear Unit: outputs the input if positive, else zero.
Leaky ReLU
BasicA ReLU variant that lets a small gradient flow for negative inputs.
Tanh Activation
BasicSquashes input to the range (-1, 1); zero-centred activation.
Softmax
BasicTurns a vector of scores into a probability distribution that sums to 1.