jsat.math.optimization.stochastic
Class AdaGrad
- java.lang.Object
-
- jsat.math.optimization.stochastic.AdaGrad
-
- All Implemented Interfaces:
- java.io.Serializable, GradientUpdater
public class AdaGrad extends java.lang.Object implements GradientUpdater
AdaGrad provides an adaptive learning rate for each individual feature
See: Duchi, J., Hazan, E.,&Singer, Y. (2011). Adaptive Subgradient Methods for Online Learning and Stochastic Optimization. Journal of Machine Learning Research, 12, 2121–2159.- See Also:
- Serialized Form
-
-
Constructor Summary
Constructors Constructor and Description AdaGrad()Creates a new AdaGrad updaterAdaGrad(AdaGrad toCopy)Copy constructor
-
Method Summary
All Methods Instance Methods Concrete Methods Modifier and Type Method and Description AdaGradclone()voidsetup(int d)Sets up this updater to update a weight vector of dimensiondby a gradient of the same dimensionvoidupdate(Vec x, Vec grad, double eta)Updates the weight vectorxsuch that x = x-ηf(grad), where f(grad) is some function on the gradient that effectively returns a new vector.doubleupdate(Vec x, Vec grad, double eta, double bias, double biasGrad)Updates the weight vectorxsuch that x = x-ηf(grad), where f(grad) is some function on the gradient that effectively returns a new vector.
-
-
-
Constructor Detail
-
AdaGrad
public AdaGrad()
Creates a new AdaGrad updater
-
AdaGrad
public AdaGrad(AdaGrad toCopy)
Copy constructor- Parameters:
toCopy- the object to copy
-
-
Method Detail
-
update
public void update(Vec x, Vec grad, double eta)
Description copied from interface:GradientUpdaterUpdates the weight vectorxsuch that x = x-ηf(grad), where f(grad) is some function on the gradient that effectively returns a new vector. It is not necessary for the internal implementation to ever explicitly form any of these objects, so long asxis mutated to have the correct result.- Specified by:
updatein interfaceGradientUpdater- Parameters:
x- the vector to mutate such that is has been updated by the gradientgrad- the gradient to update the weight vectorxfrometa- the learning rate to apply
-
update
public double update(Vec x, Vec grad, double eta, double bias, double biasGrad)
Description copied from interface:GradientUpdaterUpdates the weight vectorxsuch that x = x-ηf(grad), where f(grad) is some function on the gradient that effectively returns a new vector. It is not necessary for the internal implementation to ever explicitly form any of these objects, so long asxis mutated to have the correct result.
This version of the update method includes two extra parameters to make it easer to use when a scalar bias term is also used- Specified by:
updatein interfaceGradientUpdater- Parameters:
x- the vector to mutate such that is has been updated by the gradientgrad- the gradient to update the weight vectorxfrometa- the learning rate to applybias- the bias term of the vectorbiasGrad- the gradient for the bias term- Returns:
- the value to change the bias by, the update being
bias = bias - returnValue
-
clone
public AdaGrad clone()
- Specified by:
clonein interfaceGradientUpdater- Overrides:
clonein classjava.lang.Object
-
setup
public void setup(int d)
Description copied from interface:GradientUpdaterSets up this updater to update a weight vector of dimensiondby a gradient of the same dimension- Specified by:
setupin interfaceGradientUpdater- Parameters:
d- the dimension of the weight vector that will be updated
-
-
DataMelt 3.0 © DataMelt by jWork.ORG