Glossary › Training

Gradient Descent

Optimization algorithm that iteratively updates parameters in the direction opposite to the gradient of the loss: θ = θ - lr × ∂L/∂θ. Variants: SGD, Adam, AdamW.

Where this is taught

Related terms

Learning Rate · Backpropagation

More in Training

Backpropagation · Batch Size · Fine-Tuning · LoRA (Low-Rank Adaptation) · RLHF · DPO (Direct Preference Optimization) · QLoRA · Learning Rate

Learn this by building

DeVenture Academy teaches AI engineering through projects — every lesson is paired with a hands-on lab you run on your own machine with real tools.

Create a free account — the opening phases of 24 of 30 courses are free, no credit card. Or see Pro pricing.

All courses · Pricing · About · FAQ · Glossary