Glossary › Math
Function that converts a vector of real numbers into a probability distribution: softmax(x_i) = e^(x_i) / Σe^(x_j). Used in attention scores and classification outputs. Temperature scales logits before softmax.
Temperature · Attention Mechanism
Cosine Similarity · Cross-Entropy Loss · Perplexity
DeVenture Academy teaches AI engineering through projects — every lesson is paired with a hands-on lab you run on your own machine with real tools.
Create a free account — the opening phases of 24 of 30 courses are free, no credit card. Or see Pro pricing.
All courses · Pricing · About · FAQ · Glossary