JA EN

#gradient-descent

1 articles

01 ·Machine Learning Basics·FREE·8 min read Loss Functions and Optimization — How a Model Learns From Being Wrong Why MSE and cross-entropy have the shapes they do, what the gradient actually points at, and one step of gradient descent taken apart with equations, a draggable figure, and ten lines of numpy — divergence included.