Tagged “Neural-Networks”

  1. Extending makemore with an MLP, and running some experiments to minimize validation loss

  2. A neural character-level bigram language model, optimized with gradient descent rather than counted from the training data.

  3. Extending the autograd engine with more arithmetic operations and nonlinear activations, then building a small neural network library on top of it.

  4. Implementing a scalar-valued automatic differentiation engine from scratch. First post following Karpathy's Neural Networks: Zero to Hero.