Stage 3 · Expert · A2
Training Your Own Small Model
A model of 449 numbers, trained in under a second, can be all your one task needs.
12 lessons · 124 minDemanding
About this chapter
Frontier models are general, and your problem is not. That gap is your advantage. This chapter takes one narrow question, starts with honest label work, and trains a model small enough to read in a text file. You will overfit on purpose to see the shape of it, fix it three ways, borrow a model trained for something else, let a big model teach a small one, and squeeze the weights down to a few bits. It ends by comparing your model with a giant the way an honest person would, cost included.
What you will be able to do
- 19 min
A Task Worth a Small Model
Tell whether a task is narrow enough for a model you can train yourself.
- 210 min
A Few Hundred Labels
Write a label guide, measure disagreement, and know the ceiling it puts on your score.
- 39 min
The Number You Have to Beat
Ship three untrained answers first and write down the one score a model has to clear.
- 412 min
Training a Tiny Classifier
Run a full training loop on your own labels and read both scores it produces.
- 510 min
Overfit on Purpose
Make a model memorize a tiny pile on purpose and read the three things it tells you.
- 611 min
Three Ways to Fix It
Pick between more labels, weight decay and stopping early by reading what each one buys.
- 712 min
A Language Model You Can Read
Train a character-level language model on this device and tell what it learned from what it memorized.
- 810 min
Fine-Tune or Start From Scratch
Decide between borrowing a trained model and starting fresh, from the number of labels you have.
- 910 min
LoRA, by Picture
Adapt a frozen model by training two skinny matrices instead of a whole layer.
- 1010 min
Letting the Giant Teach
Train a tiny model on a bigger model's beliefs instead of on human labels.
- 1110 min
Shrinking It to Ship
Round a model's weights onto a coarse ladder and find the point where shrinking starts to cost score.
- 1211 min
Judging It Against the Giant
Compare two models on the same held-back set honestly, and put the cost of an answer beside the score.