Show HN: Microgpt is a GPT you can visualize in the browser (microgpt.boratto.ca)
very much inspired by karpathy's microgpt of the same name. it's (by default) a 4000 param GPT/LLM/NN that learns to generate names. this is sorta an educational tool in that you can visualize the activations as they pass through the network, and click on things to get an explanation of them.
17 comments
[ 5.1 ms ] story [ 38.8 ms ] thread"Microgpt's larger cousins using building blocks called tokens representing one or more letters. That's hard to reason about, but essential for building sentences and conversations.
"So we'll just deal with spelling names using the English alphabet. That gives us 26 tokens, one for each letter."
Can't find it now though (maybe the link rotted?), anyone happen to know what that was?
To give a sense of what the loss value means, maybe you can add a small explainer section as a question and add this explanation from Karpathy’s blog:
> Over 1,000 steps the loss decreases from around 3.3 (random guessing among 27 tokens: −log(1/27)≈3.3) down to around 2.37.
to reiterate that the model is being trained to predict the next token out of 27 possible tokens and is now doing better than the baseline of random guess.
https://karpathy.github.io/2026/02/12/microgpt/
It was submitted to hn a few days ago but only received a few comments. https://news.ycombinator.com/item?id=47000263