Agent Skills: nanogpt
Educational GPT implementation in ~300 lines. Reproduces GPT-2 (124M) on OpenWebText. Clean, hackable code for learning transformers. By Andrej Karpathy. Perfect for understanding GPT architecture from scratch. Train on Shakespeare (CPU) or OpenWebText (multi-GPU).
UncategorizedID: benchflow-ai/skillsbench/nanogpt
278174
Install this agent skill to your local
Skill Files
Browse the full folder contents for nanogpt.
Loading file tree…
Select a file to preview its contents.