3 ms·
Heavily optimized single C file that can train the same model as Karpathy's microgpt to lower loss in under a second on a single Mac core.
by easygenes 7mo ago
Heavily optimized single C file that can train the same model as Karpathy's microgpt to lower loss in under a second on a single Mac core.