4 posts
Blog
Where a token actually spends its time
A new explainer that follows one token through a GPU cluster, and makes the case that prefill and decode are different problems sharing hardware.
A site you update with git push
The old abusing.technology was a CMS starter I never finished configuring. The new one has no CMS, no login and no server, just markdown in a repo.
Watching GPT-2 pick a word, one step at a time
An explainer that runs a real model in your browser and shows the eight steps between a prompt and a single token, including the one everybody gets wrong.
Bridging APRS and Meshtastic without losing the replies
Relaying packets one way is easy. Getting a reply back to the right device, for the right operator, is where the interesting problems live.