Becoming the Next Anthropic With a 2020 GPU

Why This Project In the previous post, I ended up running inference on a Tesla T10 picked up for €300 — a card that, in another life, streamed cloud gaming for GeForce Now. Fun, but let’s be honest for a second: running a model someone else trained is the bare minimum. Anyone with ollama pull can do that in thirty seconds. What was itching at me was the part before that. Where the model actually comes from. How you go from a pile of raw text to something that strings together sentences that hold up. When you use Claude, ChatGPT, or Mistral day to day, you type a prompt, out comes an answer, and everything in between is black magic to pretty much everyone who uses it. ...

August 13, 2026 · 21 min · Léo Nonnenmacher