Karl
@karlkfi
SF Tech Gamer Car Nerd. AI Infra Tech Lead @ AMD (opinions my own). Previously Google, Cruise, Mesosphere, Pivotal.
Welp… guess I’m not gonna get that 700 day Reddit streak cheevo. Counter reset because I worked all day then passed out from not sleeping the previous night. Maybe this is a sign.
I posted the cache chart in a previous thread in the quote chain, but stopped posting it because it's just silly. The tokens reported above are non-cache. Here's the latest one with reads & writes.
Surprisingly, Opus 5 runs 21% cheaper per message than 4.8. But a beefy new toy and a growing addiction meant a lot more tokens spent. 96 hours at the keyboard became 286 hours of session time in 9 days. ~3 at once, peak 13. Parallel dispatch is the trick. One task, one session, one branch.
Day 80 of building a Kubernetes operator with Claude Code: 1,654 commits 1,876 tests 424M tokens 198k lines authored ~2,140 tokens/line, flat since day 70 Daily token spend jumped 2.9× after I got my new M5 Max 128GB MacBook Pro!
The headline number undersells it. Counting cache reads: 10.5B tokens. 227M written to cache, each replayed ~45 times. The cache is the only reason this is affordable.
Day 70 of building a Kubernetes operator with Claude Code: 1,221 commits 1,394 tests 284M tokens 137k lines authored ~2,050 tokens/line — barely rising now Scaffolding was cheap. Logic, tests, and review are where the tokens went. Now the conventions are paying rent.
Day 48 of building a Kubernetes operator with Claude Code: 965 commits 976 tests 188M tokens 98k lines authored ~1,900 tokens/line and rising The cost-per-line climbs ~5× as the work shifts from scaffolding to logic, tests, and debugging.
I had way too much fun marking this! "AI generated", but frame-by-frame with code.
35 days building github-actions-gateway with Claude Code: 📈 115M tokens spent 📦 60K lines authored 🔁 786 commits 🐹 28k lines of Go 📚 20k lines of Documentation 🧪 547 test cases Each new line cost more tokens! Most of the work is reading, writing, testing & reviewing, not cranking out code.
Claude's cache doing heavy lifting here. 1.89B cache reads. 42M cache writes. 22 days. Every token I cached got replayed ~45 times!
Upgrading to Max made using Opus for everything viable. Every gold bar before the line is Sonnet 4.6 on Pro. The second Max kicked in, I switched to Opus — and rode the 4.7 → 4.8 upgrade straight through. Daily token use jumped with it.
22 days into this side project with Claude Code... Day 7 → Day 22: 10M → ~56M tokens 232 → 617 commits 15.5k → 20.9k lines of Go 2.3k → 4.2k lines of comments 14.3k → 18.3k lines of Markdown 1.5k → 1.8k lines of YAML (without CRDs) 269 → 393 tests Mostly Sonnet → Mostly Opus Pro → Max
Helios racks are coming. - CPUs: 18x EPYC “Venice” (2nm, 256 physical cores each) - GPUs: 72x MI455X (432 GB HBM4 each) - NICs: 36x Pensando Vulcano (800 Gbps each) - RAM: 288x DDR5 (1.6 TB/s per-socket) Venice production is ramping up now. www.tomshardware.com/tech-industr...
I've joined AMD! The future for me looks a lot like the past: building platforms for internal developers. Kubernetes, containers, virtual machines, bare metal, hybrid, multi-cloud... anything AMD needs to build, test, and ship GPU drivers, AI tools, inference, training, demo clusters, you name it!
Frieren is great. What if the elf that went on the hero’s quest and helped defeat the Demon King outlived the rest of her party? What would her next adventure be? On Crunchyroll and Netflix.
GKE just landed some nifty features: - HPA decision logging - Node & Pod startup latency monitoring cloud.google.com/kubernetes-e...