Aug 19, 2026
Latest PostHow Kubernetes probes work
Learn how probes work interactively on simulated Kubernetes clusters running in your browser.
Jun 30, 2026
I ported Kubernetes to the browser
Almost 100,000 lines of LLM-generated code in 2 months, and none of it is slop.
Mar 25, 2026
Quantization from the ground up
A complete guide to what quantization is, how it works, and how it's used to compress large language models
Jan 29, 2026
What those AI benchmark numbers mean
A guide to the 14 benchmarks you'll see cited in every new model release: what they test, how they're built, and why their scores are easy to misread.
Dec 16, 2025
Prompt caching: 10x cheaper LLM tokens, but how?
A far more detailed explanation of prompt caching than anyone asked for: how tokens, embeddings, and attention make cached LLM tokens 10x cheaper and faster.
Nov 17, 2025
Migrate from ingress-nginx to the ngrok Operator
A step-by-step guide to migrating from ingress-nginx to the ngrok Operator without downtime.
Oct 27, 2025
Self-hosting with and without ngrok
A step-by-step guide to self-hosting your web app using a VPS, Caddy, and systemd, with tips for using ngrok as an alternative.
•~5.2K words
