20 Jul 2026 • 07 Mins read
The Model Is the Brain. The Harness Is the Body.
What I learned teaching vibe coding through context limits, phase prompts, validation loops, and the strange vector geometry inside language models.
Continue Reading20 Jul 2026 • 07 Mins read
What I learned teaching vibe coding through context limits, phase prompts, validation loops, and the strange vector geometry inside language models.
Continue Reading10 May 2026 • 15 Mins read
I had a hunch the standard system → user prompt order was wrong for negative constraints. Every standard answer told me to drop it. I ran 1500 calls on 30 hard tasks and the flipped version beat the standard one by 14 to 46 points. The system slot is the worst place for your 'do not' rules and the docs back it.
Continue Reading14 Apr 2026 • 08 Mins read
A friend asked what MCP actually gives you over plain APIs. I tried to defend it. Every defense line collapsed under one simple question. Here's the honest walk through why MCP is a packaging standard, not a technical innovation, and why that's still valuable.
Continue Reading12 Apr 2026 • 07 Mins read
Two researchers proved mathematically that rational firms will over-automate with AI even when they know it destroys the demand they depend on. Only one policy instrument out of six actually fixes it.
Continue Reading09 Apr 2026 • 13 Mins read
I read Anthropic's 243-page Claude Mythos system card. It starts as safety science and ends as 20 pages of employees marveling at their model's creative writing. The real story is buried in section 5.8.1, where they admit the circularity.
Continue Reading05 Apr 2026 • 05 Mins read
Karpathy shipped autoresearch and named a pattern everyone could already run. The real insight is that the loop works anywhere you can define a verifiable output. You can apply it today, in any field, without waiting for a framework.
Continue Reading04 Apr 2026 • 14 Mins read
Gemma 4 dragged Per-Layer Embeddings back onto the timeline, so I rebuilt PLE from scratch, hit the dead ends honestly, and measured what actually worked.
Continue Reading03 Apr 2026 • 12 Mins read
A Russian mathematician picked a fight over free will, accidentally invented the most important idea in probability, and 120 years later it powers Google, nuclear physics, and every large language model you've ever used. Here's how Markov chains actually work.
Continue Reading01 Apr 2026 • 08 Mins read
The Claude Code source map story turned into a public dunk fest, but the real lesson is about build hygiene, trust boundaries, and how exposed modern AI tools can get.
Continue Reading28 Mar 2026 • 07 Mins read
A practical breakdown on what actually goes wrong with AI coding agents in production, from instruction overload and horizontal planning to why humans still need to do the real thinking. Written from experience shipping real systems at scale.
Continue Reading14 Feb 2026 • 11 Mins read
A rebuttal to AI replacement hype from someone who builds production AI systems at scale, explaining why human judgment, architectural thinking, and oversight matter more than ever, and why trusting AI to replace you is the actual danger.
Continue Reading