Blog 385 posts

Notes from the trenches on AI engineering, LLM apps, and the full-stack work that holds it all together.

Microsoft FastContext cuts coding agent tokens by offloading repo search
5 min read

Microsoft FastContext cuts coding agent tokens by…

FastContext is a 4B–30B exploration subagent that returns file-line citations instead of dumping whole files into the main agent. Mini-SWE-Agent gains up to 5.5% success with up to 60% fewer main-agent tokens.

Kimi K2.7 Code thinks 30% less and still ships harder on long coding tasks
5 min read

Kimi K2.7 Code thinks 30% less and…

Moonshot's open-weight Kimi K2.7 Code keeps the 1T MoE backbone but cuts thinking tokens ~30% versus K2.6 while jumping +21.8% on Kimi Code Bench v2. Here is when I would route agents to it.

Anthropic filed for IPO at $965B, OpenAI declared chat dead, and Gary Marcus called it a bubble. Something's gotta give.
9 min read

Anthropic filed for IPO at $965B, OpenAI…

A $965B confidential S-1, OpenAI's 'chat is dead' pivot, and a half-trillion-dollar chip rout, all in one week. One capital cycle, one market that can't decide what to believe.

What Claude Fable 5's leaked system prompt actually reveals about Mythos
5 min read

What Claude Fable 5's leaked system prompt…

A near-complete Claude Fable 5 product prompt surfaced on GitHub in June 2026. The Mythos tier, artifact storage API, and model-switch rules are the parts that matter for builders.