Skip to content

Recent entries

Cut the Rules, Delegate the Judgment: Anthropic's Six Shifts for "Context Design in the Claude 5 Era"—and Why Accuracy Held Even After an 80% Cut

Anthropic's new rules for "context design in the Claude 5 era," published July 24, 2026. Cutting 80% of Claude Code's system prompt caused no drop in accuracy. This piece lays out six shifts—lightening CLAUDE.md and moving procedures into skills—along with the concerns around automatic memory and de

Loosening a Freshly Tightened Rein in Just Two Days — Claude Code v2.1.219 Restores Nested Subagents to "Depth 3 by Default" and Adds a "Fewer Than 15" Guideline for Workflows

Claude Code v2.1.219 restores nested subagents — turned off by default just two days earlier — to "depth 3 by default," and introduces a new "fewer than 15" guideline for dynamic workflows. A look at the tug-of-war between autonomy and safety, and where these changes matter most for unattended opera

Reviews Move Out of the Conversation, Deep Dives Wait to Be Called — How Claude Code v2.1.218 Draws the Line at "No Running Off on Its Own"

Claude Code v2.1.218 looks like a bundle of minor fixes, but a single thread runs through it. /code-review now runs in the background, /deep-research launches only when invoked manually, and fork skills default to running behind the scenes—pushing heavy automated work out of the conversation so noth

Leaving the 'Pro' Seat Empty, Three Cheap-and-Fast Flash Models Instead — How Gemini's Refresh Targets the Token Bill of Keeping Agents Running

On July 21, Google refreshed three Gemini Flash models. It again held back the flagship 3.5 Pro, instead strengthening the mid-tier — including a 3.6 Flash that cuts output tokens by 17% — to target "the cost of keeping agents running." A developer-focused rundown covering price tags, benchmarks, an