Tool Monsters View on GitHub →

Spotify cut Claude Code token usage by 90%. Here's what I got when I tested it.

A cheap model does the reading and the boilerplate. Claude keeps the thinking. I rebuilt the setup for plain Claude Code, then ran it on 4 real tasks, with and without.

Based on the setup a Spotify product manager published in September 2026. Theirs runs on Portal and Gemini Flash. This one needs only the claude CLI and uses Haiku. Not affiliated with Spotify.

0reads blocked by the hook, across 4 tasks
0calls to the cheap reader. Claude always found a cheaper path
-28%on boilerplate with the cheap writer, and a more complete result
1 · What's inside

code-writethe part that pays

Haiku writes predictable code (tests, type stubs, config) from a spec plus a reference file, straight to disk. Claude never reads the reference or the output.

bulk-readnever called

Sends whole files to Haiku in one call and returns dense bullets with line numbers. The files never enter Claude's context.

block-big-readsnever fired

A PreToolUse hook that stops any full read over 350 lines and points Claude to the two scripts. Targeted reads pass.

2 skills + benchmark

Skills that tell Claude when to call each script, and the 4 scenarios with the runner, so you can reproduce every number below.

2 · Install
# clone it, install it, restart Claude Code
git clone https://github.com/ToolMonsters/claude-code-routing
cd claude-code-routing && ./install.sh
3 · What I measured

Claude Code 2.1.270, Opus 5, on psf/requests. Each task ran once without the setup and once with it. Total cost includes the Haiku calls.

TaskWithoutWithWhat happened
S1 Inventory of 110 classes and functions across 3 files$0.75$0.71110/110 both. Neither run opened a file.
S2 Every raise and except in 2 files$0.74$0.85Same answer. Without: 2 full file reads. With: ranged reads instead.
S3 Write tests matching a 3,094-line test file$0.96$1.22Worse. 17 tests down to 11. 90s up to 243s.
S4 Write a type stub for a 1,184-line module$0.82$0.59Cheaper and better. 51 of 57 functions covered vs 42. 54s up to 164s.

One run per cell. Treat small gaps as noise.

4 · What that means
5 · Known gaps
Reproduce it: benchmark/run.sh opus clones psf/requests and reruns the 8 sessions. Budget about $6 to $7 with Opus.

Everyone copies the reader.

The writer is the part that pays.

Your next customers are already around you.

We build GTM systems that turn your audience, network and market signals into sales conversations. toolmonsters.com