Integrate Pxpipe for economic use of Fable-class models

Feature request for product/service

– Other –

Describe the request

GitHub - teamchong/pxpipe: cut Fable 5 token usage by rendering text context as images · GitHub – a token processor that, in Claude Code, allows Fable 5 to be used at an effective price lower than Opus 4.8.

I really like Fable, but it’s too expensive. With pxpipe, the price drops to a reasonable level.

Screenshot / Screen Recording

To give you an idea of ​​the savings: I connected Fable after Opus, which had already accumulated 76k context, but the context shrank to 34k, and after 10 minutes of bug investigation, it only rose to 56k. Without Pxpipe, it would have been 120-180k tokens.

That context drop is useful evidence, but I’d separate prompt-window compression from complete-task savings. The 76k → 34k snapshot does not include later retries, cache pricing, tool calls, or whether the run still solves the same bug. A convincing integration benchmark would repeat the same issue with and without Pxpipe and report success, total billed usage, elapsed time, and run-to-run variance.

Evidence is in the original repo. Furthermore, this method allowed me to use Fable with same or slower quota usage than with Opus 4.8. The very fact that the context doesn’t fly beyond the 250k tokens upon task completion, which often happens to me on Opus 4.8, speaks for itself.

I didn’t notice any poor performance. Furthermore, Fable itself understands what Pxpipe is doing and, when necessary, starts requesting data so that it arrives as text.


Besides, if you have at least Claude Code Pro, you still have time to test it yourself.

Gemini 3.6 Flash is the second model, which can work with Pxpipe

Update for Opus 5

Two five-hour limits of Claude Pro in Claude Code Desktop with Opus 5 XHigh


The estimate is lower than it should be, since in most cases the input should have been counted as a cache write for $12.5/1M.