releases

what’s new,in plain words.

What each release does, how to turn it on, and what it showed on my own machine. For every detail, optimaizr changelog --all prints the full list offline.

0.10.0 · oct 7, 2026 · latest

know when thecache goes cold.

Claude Code keeps your conversation in a cache for up to an hour. While it is warm, each message re-reads it at the cheapest rate there is. Once it expires, the next message writes all of it again. In the second conversation below that is $1.02 for one message, where re-reading it costs $0.025. 0.10 shows you which one you’re in, before you send.

under claude code’s promptmy machine, Oct 7, 2026
// a conversation you’re in
❯
◉ optimAIzr · 122.7K context, $0.025/call to re-read · cache warm 55m
// another one, left for over an hour
❯
◉ optimAIzr · 127.2K context, $0.025/call to re-read · cache expired, next message writes it all again ($1.02) · new task? /clear first
turn it on, in a terminal
$ optimaizr statusline on
// optimaizr statusline off takes it out
// optimaizr statusline alone shows a preview
// no mod needed; never replaces a line you have
how to read it
cache warm 55m

The conversation is cached. Each message re-reads it cheaply, and the minutes are how long that lasts if you stop now.

last 15 minutes

It counts down, and says what coming back after will cost. Leaving? Ask for a handoff note, then /clear.

cache expired

Your next message writes the whole conversation again. Starting a new task? /clear first, and it starts small.

5h

Your 5-hour window, at the end of the line.

in vs code

The chat panel draws no status line. Run optimaizr live in a terminal beside it for the same countdown.

new command · optimaizr sessions

where re-reading takes over.

Every message re-reads the whole conversation, so the longer it runs, the more of each message is re-reading. sessions measures where that happens in your own logs, with nothing estimated.

optimaizr sessionsmy machine, 182 days
110 sessions · 8,399 calls · $1,057 at list prices
a median session: 30 calls, 77K at its largest
re-reading’s share of a call, by context
0-50K24%50K-100K35%100K-150K46%150K-200K52%200K-300K59%300K-500K64%500K+60%
From 150K on, re-reading is half the cost of a call or more.
30 sessions passed 200K and took 86% of spend
cold returns
38 returns to an expired cache
rewrote 12.9M tokens for $102.59
a warm cache would have read them for $3.15
// now a rule of its own: profile counts it as likely
everything else in 0.10.0
live

Counts down too, warns 5 minutes before an expensive cache expires.

cold returns

A new rule, for Claude Code and Codex, priced against a fresh start.

handoff

/optimaizr handoff: a short note for cents, and after /clear the next conversation starts from it.

the mod

Its band counts down the last 15 minutes of the cache.

fairer lever

Compacting prices the reload at the 1-hour rate your conversations use.

--help

Grouped: start here, change things, go deeper.

0.9.0 · oct 6, 2026 · savings you can act on

find the waste.then the real savings.

Before 0.9 a finding was one number. Now each saving says how sure it is: clear waste you can drop with nothing to lose, likely savings if you try them, and levers that save up to an amount, never added together. On my machine that day: $13.07 a month of clear waste, and up to 28% less by compacting long conversations earlier.

optimaizr profilemy machine, Oct 6, 2026
Your Claude Pro did $941.85/mo of work at API prices: 47x what you pay
Savings found
Clear waste$13.07/mo$159.03/yr · nothing to lose
Likely, if you try+$1.54/moquick tasks, less reasoning
Up to, if you test$264.94/mocompact earlier
$54.46/mosmaller default model
Biggest win  compact earlier  TEST · possible
  Up to 28% less usage, about 1.4x the work per 5-hour window
optimaizr apply context-compaction
~/.claude/settings.json
"env": { "CLAUDE_CODE_AUTO_COMPACT_WINDOW": "200000" }
~/.codex/config.toml
model_auto_compact_token_limit = 200000
// from the next session · optimaizr undo context-compaction
what changed
no setup

Your Claude plan is read from Claude Code itself. No --plan.

exact dollars

Clear waste, likely savings and levers, each labelled, never added up.

the big lever

Every call re-reads the conversation. Compact earlier, undo any time.

live

A status line, context warnings, and Y to compact.

codex too

Its own settings and compactions, GPT-5.4 to 5.4 mini.

stricter

Whole tasks, not single calls. Numbers that hold up.

0.8.0 · oct 2, 2026 · claude code mods

now insideclaude code.

Claude Code runs mods. Press Y in optimaizr live, or let --auto do it, and the session you’re in moves to a cheaper model when that pays. The HUD shows what it saved. The mod is optional: everything else works from the terminal.

claude · eatmaxxinga real session, 2 Oct
❯ read the README and list the API routes
✻ Churned for 16s
⏺ optimaizr: this turn $0.18 · 4 requests · 5h 55% → 56%
❯ what does apps/web do? one sentence
✻ Churned for 2s
⏺ optimaizr: this turn $0.06 on Haiku 4.5 · 1 request · $0.05 more than Opus 5.5 once, to load the conversation · 5h 56% → 56%
❯ list the files in apps/api/src and say what each one does
✻ Cogitated for 28s
⏺ optimaizr: saved $0.11 vs Opus 5.5 · this turn $0.05 on Haiku 4.5 · 5 requests · 5h 58% → 60%
optimAIzr  ██████████░░░░░░  60% of 5h · lasts to the 15:00 reset
switched  eatmaxxing: Opus 5.5 → Haiku 4.5 · saved $0.06 · harder task? /optimaizr off
❯
/optimaizr hudlater, on 0.8.1
optimAIzr    ● waiting
saved $0.05    48% cheaper    ≈0.8% of your 5h window
spent $2.11 · $1.21/h · 6 requests · cache 93%
5h  ███████░░░░░░░  48%
        ~1h 37m left at this pace · resets 02:40
7d  ████████████░░  87%
turns  ▆▃    priciest #1 $0.05 · last 2
waiting    eatmaxxing: Opus 5.5 → Sonnet 5.5 · reload $0.27 pays back in ~51 requests · subagents switch now
[ pause switch ] p · Esc closes
// the subagents' requests, priced from their tokens
Opus 5.5100%Sonnet 5.552%
48% less on Sonnet · the HUD shows it as it happens
in claude code 2.1.287 or later
/plugin marketplace add blendbunjaku/optimaizr
/plugin install optimaizr@optimaizr
// local: usage figures in, no prompts, no network
what it adds
optimaizr live

Y, or --auto, switches the session you’re in, when it pays.

while it works

Turn cost and your 5h window, live. /optimaizr hud for the rest.

every answer

What it cost, and what a switch saved, in dollars and of your window.

less thinking

Lower effort, same model.

retry loops

A command that failed twice, held once.

harder task?

/optimaizr off goes back.

before the mods

earlier releases.

0.7.0
Open source, with Team plans
  • The CLI and its engine are open source under MIT, developed in the open on GitHub.
  • --plan team and --plan team-premium show 5-hour sessions for Claude Team seats; every ChatGPT plan Codex reports is named and priced.
  • Claude Sonnet 5.5 is priced as its own model, and verify's judge no longer runs out of room before its verdict.
0.6.1
Live applies to your app's next request
  • Press Y in live and apps using wrap() switch model from their very next request, no restart; optimaizr undo <rule> takes it back.
  • Large Claude Code histories, past roughly 100K responses, are read in full.
  • A source that cannot be read is a red warning at the top of every command, not a note at the end.
0.6.0
Budgets, plans, and the recent rate
  • --budget 300 names the day a monthly cap runs out, and how many days the fixes buy back; live warns at 50, 80, 95 and 100%.
  • ChatGPT limits read straight from Codex, and --plan pro for Claude, with waste per 5-hour window.
  • Monthly figures project the last 30 days instead of averaging your whole history.
  • optimaizr card turns your last 30 days into an image to post: totals only, no project names.
0.5.0
Your whole profile, one command
  • optimaizr profile puts spend, waste and your single biggest opportunity on one screen, with the next command to run.
  • Claude Opus 5.5 and Fable 5.1, GPT-6 and GPT-5.1 through 5.6 are priced at their own rates, long-context and cache writes included.
try it

try 0.10.0 on your own history.

then optimaizr, and optimaizr statusline onRead the docs