🔥 Follow-up: Fable found the leaks. I fixed them. 48 hours later, I re-measured... and one of them is already growing back. 📏
📝Note: This post is not about ICM, it's about tuning the AI systems you are using with ICM, if you're new to ICM, you can start Here: https://www.skool.com/cliefnotes/welcome-to-clief-notes-heres-where-to-start-2?p=f8f85a09
Let's dive into the results just 48 hours later! 👇
Two days ago, I posted about pointing Fable at my own environment and finding out where my tokens were really going.
A lot of you went hunting in your own setups after that post (and found GOLD 👏).
So, here's the part two: what happened AFTER the audit.
📌 TL;DR - The fixes shipped, the savings are real, and the most important lesson came from re-measuring two days later: optimization is not an event. Drift never sleeps. The audit is worth nothing without a cadence behind it.
🧾 First, the receipts. Every fix from the audit is now live:
✅ The file that lied. My startup menu file claimed "450 tokens" in its own header and was actually ~7,000. It got rebuilt down to ~750, and the detail it was hoarding moved to where it already lived anyway (each project's own handoff file).
💡One fact, one place, and everything else that needs it points to it by design.
✅ The silent tax. My safety hook was injecting ~800 tokens into EVERY prompt, repeating rules that already load once at session start. It's now ~85 tokens: a short pointer to the rules instead of a full copy of them.
💡The enforcement never lived in the repetition; it lives in the hooks that physically block bad actions. Cut the prose, keep the mechanism. That one saves on every prompt, in every session, forever.
✅ The bloated handoff. My portfolio handoff file went from 71KB to 21KB.
Nothing was deleted (nothing is EVER deleted in my environment 😅), the history moved to an archive that loads on demand instead of every session if it is ever needed.
✅ The grunt work tax I was paying full price for. My recon and ritual agents were all silently running on the flagship model.
💡Pre-defining my frontmatter: Read-only file recon now runs on Haiku. Structured ritual writing (plans, reviews, handoffs) runs on Sonnet. The main loop verifies their output, so quality holds while the meter drops. Reasoning models are for reasoning. 🧠
✅ A 176KB session archive that was getting re-read: rotated down to 34KB.
🛡️ Now the part I actually want you to take away from this experience, because THIS is what made it safe to do:
💡Before cutting a single token, I made Fable name the four things that were NOT allowed to break: continuity, identity, rule enforcement, and quality.
Every proposed cut had to state, in writing, which mechanism preserved each one. If a cut couldn't name its preservation mechanism, it didn't happen.
And every single change went through the same gate: dry-run diff ---> my eyes ---> my approval ---> original moved to a dated backup ---> then apply.
The rollback for any change is "put the old file back." Boring. Reliable. Zero fear. 🥱👌
I also defined the acceptance test BEFORE touching anything:
In one week, I compare my usage report against the baseline. Not vibes. Numbers.
If the heavy context share and the ritual overhead don't drop, the plan failed, no matter how good it felt.
(For context, my baseline was ugly: 73% of my weekly usage was happening above 150k context, and 56% came from 6+ hour marathon sessions. The marathons I was proudest of were the leak. 💸)
😂 And now, the plot twist. Today I re-measured everything. (Only 2 days, I'll run again at the 7-day mark)
The menu file that was rebuilt to a 450-token contract two days ago? It's already sitting at roughly 700. The handoff file that was trimmed to 21KB? It grew 15KB back in ONE DAY of normal work.
💡My own environment re-proved the original finding in 48 hours: labels lie because files grow and headers don't. And that's the real lesson of this whole experiment. The cuts buy you the win ONCE. What keeps the win is structure:
📏 Size contracts on every auto-loaded file, re-measured on a schedule (not trusted from the header, we know how that goes 🤣)
🧲 Pull instead of push: the task recalls the 3-5 memories it needs instead of every session loading everything I'm scared to lose 🗂️ One fact lives in ONE place, everything else points to it 🧹 Archives are for archives: if it's not needed to act, it doesn't auto-load
📈 Willpower does not scale. Contracts and cadence do.
🆓 Also, the cheapest wins of all cost zero tokens to implement, because they're behavior, not files:
1️⃣ One task per session. Clear between tasks.
2️⃣ Plan in one session, execute in a FRESH one. The plan file carries the context so the executor starts small and sharp.
3️⃣ No agent fan-outs by default. (UltraCode) A swarm is a tool you reach for on purpose, not a default you drift into. (Ask me about my 62-agent UltraCode adventure 🤠)
These three alone changed the shape of my sessions more than any file edit.
🔭 What's next:
The 7-day acceptance test lands this week, and I'll share the before/after report, wins or losses, either way. (That's the winning together part)
⚡Then the deeper rewiring:
💡My ACTIVE memory tier picks up session pickups, so handoffs shrink to deltas, and intake pulls by task keywords instead of reading fat files.
And this weekend my local models are coming online, which also answers the best question from the last thread (thanks 🙏) and thank you for raising the question of if it was Fable being smarter, or a COLD model with no stake in my setup and no story to defend that made the difference?
A freshly tuned local model is as cold as it gets. And we are all going to find out. 🧊
⭐ If you ran the audit after the last post: don't stop there. Put a date on your calendar and measure again.
💡The gap between what your files claim and what they actually are, is the leak, and it reopens quietly.
The audit finds it. The cadence keeps it closed.
🤝💪🏆 We learn together, we grow together, we win together.
Did anything you fixed after the last post already start drifting back? And what's YOUR re-measure cadence: weekly, monthly, or "when it hurts"? 😄 Tell me in the comments!
23
25 comments
Bas Rosario
7
🔥 Follow-up: Fable found the leaks. I fixed them. 48 hours later, I re-measured... and one of them is already growing back. 📏
Clief Notes
skool.com/cliefnotes
What we give away free beats most paid courses. Build durable AI systems with a Marine vet and Edinburgh researcher. 40+ lessons, growing.
Leaderboard (30-day)
Powered by