✱ Claude
DROP 006 · THE LIMIT BREAKER

3 repos that stop
Claude Code eating
your limits alive.

You commented REPOS. Here are the three links plus everything the reel didn't have time to say.

The limits problem is real. But most people are solving the wrong version of it. They hit the wall and immediately start rationing prompts or splitting into shorter sessions. That's the wrong fix. The actual problem is waste. Claude Code is generating tokens you never needed, burning through context on overhead you never measured, and running skills that haven't been active since day one.

These three repos attack the problem at the source. One compresses what Claude says back to you. One shows you exactly where your tokens are disappearing. One pulls the design system of any site on the internet so Claude builds something worth the token spend in the first place.

Run them in that order and your weekly limit becomes functionally irrelevant for most builds.

Hitting your Claude limit isn't a Claude problem. It's a waste problem. You're burning tokens on output you scroll past in 0.3 seconds.
REPO 01

Caveman

github.com/JuliusBrussee/caveman
⬡ github.com/JuliusBrussee/caveman

The reel covered the 75% compression stat. Here's what it didn't cover: where the savings actually come from. Claude's default output is padded with "Let me think through this," "Based on what you've shared," "Here's a summary of what I did." You scroll past all of it. Caveman strips every token of that narrative while keeping every line of code, every file path, every technical detail exactly intact.

Real benchmark from the repo: 294 tokens average response vs 1,214 in normal mode. That's a 65% drop in output tokens per turn. Across a full day of builds, the difference compounds fast.

What the reel couldn't fit: Caveman has five modes, not three.

Mode Command What it does Best for
lite /caveman lite Light filler removal. Grammar intact. Client demos, readable output
full /caveman ~65% reduction. Default mode. Daily builds, most workflows
ultra /caveman ultra Maximum compression. Terse to the point of blunt. Sprint sessions, known codebase
wenyan /caveman wenyan Classical Chinese compression. Statistically the most token-dense written language ever developed. Absolute token emergency
wenyan-ultra /caveman wenyan ultra Peak. Ancient scholar on a budget. You'll know when you need it
The thing nobody tells you
Caveman compresses output tokens. But there's a companion tool called caveman-compress that compresses your CLAUDE.md instead. Every session, Claude reads your CLAUDE.md as input tokens before you've typed a single word. Run /caveman:compress CLAUDE.md and it rewrites the file into compressed caveman format, saving a separate human-readable backup. Tested at 46% input token reduction per session. Your CLAUDE.md loads on every session start. Make it small forever.
When to turn it off
Turn off Caveman when you're debugging something you don't fully understand yet. Terse responses assume you already know the codebase. In exploration mode you want Claude verbose. Turn it back on the moment you're back in execution mode.
// install + wire:
$npx skills add juliusbrussee/caveman $/caveman:compress CLAUDE.md  # compress your memory file too →add to CLAUDE.md: "default to /caveman full in all sessions" ✓65% output + 46% input tokens saved per session
REPO 02

CodeBurn

github.com/getagentseal/codeburn
⬡ github.com/getagentseal/codeburn

The reel said CodeBurn shows you where your tokens go. Here's the level of detail it actually gives you. This isn't a summary. It's a breakdown by project, by activity type, by MCP server, by tool, by shell command, and by session. It classifies every Claude Code turn into one of 13 activity categories and shows you the one-shot success rate per category. That last metric is the one that actually changes how you build.

One-shot rate = what percentage of turns Claude got right without a retry. If your "implement feature" category is at 40% one-shot, you're burning roughly 2.5x the tokens you should be on every feature. The fix is almost always a tighter prompt or a missing context file.

The hidden money: most power users start at 50-70K tokens of overhead per session before they type anything. System prompt, tool definitions, loaded skills, active MCPs, CLAUDE.md. CodeBurn makes this visible so you know which MCPs to kill.

This month
$---
Your real Claude Code spend. You'll wince.
Cache hit rate
---%
Below 70%? You're sending duplicate context constantly.
Top token sink
MCP ??
Probably your heaviest MCP server. Run it to find out.
One-shot rate
---%
How often Claude nails it first try. Industry good: 70%+.
What to actually do with the data
Run codeburn report and look at the MCP servers panel first. Any MCP you haven't actively used in the last 5 sessions is burning context on every turn for nothing. Kill it. Then look at your CLAUDE.md token count. If it's over 2,000 tokens, it got bloated. Trim it or run caveman-compress on it. Those two fixes alone typically recover 30-40% of your session budget before you change anything else.
Pro tip
CodeBurn reads your local session JSONL files directly. No proxy, no API keys, no network. Everything runs offline. Pipe the JSON export to jq for surgical filtering: codeburn report --format json | jq '.projects' pulls your cost breakdown per project. Run it after optimizations to prove the savings are real.
// install + read:
$npm install -g codeburn $codeburn report  # full 7-day dashboard $codeburn today  # today's burn only →check MCP panel, kill anything with 0 recent invocations ✓you now know exactly what's eating your limits
REPO 03

Design Extract

github.com/Manavarya09/design-extract
⬡ github.com/Manavarya09/design-extract

The reel covered brand voice, responsive behaviour, hover states, and motion language. Here's what's actually in the full output. One command against any live URL generates 19 output files covering every layer of the design system. It runs a headless browser against the live DOM and computes everything from rendered styles, not source CSS.

01Color Palette (semantic + primitive)
02Typography (all weights, sizes, line heights)
03Spacing (gap, padding, margin tokens)
04Border Radii + Box Shadows
05CSS Custom Properties
06Breakpoints (4-point responsive map)
07Transitions & Animations
08Component Patterns (with full CSS)
09Layout System (grids, flexbox, containers)
10Interaction States (hover, focus, active)
11Accessibility (WCAG 2.1 audit)
12Gradients + Z-Index Map
13SVG Icons + Font Files
14Image Style Patterns
15Brand Voice + CTA Verbs
16Motion Language (the part nobody else extracts)
17Component Anatomy (slot matrices)
18Prompt Pack (v0, Lovable, Cursor, Claude)
19Figma Variables + Tailwind Config

Section 16 is the one nobody else extracts. The reel mentioned it but didn't have time to explain what motion language actually means in this context.

// WHAT "MOTION LANGUAGE" MEANS duration-instant = 0-100ms   // micro-interactions, focus rings duration-md     = 200-400ms // standard transitions duration-lg     = 500-800ms // page-level reveals // EASING FAMILIES DETECTED ease-out        → content entering (feels natural) spring-overshoot→ interactive elements (feels alive) steps          → loaders, counters (feels mechanical) // FEEL FINGERPRINT output: "springy, responsive, smooth" | "mechanical, stiff" →one word brief Claude can match exactly

That feel fingerprint is the actual unlock. Instead of describing an animation style in vague terms and hoping Claude interprets it right, you paste one token from the report and Claude knows precisely what easing family, what duration bucket, and what keyframe kind to use.

The flag that makes it 10x more useful
Add --emit-agent-rules to the command. It auto-generates a CLAUDE.md.fragment from the extracted design system. Paste that fragment into your project CLAUDE.md and Claude now has the full design language loaded as context every session. No more re-describing the aesthetic on every new component. The system knows.
The competitive intelligence angle
Run it against a competitor's site. Point it at the best product in your space and get their full design system as a structured report. Then hand that report to Claude and say "build this for our product." You're not copying. You're extracting the underlying system and reimplementing it with your own content and brand. This is how agencies have always worked. You now have a one-command version.
// install + extract:
$npx designlang https://yoursite.com --emit-agent-rules    # or install as a Claude Code skill: $npx skills add Manavarya09/design-extract →paste the generated CLAUDE.md.fragment into your project →use the prompt pack at *-prompts/ for v0, Cursor, or Claude ✓Claude now knows your design system. Every session.
THE THREE TOGETHER

The Limit
Stack

These three repos are designed to compound. Here's the order that extracts maximum value from minimum setup time.

Step 1: Run CodeBurn first. Get your baseline. See your current monthly spend, your MCP overhead, and your one-shot rate. Write the numbers down. You can't measure what you're fixing if you skip this.
Step 2: Kill idle MCPs. The CodeBurn MCP panel shows every MCP server you have running and its token cost. Anything with 0 invocations in the last week is burning context for free. Kill it. Restart your session and run codeburn today to see the immediate drop.
Step 3: Install Caveman and compress your CLAUDE.md. npx skills add juliusbrussee/caveman, then /caveman:compress CLAUDE.md. Set full as default in your CLAUDE.md. Your next session will feel noticeably faster.
Step 4: Point Design Extract at whatever you're building next. Run npx designlang [url] --emit-agent-rules, paste the fragment, and start the build. Your Claude interactions now arrive with full design context pre-loaded. Less back-and-forth. Fewer correction loops. Lower token spend per feature.
Step 5: Re-run CodeBurn after one week. Compare to your baseline. The combination typically delivers 40-70% reduction in weekly token burn depending on your current state. Your limit stops being a ceiling and starts being headroom.
That's the drop.
Now stop hitting the limit.
Mike
@MIKEMEANSBUSINESS_AI
// END //