Adds quiet flags to noisy shell commands before they run (git status, pytest, cargo, npm install, mvn, curl, wget, docker pull), so Claude reads short output…

Make the $20 Claude Pro plan last longer in Claude Code.
Twenty-two small mods that show you exactly where your usage goes and cut the waste. One mod makes a model call: prompt-polish, once each time you press Improve, and never on its own. The others make none. None adds anything to the system prompt (read-cap adds one tool, which Claude Code lists by name only until Claude first uses it; write-guard adds one line to a tool result at most once per conversation; compact-keeper adds a note of at most 4,000 characters after a compaction): every figure on screen is one Claude Code already reports, or a time the mod measured.
| Mod | What it does | Where |
|---|---|---|
| tool-diet | Loads tools you have not used lately on demand instead of with every request | Everywhere |
| skill-diet | Lists skills you have not used lately in this project by name only, without their descriptions | Everywhere |
| agent-diet | Runs Explore subagents on Haiku instead of your main model | Everywhere |
| context-xray | /xray opens the exact breakdown of what fills your context window | Everywhere |
| pro-hud | Live meters above the prompt for your 5-hour session, your week and the context window, plus a per-turn receipt of tokens in, cached and out | Claude desktop app |
| output-diet | Trims long shell output, search results and subagent reports before Claude reads them, keeping the head, the tail and the error lines; the untrimmed text is saved to a file Claude can open without a permission prompt | Everywhere |
| reread-guard | Skips Claude re-reading a file it already read when the file has not changed, with Read or a plain cat, sed -n, head, tail or Get-Content; a deliberate retry still goes through | Everywhere |
| write-guard | Steers Claude to Edit instead of rewriting an existing file in full with Write; a deliberate rewrite still goes through | Everywhere |
| loop-guard | Holds back a shell command that already failed twice in a row, so Claude changes approach | Everywhere |
| cmd-diet | Adds quiet flags to noisy shell commands before they run, so Claude reads short output from the start; errors, failures and warnings stay in full | Everywhere |
| gh-account | Runs each git push, pull, fetch, clone and gh command as the logged-in GitHub account that can see the repository, without switching the active account | Everywhere |
| cache-clock | Counts down until the prompt cache expires; once it has, shows exactly how many tokens your next message will re-send uncached | Everywhere |
| read-cap | Stops Claude reading a file over 1,000 lines whole (with Read, cat or Get-Content) and gives it an outline tool, so it reads only the lines it needs; a retry still reads the whole file | Everywhere |
| session-receipt | /receipt opens a pane with the exact tokens every turn of the session spent, and the costliest turns | Everywhere |
| peek | Typing status while background tasks run opens a pane with each one's exact elapsed time and last output lines, instead of sending the prompt to Claude | Everywhere |
| budget-guard | Holds a prompt back once your 5-hour or weekly usage reaches your limit (90% by default); sending it again goes through | Everywhere |
| turn-budget | When one turn uses more than +5 session points, writes a handoff and continues in a fresh session by itself; /handoff any time | Everywhere |
| compact-keeper | After a compaction, adds a note of exact facts from before it: your latest prompt in full, the todo list, the last failed command, files edited | Everywhere |
| collision-guard | Asks before Claude edits a file another chat on this machine changed in the last 30 minutes | Everywhere |
| answer-pane | Explain, plan and ELI5 pages drawn natively in a side pane; plans have decision buttons and Respond fills the prompt box | Desktop app (no diagrams in the terminal) |
| prompt-polish | An Improve button beside Send (above the prompt in the terminal) rewrites your draft with Opus at low effort and puts it back in the box; Undo restores it. One model call per press | Everywhere |
| kit-updates | Tells you when an installed mod from this kit has a newer version or a new mod joins the kit; /kit-update installs updates and new mods | Everywhere |

In Claude Code:
/plugin marketplace add VedantAndhale/claude-pro-kit
/plugin install tool-diet@claude-pro-kit
/plugin install skill-diet@claude-pro-kit
/plugin install agent-diet@claude-pro-kit
/plugin install context-xray@claude-pro-kit
/plugin install pro-hud@claude-pro-kit
/plugin install output-diet@claude-pro-kit
/plugin install reread-guard@claude-pro-kit
/plugin install write-guard@claude-pro-kit
/plugin install loop-guard@claude-pro-kit
/plugin install cmd-diet@claude-pro-kit
/plugin install gh-account@claude-pro-kit
/plugin install cache-clock@claude-pro-kit
/plugin install read-cap@claude-pro-kit
/plugin install session-receipt@claude-pro-kit
/plugin install peek@claude-pro-kit
/plugin install budget-guard@claude-pro-kit
/plugin install turn-budget@claude-pro-kit
/plugin install compact-keeper@claude-pro-kit
/plugin install collision-guard@claude-pro-kit
/plugin install answer-pane@claude-pro-kit
/plugin install prompt-polish@claude-pro-kit
/plugin install kit-updates@claude-pro-kit
Updates are off by default for marketplaces you add yourself. With kit-updates installed you are told when a fix ships and /kit-update installs it, from the desktop app or the terminal. Without it, turn on auto-update once: in a terminal, run claude, then /plugin → Marketplaces → claude-pro-kit → Enable auto-update; the desktop app has no toggle for it.
Install any one on its own; they do not depend on each other. Mods are not sandboxed, so read the code before installing: each mod is a single file under plugins/<name>/hooks/.
Every tool listed in front sends its whole description and schema with every request. A deferred tool is listed by name only, and Claude loads it through ToolSearch when it needs it; Claude Code already does this for most MCP tools. tool-diet does it for the rest of the tools you are not using: anything outside the core set (Bash, PowerShell, Read, Edit, Write, Glob, Grep, Agent, Skill, ToolSearch, TodoWrite, AskUserQuestion) that you have not used in your last five sessions.
Measured with one prompt, "Reply with just OK.", in a fresh session, as the API reported each request:
| Prompt tokens per request | |
|---|---|
| Without tool-diet | 43,859 |
| With tool-diet | 27,905 |
| Change | −15,954 (−36%) |
On that setup, the largest tool moved was Artifact, whose description alone is 19,870 characters. What moves depends on your tools and your habits; /xray shows yours.
/tool-diet lists what is on demand this session, grouped by where each tool comes from; /tool-diet keep <tool> always loads one, unkeep undoes it, and /tool-diet off|on switches it. Answers are toasts, so they add nothing to the conversation. The status line keeps the count: 38 tools on demand.
Every request carries the skill listing: each installed skill's name and its whole description. With a few plugins installed that is dozens of skills, most of which a given project never uses (video skills in a backend repo, document skills in a game). skill-diet keeps the skills you used lately in this project listed in full and lists the rest on one line by name only. The Skill tool still loads any of them, and typing /name still works.
Measured with one prompt, "Reply with just OK.", in a fresh session with 45 skills installed and the other mods on, as the API reported the request:
| Prompt tokens per request | |
|---|---|
| Without skill-diet | 27,423 |
| With skill-diet | 19,197 |
| Change | −8,226 (−30%) |
In the same setup, asked to fill in a PDF form, Claude found anthropic-skills:pdf from its name alone and loaded it with the Skill tool. What moves depends on your skills; /xray shows yours.
/name or called by Claude through the Skill tool./skill-diet shows what is listed by name only this session and how many characters left the listing; /skill-diet keep <skill> always lists one in full, unkeep undoes it, and /skill-diet off|on switches it. Answers are toasts, so they add nothing to the conversation. The status line keeps the count, in the form <n> skills by name only, <n> characters off./xray shows the skills line before and after, in tokens.A subagent runs on your main model unless its definition or Claude's Agent call names another. Explore only searches and reads, yet in the author's own 85 subagent transcripts, every one of the 2,043 Explore requests ran on Opus or Sonnet, reading 147,298,992 input tokens. agent-diet starts the agent types you list (Explore by default) on Haiku.
/agent-diet shows the setting and how many subagents it moved this session; /agent-diet model haiku|sonnet|opus picks the model, /agent-diet add|remove <agent type> changes the list, and /agent-diet off|on switches it. Answers are toasts, so they add nothing to the conversation. The status line keeps the count, in the form 2 subagents on haiku.Measured with claude -p --model sonnet --output-format json on a task that sends one Explore subagent to find a file, three runs each in alternating order. Every run found the file. Figures are the cost and tokens Claude Code reported per model:
| Run | Without: cost | Without: Sonnet tokens in | With: cost | With: Sonnet tokens in | With: Haiku tokens in |
|---|---|---|---|---|---|
| 1 | $0.05692 | 68,395 | $0.03724 | 40,213 | 27,848 |
| 2 | $0.05577 | 68,111 | $0.03764 | 39,593 | 27,872 |
| 3 | $0.05441 | 68,123 | $0.03736 | 40,184 | 42,770 |
| Average | $0.05570 | $0.03741 |

/xray opens a pane with the exact breakdown /context computes: what is sent with every request (system prompt, tools, MCP tools, memory files, skills, messages), what is loaded on demand, which MCP tools load every time, and each memory file's size. It measures when the pane opens and when you press Refresh (or r), never in the background, because the exact count sends one token-count request per tool and memory file.
Opens by itself when the context crosses 60% and again at 80%, with a toast pointing at /compact and /handoff. /xray auto off keeps it to /xray only.
The recording at the top of this page is the band. In text:
Session ━━━━━━━━━━━━━━────── 69% resets in 32m
Week ━━━━━━━━━━━━━━━───── 74% resets in 4d 17h
Context ━━━━━━━━──────────── 44% 436,034 tokens
This turn 1m 35s · 1 tool · 0 files edited · 436,034 in · 431,260 cached · 580 out
On a wide window the three meters sit side by side on one row; on a narrow one each figure on the turn line wraps whole rather than being cut off.
/compact when it climbs.· session 20%, and finished tool calls draw as one line: status dot, tool, target, time./hud shows what is on; /hud all on|off, or /hud band|spinner|cards on|off. The answer is a toast, so toggling adds nothing to the conversation. It draws in the desktop app only and leaves the terminal as it is. Rows other mods put above the prompt still draw beneath the band.
Weekly toasts at 50%, 75% and 90% of the week, once each, because the weekly limit drains quietly across many sessions.
When a Bash or PowerShell result runs past 120 lines or 8,000 characters, Claude reads:
[output-diet: 82/500 lines shown; all 500 in ~/.claude/projects/<project>/<session>/tool-results/output-diet-<call>.txt]
<first 30 lines>
… [lines 31–450 omitted; error/warning lines from them:]
250│ ERROR: build failed in src/app.ts
300│ Warning: deprecated API
…
<last 50 lines>
3 trimmed · 41,200 chars saved.tool-results folder, beside the transcript, where Claude Code keeps the outputs it saves itself, so Claude can open it without a permission prompt. When that folder cannot be found, it goes under ~/.claude/output-diet/ (or CLAUDE_CONFIG_DIR).Grep result past the same limits keeps its first 100 lines, with a note to narrow the search: [output-diet: first 100/252 lines shown; narrow the search, or read all in …]. In the author's 1,048 past Grep results, 26 went past the limits, and keeping the first 100 lines would have cut 104,052 characters. Glob is left alone: none of 128 went past.When Claude asks to Read the same range of the same file again in the same conversation, and the file's size and modification time have not changed, the read is skipped and Claude is told to use the copy it has. Any edit, a different range, or a subagent (which has its own context) reads freely. Claude Code can clear old tool results from context, so retrying the identical read straight after a skip always goes through. The record resets on /compact and /clear.
What Claude is told in place of the repeat read, kept to one line because the model reads it:
api.ts unchanged since you read it; use that copy. If it's gone from context, retry the same Read.
Shell reads too. Claude often reads a file with the shell instead of Read: in the author's transcripts, 318 sed -n runs went past the guard. A shell command that only prints one file is now treated the same way: cat FILE, sed -n 'A,Bp' FILE, head -n N / tail -n N, and in PowerShell Get-Content, gc, cat or type (with -TotalCount, -Head, -First, -Tail, -Last or -Raw). The same range of the same unchanged file is skipped once, and retrying the same command goes through. A whole cat after a whole Read that showed every line counts as a repeat. A pipe, a chain, a variable, a glob or any other flag runs untouched. A shell read never counts as a Read, since Edit needs a real one first.
The status line counts them: 2 re-reads skipped.
Write sends the whole file as Claude's output, the most expensive kind of token; Edit sends only the lines that change. Claude sometimes rewrites a file it has already read in full to change a few lines. In the author's own 272 session transcripts, 141 writes rewrote a file Claude had already read or written (999,273 characters), and 112 of them followed an earlier rewrite in the same session. Of the 512,847 characters in the rewrites whose previous version was in the transcript, 171,325 had changed.
A hook runs after Claude has written the content, so write-guard cannot save the rewrite it sees; it stops the ones after it:
You rewrote all of api.ts. For changes to an existing file use Edit: it sends only the changed lines.api.ts exists; change it with Edit, not a full Write. If a full rewrite is intended, retry the same Write. Retrying the same Write goes through./clear starts over.1 full rewrite held back.Each retry of a failing command re-sends the whole conversation, and the same command usually fails the same way. In the author's 272 session transcripts, 10 commands failed 3 or more times, 38 runs between them.
Once the exact same Bash or PowerShell command has failed twice in a row, the next try is held back once, and Claude is told:
This exact command failed 2 times in a row. Change the approach instead of rerunning it. If a rerun is intended, retry the same command.
/clear starts over.1 failing retry held back.output-diet trims long output after a command has run; cmd-diet keeps the noise from being printed at all. Before a Bash or PowerShell command runs, a known noisy command gets its own quiet flags, which drop progress lines and keep errors, test failures and warnings:
| Command | Runs as | Output, measured |
|---|---|---|
git status | git status --short --branch | 415 to 58 characters |
pytest | pytest -q | 1,063 to 495 characters, failure details unchanged |
cargo build / test / check / clippy / run | cargo build -q | 197 to 0 characters on success, warnings unchanged |
npm install / ci | npm install --no-audit --no-fund | 62 to 18 characters |
curl (Bash only) | curl -sS | 1,049 to 577 characters, the progress meter removed |
mvn | mvn -B -ntp | not measured: drops download progress |
wget (Bash only) | wget -nv | not measured: one line per file |
docker pull | docker pull -q | not measured: drops layer progress |
In a live headless run (git status && python -m pytest on 41 test files, one failing), the request cost 41,535 tokens without cmd-diet and 40,571 with it, by the API's usage, and both runs named the failing test and its reason correctly. The shorter output stays in context, so every later request in the session sends less too.
&&, || or ; chain is handled on its own. A step that pipes, redirects or substitutes (| grep, > file, $(...)) is left alone, since something else reads its output; 2>&1 is fine.git status --porcelain, pytest -v, curl -fsSL, npm install --silent) is left alone, and a flag already present is not added twice.3 commands quieted.With two GitHub accounts logged in to gh, a push to a repository the other account owns fails, and Claude spends turns on gh auth status and gh auth switch. Before a Bash or PowerShell command runs, gh-account matches each git push, pull, fetch, clone, ls-remote and gh step to the repository's owner, and the owner to a logged-in account. When that account is not gh's active one, that step alone runs with its token:
GH_TOKEN="$(gh auth token --user ACCOUNT)" git push
The token itself never appears in the command or its output, and the active account is not switched. For git it also asks gh for the credential.
-R owner/repo, a gh api repos/<owner>/... path, or the folder's remotes. With no remote named, every remote must point at the same owner.gh auth status, and organizations from gh api user/orgs. What it learns is kept across sessions.gh auth and other gh commands that reach no repository, a command that already sets GH_TOKEN, a token from the environment, git over ssh, hosts other than github.com, and subshells or substitutions./gh-account shows which account each owner uses, and /gh-account forget clears it. Answers are toasts. The status line counts them: 2 commands sent as <account>.Claude's prompt cache keeps your conversation for a fixed time after each request. Reply within it and the context is read from cache; reply after it and the whole context is sent again at full price. That message pays the cache-write price (1.25 times the input price for a 5-minute cache, 2 times for a 1-hour one) on every token instead of the cache-read price (0.1 times): 12.5 to 20 times more for the same context, and nothing on screen says so.
cache-clock puts the countdown in the status line, restarted by every response from the main conversation:
cache warm · 3m left
cache cold · next message re-sends 61,204 tokens
When the cache runs out, a toast says so once (after the lifetime is known, see below). The token figure is the previous response
hooks/register.ts 178 lines1import type { Register } from 'claude-code'
2
3// Some commands print pages of progress lines Claude has to read. Before a
4// shell command runs, known noisy ones get their own quiet flags, which drop
5// the progress lines and keep errors, failures and warnings. Only a plain
6// command is touched: one that pipes, redirects or substitutes is left alone,
7// since something else reads its output. If a tool rejects a flag (an old
8// version), that rule stops for the session and Claude's retry runs as typed.
9
10type Rule = {
11 name: string
12 // The command's start; flags go right after it, before any `--`.
13 prefix: RegExp
14 // Each entry is a flag and its spellings; one already present is not added.
15 add: string[][]
16 // A token that means the person chose the output level: leave the command alone.
17 conflicts: RegExp
18 bashOnly?: boolean
19}
20
21const RULES: Rule[] = [
22 {
23 name: 'git status',
24 prefix: /^git\s+status(?=\s|$)/,
25 add: [['--short', '-s'], ['--branch', '-b']],
26 conflicts: /^(--porcelain.*|-v+|--verbose|--long|-z)$/,
27 },
28 {
29 name: 'pytest',
30 prefix: /^(pytest|py\.test|python3?\s+-m\s+pytest|py\s+-m\s+pytest|uv\s+run\s+pytest)(?=\s|$)/,
31 add: [['-q']],
32 conflicts: /^(-[a-zA-Z]*[qv][a-zA-Z]*|--quiet|--verbose)$/,
33 },
34 {
35 name: 'cargo',
36 prefix: /^cargo\s+(build|b|test|t|check|c|clippy|run|r|doc|bench)(?=\s|$)/,
37 add: [['-q', '--quiet']],
38 conflicts: /^(-v+|--verbose|--message-format.*)$/,
39 },
40 {
41 name: 'npm install',
42 prefix: /^npm\s+(install|i|ci|add)(?=\s|$)/,
43 add: [['--no-audit'], ['--no-fund']],
44 conflicts: /^(--audit.*|--fund.*|--silent|--loglevel.*|-d+|--verbose)$/,
45 },
46 {
47 name: 'mvn',
48 prefix: /^(mvn|mvnw|\.\/mvnw|mvnw\.cmd|\.\\mvnw(\.cmd)?)(?=\s|$)/,
49 add: [['-B', '--batch-mode'], ['-ntp', '--no-transfer-progress']],
50 conflicts: /^(-q|--quiet|-X|--debug)$/,
51 },
52 {
53 name: 'docker pull',
54 prefix: /^docker\s+pull(?=\s|$)/,
55 add: [['-q', '--quiet']],
56 conflicts: /^$/,
57 },
58 // In Windows PowerShell, curl and wget name Invoke-WebRequest.
59 {
60 name: 'curl',
61 prefix: /^curl(?=\s|$)/,
62 add: [['-sS']],
63 conflicts: /^(-[a-zA-Z]*[sv][a-zA-Z]*|-#|--silent|--verbose|--progress-bar|--trace.*)$/,
64 bashOnly: true,
65 },
66 {
67 name: 'wget',
68 prefix: /^wget(?=\s|$)/,
69 add: [['-nv', '--no-verbose']],
70 conflicts: /^(-[a-zA-Z]*[qvd][a-zA-Z]*|--quiet|--verbose|--debug)$/,
71 bashOnly: true,
72 },
73]
74
75const SHELLS = new Set(['Bash', 'PowerShell'])
76
77const REJECTED = /unknown|unrecognized|unexpected|invalid|no such option|not recognized|unknown flag/i
78
79// Splits at top-level `&&`, `||`, `;` and newlines, outside quotes.
80// Returns undefined for anything it does not fully understand.
81export function segments(command: string): string[] | undefined {
82 const out: string[] = []
83 let cur = ''
84 let quote: string | undefined
85 for (let i = 0; i < command.length; i++) {
86 const c = command[i]
87 if (quote !== undefined) {
88 cur += c
89 if (c === quote) quote = undefined
90 continue
91 }
92 if (c === "'" || c === '"') {
93 quote = c
94 cur += c
95 continue
96 }
97 const two = command.slice(i, i + 2)
98 if (two === '&&' || two === '||') {
99 out.push(cur, two)
100 cur = ''
101 i++
102 continue
103 }
104 if (c === ';' || c === '\n') {
105 out.push(cur, c)
106 cur = ''
107 continue
108 }
109 cur += c
110 }
111 if (quote !== undefined) return undefined
112 out.push(cur)
113 return out
114}
115
116// Pipes, redirects, background jobs, substitutions and heredocs: leave alone.
117const PLUMBING = /[|<>&`]|\$\(/
118
119// `2>&1` still sends everything to Claude, so it does not count as plumbing.
120function unquoted(segment: string): string {
121 return segment.replace(/'[^']*'|"[^"]*"/g, '""').replace(/\s2>&1(?=\s|$)/g, '')
122}
123
124export function rewrite(command: string, shell: string, off: ReadonlySet<string> = new Set()): { command: string, rules: string[] } {
125 const parts = segments(command)
126 if (parts === undefined) return { command, rules: [] }
127
128 const rules: string[] = []
129 const rebuilt = parts.map((part, index) => {
130 if (index % 2 === 1 || PLUMBING.test(unquoted(part))) return part
131 const lead = part.match(/^\s*/)?.[0] ?? ''
132 const body = part.slice(lead.length)
133 for (const rule of RULES) {
134 if (off.has(rule.name) || (rule.bashOnly === true && shell !== 'Bash')) continue
135 const hit = body.match(rule.prefix)
136 if (hit === null) continue
137 const tokens = unquoted(body).trim().split(/\s+/)
138 if (tokens.some(t => rule.conflicts.test(t))) return part
139 const flags = rule.add.filter(spellings => !spellings.some(s => tokens.includes(s))).map(s => s[0])
140 if (flags.length === 0) return part
141 rules.push(rule.name)
142 return `${lead}${hit[0]} ${flags.join(' ')}${body.slice(hit[0].length)}`
143 }
144 return part
145 })
146
147 return { command: rebuilt.join(''), rules }
148}
149
150export const register: Register = on => {
151 const off = new Set<string>()
152 let shortened = 0
153
154 on('tool.call', async ($, e, next) => {
155 if (!SHELLS.has(e.tool) || typeof e.command !== 'string') return next(e)
156
157 const slim = rewrite(e.command, e.tool, off)
158 if (slim.rules.length === 0) return next(e)
159
160 const ran = await next({ ...e, command: slim.command })
161 if (ran.deny !== undefined) return ran
162
163 // A tool too old for a flag says so; stop that rule so the retry runs as typed.
164 if (ran.isError === true && typeof ran.text === 'string' && REJECTED.test(ran.text)) {
165 const added = slim.command.split(/\s+/).filter(t => !e.command.split(/\s+/).includes(t))
166 if (added.some(flag => ran.text.includes(flag))) {
167 for (const rule of slim.rules) off.add(rule)
168 return ran
169 }
170 }
171
172 shortened += 1
173 $.ui.status(`${shortened} ${shortened === 1 ? 'command' : 'commands'} quieted`)
174
175 return ran
176 })
177}
178