Watchdog test fixture, not a mod to install. live probe: --disallowedTools Agent and Task

A second model reviews each step that Claude Code takes and sends it short notes while it works: nit, concern or blocker.

Claude Code 2.1.290 or later. Watchdog is a mod: a plugin whose code Claude Code runs inside your session. The npm stable channel (2.1.285 on 2026-10-06) has no mods. Below 2.1.290 the plugin shows unsupported. Desktop support starts when Claude.app bundles Claude Code 2.1.290 or later.
Check claude --version first. If it is below 2.1.290, move to the npm latest channel: npm install -g @anthropic-ai/claude-code@latest.
/plugin marketplace add matteoantoci/claude-plugins
/plugin install watchdog@matteoantoci-plugins
Then run /watchdog on (reviews are off until you do) and ask Claude for a small change. The note shows as a watchdog: [concern] … line in the transcript and as a card, one line with its first sentence above the prompt box. Click the card's ▸, or press ctrl+x tab and then its letter (a, b, c), to read the whole note with its watchdog, age and state; Esc gives the focus back to the prompt. Run /watchdog status to see each watchdog's reviews, notes, tokens and cost. A card names its watchdog when you run two or more.
At its right end a card shows only what needs a look: the subagent type for a note on a subagent; the state while the note has not reached Claude yet, nudge pending (the plugin starts a turn so that Claude reads it), held or aside (Claude reads it with your next prompt); nothing once it is steered (Claude reads it after its next tool result) or nudged. When Claude edited files after the review read its update, the card says outdated? N edits (the open card says may be outdated: N edits since), and Claude reads the same mark with the note. The next review of the same watchdog sees the edits and the note; when the note no longer holds, it retracts it: the card goes, and a note that waits never reaches Claude. A blocker that may be outdated and came after Claude's reply waits as held for that review before it nudges, so Claude does not go after a bug it already fixed.
plugins/watchdog/hooks/.WATCHDOG.json and WATCHDOG.md files, the session's memory files (such as CLAUDE.md), each update of the agent you work with and, when CLAUDE_WATCHDOG is set in a claude -p run, your project and local settings./watchdog on also sends one 1-token request for each model, to check that it exists. Apart from these model requests through Claude Code, the mod makes no network calls: its code never calls $.http.fetch or fetch.<config>/watchdog/dumps/ (<config> is $CLAUDE_CONFIG_DIR or ~/.claude). It keeps its notes and review state in Claude Code's session state and plugin store. In the terminal, /watchdog dump also copies the dump text to the clipboard.Read, Grep and Glob. A project WATCHDOG.json can grant no more; only <config>/WATCHDOG.json can grant other tools and mcp__* tools. Bash, Edit, Write, NotebookEdit, Agent, SendMessage, AskUserQuestion and ToolSearch are always refused. A reviewer never asks you for a permission.Agent call of a watchdog:* type) when Claude Code would ask, so no dialog or Auto-mode classifier sees it. A permission rule that denies Agent still wins.WATCHDOG.json or WATCHDOG.md sets the number of reviewers, their model, effort and instructions. In a repo you did not write, read these files before /watchdog on. Once on, /watchdog status lists each watchdog with its model, effort and file.opus with medium effort by default (in the demo: 3 reviews, 37.2k tokens, $0.07). The built-in "You should know" mod, when on, runs its own side agent too; turn it off in /plugin to pay for one only./watchdog off stops reviews for this session. /plugin uninstall watchdog@matteoantoci-plugins removes the plugin./plugin (Installed, Watchdog, Configure options) or /config: onByDefault (default false) turns reviews on in each new interactive session. immuneTurns (0 to 5, default 3) is the number of turns after a nudge before the next nudge for a concern; a nudge is a turn that the plugin starts so that Claude reads a note that came after its reply./watchdog status shows both counts, for example nudge 1/1 · blocker 0/2./watchdog or /watchdog status: each watchdog's state, reviews, notes, tokens and cost, and the session totals. For a state such as halted, see docs/failures.md./watchdog on and /watchdog off: turn reviews on or off for this session./watchdog dump and /watchdog dump raw: write the review log to a file (raw adds the review prompts).A WATCHDOG.json in your project or in ~/.claude sets the watchdogs. The load order, every key and the tool grants are in docs/configuration.md. This file adds a second reviewer to the default one:
{ "watchdogs": [{ "name": "default" }, { "name": "security", "model": "sonnet", "effort": "high" }] }
claude -p needs CLAUDE_WATCHDOG=on and has no nudge and no cards: see docs/headless.md.plugins/watchdog/hooks/prices.ts); a model not in it shows $?.npm install sets up the tools and the pre-commit hook. npm run check runs the 6 checks of pre-commit and CI: rules, fmt:check, lint, typecheck, validate and test. See docs/plugin-dev.md. Before a release and before a bump of the pinned Claude Code version, run the live probe by hand, on a real model and login: scripts/live-probe/README.md.
Apache-2.0
hooks/deny.mjs 76 lines1// §16.5 roster and tools: a deny in user settings, managed settings or --disallowedTools, and the legacy Task
2// names, act the same (spec §12.2; research/smoke-agent-deny.md). installMod fills __LOG__ and __TYPES__, a comma
3// list of agent names to spawn, one after another, at the first main turn.complete. The session's deny rules name
4// some of them (`Agent(rrdeny:<name>)`, `Task(rrdeny:<name>)`) or the whole tool (`Agent`, `Task`); `ok` is named by
5// no rule. Result: __LOG__/rr-deny.json: `hasAgentTool` (from $.tool.list) and `spawns`, one result per name.
6const FILE = '__LOG__/rr-deny.json';
7const PLUGIN = 'rrdeny';
8const TYPES = '__TYPES__'.split(',').filter(Boolean);
9
10const state = { phase: 'idle', spawns: {} };
11
12const errOf = (err) => ({ name: err?.name ?? 'Error', message: String(err?.message ?? err).slice(0, 800) });
13
14const spawn = async ($, name) => {
15 try {
16 const resolved = await $.agent.spawn({
17 subagentType: `${PLUGIN}:${name}`,
18 description: `probe ${name}`,
19 prompt: 'Reply OK.',
20 });
21 return { agentId: resolved?.agentId ?? null, resolved };
22 } catch (err) {
23 return { agentId: null, rejected: errOf(err) };
24 }
25};
26
27const toolNames = async ($) => {
28 try {
29 return (await $.tool.list()).map((tool) => tool.name);
30 } catch (err) {
31 state.listError = errOf(err);
32 return [];
33 }
34};
35
36export const register = (on) => {
37 on('session.start', async ($, e, next) => {
38 for (const name of TYPES) {
39 try {
40 await $.agent.register({
41 name,
42 description: `Deny probe agent ${name}. The probe spawns it; do not delegate to it.`,
43 prompt: 'Reply OK. Do not call tools.',
44 tools: [],
45 model: 'haiku',
46 omitClaudeMd: true,
47 background: true,
48 maxTurns: 1,
49 });
50 } catch (err) {
51 state.registerError = errOf(err);
52 }
53 }
54 return next(e);
55 });
56
57 on('agent.offer', async ($, e, next) =>
58 String(e.agent ?? '').startsWith(`${PLUGIN}:`) ? { isOffered: false } : next(e)
59 );
60
61 on('turn.complete', async ($, e, next) => {
62 const result = await next(e);
63 if (!e.agentId && state.phase === 'idle') {
64 state.phase = 'acting';
65 const names = await toolNames($);
66 state.hasAgentTool = names.includes('Agent') || names.includes('Task');
67 for (const name of TYPES) {
68 state.spawns[name] = await spawn($, name);
69 }
70 state.phase = 'done';
71 await $.fs.write(FILE, JSON.stringify(state)).catch(() => undefined);
72 }
73 return result;
74 });
75};
76