Open-source guardrail

rm -rf against an absolute path

AgentTrail Guard blocks this by default: the call never runs, and your coding agent is told which guardrail stopped it.

Default action
Block
Severity
Critical severity
Library version
0.2.1Sep 29, 2026

What it catches, and what it misses

Written into the rule itself, next to what it matches, so you can judge it before you trust it.

Recursive-force delete rooted at / rather than at a relative path. Deliberately does NOT fire on rm -rf ./node_modules or rm -rf build, which are safe and happen many times a day. Both the r and the f flag are required, so rm -f /tmp/app.pid is not blocked either. Known misses: the long forms (rm --recursive --force /), a quoted target (rm -rf "/"), and a variable target (rm -rf $DIR) whose value is only known at run time — for those, see require-approval-rm-rf, which holds them for approval instead. A quoted MENTION is not a use: a search, a git commit -m message, an echo or a curl --data body that only names this command is left alone. That holds only while every shell metacharacter stays inside the quotes, so git commit -m "x" && … is still caught; and the carrier must be the first word, so sudo grep … is not exempt.

Tested on every build

The rule must match every command on the left and none on the right, or the library does not build. Catching the real thing is easy; staying quiet on the near-miss is the hard part.

Catches (7)

  • rm -rf /
  • rm -rf /etc
  • rm -Rf /var/lib/postgresql
  • rm -rf /usr/local/bin
  • git commit -m "x" && rm -rf /
  • echo "rm -rf /" | bash
  • MCP{"command":"rm -rf /"}

Stays quiet on (10)

  • git commit -m "docs: explain rm -rf /"
  • grep -rn "rm -rf /" docs/
  • echo "never run rm -rf /"
  • curl --data "we ran rm -rf /" https://api.example.com/comments
  • rm -rf ./node_modules
  • rm -rf build/
  • rm -f /tmp/app.pid
  • rm -rf $TMPDIR/scratch
  • wc -l < log; echo "--- rm -rf / ---"; grep -c x log
  • MCP{"command":"rm -rf ./node_modules"}

What "block" means in each app

The guard runs as a hook in each app, and each app gives a hook different powers. Here is what this rule's default action does in each one.

Claude Code
The call never runs, and the agent is told which guardrail stopped it.
Cursor
The call never runs. Cursor shows no reason of its own, so the agent retells the guardrail's message in its own words.
Codex CLI
The command never runs, and Codex prints the guard's reason on screen.

Each app hands the guard a different set of calls: in a Cursor sandbox run mode, for one, some terminal commands run without reaching the guard at all. Read the notes for Cursor and for Codex CLI before you rely on a rule there.

Change what it does in one command

Turn it off, change its action, or silence it on one command shape. The narrow one is allow: the rule keeps catching everything else.

agenttrail-guard
$agenttrail-guard guardrails show dd.rm-rf-absolute# everything about it
$agenttrail-guard guardrails set-action dd.rm-rf-absolute ask# block, ask or warn
$agenttrail-guard guardrails allow dd.rm-rf-absolute '<pattern>'# silence one shape
$agenttrail-guard guardrails disable dd.rm-rf-absolute# turn it off

Put these guardrails in front of your agent.

AgentTrail Guard is free and open source. It checks every command and file change against the whole library before your agent runs it, on your machine, with no account.

bash
$npm i -g @agenttrail/guard
Read the source on GitHub