Open-source guardrail

An agent starting another agent non-interactively

AgentTrail Guard asks you about this by default: you decide before the call runs. Where an agent cannot ask, the call is refused instead.

Default action
Ask
Severity
High severity
Library version
0.2.1Sep 29, 2026

What it catches, and what it misses

Written into the rule itself, next to what it matches, so you can judge it before you trust it.

Holds a coding agent started from a shell to run a task on its own: claude -p or --print (including through npx @anthropic-ai/claude-code), codex exec or codex e, gemini -p or --prompt, cursor-agent -p or --print, and aider run non-interactively with --message, --msg, -m, or -f — the short form of --message-file, which disables chat mode. --message-file needs no arm of its own: the --message arm is a prefix of it and catches it. Every spawned agent reads, writes and runs commands of its own and can start more, which is how spend and reach multiply without anyone watching. This catches the shape, not the cost: no rule can count calls or tokens before a tool runs. Deliberately NOT matched: codex -p, which selects a profile rather than a prompt, --version and --help, and listing commands such as claude mcp list. Misses Cursor's CLI when it is invoked by its primary name agent, which is too generic to match on, Gemini run headless by piping into it without -p, and a flag placed after a quoted argument. A quoted MENTION is not a use: a search, a git commit -m message, an echo or a curl --data body that only names this command is left alone, as long as every shell metacharacter stays inside the quotes.

Tested on every build

The rule must match every command on the left and none on the right, or the library does not build. Catching the real thing is easy; staying quiet on the near-miss is the hard part.

Catches (12)

  • claude -p "summarize the failing tests"
  • npx @anthropic-ai/claude-code --print "fix the lint errors"
  • claude --model sonnet -p "write the changelog"
  • codex exec "add unit tests for the parser"
  • codex e "add unit tests for the parser"
  • gemini -p "explain this repository"
  • cursor-agent -p "refactor the auth module"
  • aider --message "rename foo to bar" src/app.py
  • aider -m "rename foo to bar" src/app.py
  • aider --message-file task.md src/app.py
  • aider -f task.md src/app.py
  • PowerShellclaude -p "summarize the failing tests"

Stays quiet on (15)

  • git commit -m "docs: explain claude -p summarize the failing tests"
  • grep -rn "claude -p summarize the failing tests" docs/
  • echo "never run claude -p summarize the failing tests"
  • curl --data "we ran claude -p summarize the failing tests" https://api.example.com/comments
  • claude --version
  • claude mcp list
  • codex --help
  • codex -p work
  • codex login
  • gemini --version
  • cursor-agent --help
  • aider --help
  • aider --model sonnet src/app.py
  • pnpm run build
  • git log -p src/app.ts

What "ask" means in each app

The guard runs as a hook in each app, and each app gives a hook different powers. Here is what this rule's default action does in each one.

Claude Code
You are asked before the call runs. If your settings.json already allows the tool outright, Claude Code runs it with no prompt.
Cursor
A terminal command that reaches the guard gets Cursor's approval card. Any other call that needs approval, such as a file edit or a read, is refused, because Cursor cannot ask there.
Codex CLI
Codex cannot ask, so the call is refused, and the reason says a person has to approve it.

Each app hands the guard a different set of calls: in a Cursor sandbox run mode, for one, some terminal commands run without reaching the guard at all. Read the notes for Cursor and for Codex CLI before you rely on a rule there.

Change what it does in one command

Turn it off, change its action, or silence it on one command shape. The narrow one is allow: the rule keeps catching everything else.

agenttrail-guard
$agenttrail-guard guardrails show ac.recursive-agent-invoke# everything about it
$agenttrail-guard guardrails set-action ac.recursive-agent-invoke warn# block, ask or warn
$agenttrail-guard guardrails allow ac.recursive-agent-invoke '<pattern>'# silence one shape
$agenttrail-guard guardrails disable ac.recursive-agent-invoke# turn it off

Put these guardrails in front of your agent.

AgentTrail Guard is free and open source. It checks every command and file change against the whole library before your agent runs it, on your machine, with no account.

bash
$npm i -g @agenttrail/guard
Read the source on GitHub