New to Claude Skills? Learn how to install them →

wshobson on GitHub

Review Agent Governance

Free

Control AI agent review actions with human approval.

Get this skill

Free · Opens the source repo

What Review Agent Governance does

The Review Agent Governance skill is designed to enforce human approval for AI agent actions within Claude Code, specifically for tasks such as pull request reviews, comments, merges, and CI configuration edits. This skill ensures that every action taken by the AI agent is backed by explicit human consent, creating a secure and auditable trail of approvals. Each approval or denial generates an Ed25519-signed receipt, which can be verified later, ensuring accountability and traceability in automated workflows.

This skill is particularly useful in collaborative environments where multiple stakeholders are involved in code reviews and CI/CD processes. By implementing a human-in-the-loop mechanism, it mitigates the risks associated with automated actions that could potentially disrupt project integrity. The setup process includes installing the plugin, copying default policy files, and creating a directory for storing receipts, which are essential for maintaining a record of approvals and denials.

During usage, the skill requires a simple workflow where human reviewers must approve actions before they are executed by the AI agent. This can be done through a flag file or a slash command within Claude Code. Additionally, the skill supports a dry-run mode to evaluate policies without bypassing approvals, making it suitable for stringent auditing requirements. The integration of this skill with existing CI/CD pipelines enhances security and compliance, especially in projects that involve sensitive code modifications or deployments.

Overall, Review Agent Governance is ideal for teams looking to maintain a robust governance framework around their AI agents, ensuring that all automated actions are subject to human oversight. This skill is particularly valuable in regulated industries or projects where accountability is paramount, providing peace of mind to teams managing critical codebases.

When to use it

Use this skill in projects where AI agents perform significant actions like PR reviews or CI modifications, requiring human oversight.

When not to use it

Avoid this skill if the agent's tasks are limited to local file edits and running tests, as it may introduce unnecessary complexity.

What you can build with it

Collaborative Code Review

In a team setting, use this skill to ensure all AI-generated code reviews are approved by a human before merging.

Sensitive CI Configuration Changes

Implement this skill when AI agents need to modify CI configurations to ensure all changes are vetted by a developer.

Automated Release Management

Use this skill to control AI actions during automated release processes, ensuring human oversight for critical deployments.

How to install Review Agent Governance

View source

1. Install with the skills CLI

npx skills add wshobson/agents/review-agent-setup --agent claude-code

2. Or install it manually

Download the skill folder and drop it into ~/.claude/skills/ for all projects, or .claude/skills/ to scope it to one repo. Restart Claude Code so it picks up the new skill.

Anthropic's agentic coding CLI, and the reference implementation of Agent Skills. Drop a skill folder into ~/.claude/skills and Claude Code loads it automatically whenever a task matches the skill's description. Claude Code docs

Inside SKILL.md

Written by wshobson

review-agent-governance — Setup

Gate AI agent review actions (PR reviews, comments, merges, CI edits) behind explicit human approval. Every attempt, approved or denied, produces an Ed25519-signed receipt.

When to use this plugin

Install it in projects where a Claude Code agent:

  • Reviews, comments on, or merges pull requests (gh pr review, gh pr merge)
  • Triages issues (gh issue comment, gh issue close)
  • Publishes releases (gh release create)
  • Modifies CI configuration (.github/workflows/, .gitlab-ci.yml)
  • Pushes to protected branches (main, master, release, production)
  • Posts to external notification surfaces (Slack webhooks, Discord)

If the agent is only doing local file edits and running tests, this plugin is overkill. Use protect-mcp for general tool-call policy enforcement and skip this one.

One-time setup

1. Install the plugin

claude plugin install wshobson/agents/review-agent-governance

2. Copy the default policy to your project

cp .claude/plugins/review-agent-governance/policies/review-agent-governance.cedar \
   ./review-governance.cedar

You can edit this file to match your project's specific rules. See ../agents/review-policy-author.md for guidance on authoring review policies.

3. Create a receipts directory and sign key

mkdir -p ./review-receipts
echo "./review-receipts/" >> .gitignore
echo "./review-governance.key" >> .gitignore
echo "./.review-approved" >> .gitignore

The first invocation of protect-mcp sign will create the key. Commit the public key from the first receipt so auditors can verify later.

Per-session workflow

The Cedar policy denies review-surface actions unconditionally. To approve a specific action, open an approval window before it and close it after.

Flag file (simplest)

# Before the action you want to approve
touch ./.review-approved

# Let Claude Code run the review / comment / merge

# Immediately after
rm ./.review-approved

Slash command (from within Claude Code)

/approve-review "Reviewing PR #123 authored by contributor X"

This creates ./.review-approved with the given reason embedded as a note, and writes a human-approved receipt to the chain. A follow-up rm is still needed to close the window.

Dry-run everything (force full policy evaluation)

If you want every tool call to go through Cedar with no approval bypass:

export REVIEW_APPROVAL_FLAG=./.never-approve

Any tool call matching a forbid rule will be denied; approved windows have no effect. Useful for CI or for a locked-down audit run.

Verifying the chain

List all receipts:

ls -la ./review-receipts/

Verify the entire chain offline:

npx @veritasacta/verify ./review-receipts/*.json

Exit 0 means every receipt is authentic and the chain is intact. Exit 1 means one receipt has been tampered with. Exit 2 means a receipt is malformed.

Look at recent denials:

/list-pending

Within Claude Code this slash command walks the receipt chain and prints any recent decision: deny entries with the tool name, command pattern, and timestamp.

Example: approving a PR review

# 1. Human reviews the agent's proposed comment
$ /list-pending
  Recent denials:
  - 2026-04-17T14:23:01Z  Bash "gh pr review 42 --approve --body 'LGTM'"
  - 2026-04-17T14:23:02Z  Bash "gh pr comment 42 --body 'Looking good'"

# 2. Human decides the first one is appropriate, approves it
$ /approve-review "Approving LGTM on PR 42 after visual inspection"
  ./.review-approved created

# 3. Agent retries the action; this time it succeeds
$ agent: gh pr review 42 --approve --body "LGTM"
  [receipt: rec_XXX, decision=allow, reason=human_approved]

# 4. Human closes the window
$ rm ./.review-approved

Every step is in the receipt chain. The chain is offline-verifiable for regulators, counterparties, or downstream auditors who want to confirm that no review action bypassed the human gate.

Composing with protect-mcp

If both plugins are installed, run them side by side:

{
  "hooks": {
    "PreToolUse": [
      {
        "matcher": ".*",
        "hooks": [
          {
            "type": "command",
            "command": "npx protect-mcp@0.7.4 evaluate --policy ./protect.cedar --tool \"$TOOL_NAME\" --input \"$TOOL_INPUT\" --fail-on-missing-policy false"
          }
        ]
      },
      {
        "matcher": ".*",
        "hooks": [
          {
            "type": "command",
            "command": "if [ -f ./.review-approved ]; then exit 0; fi; npx protect-mcp@0.7.4 evaluate --policy ./review-governance.cedar --tool \"$TOOL_NAME\" --input \"$TOOL_INPUT\" --fail-on-missing-policy false"
          }
        ]
      }
    ]
  }
}

Both hooks must pass for the tool call to proceed. Cedar deny in either policy blocks it.

Standards

  • Ed25519 — RFC 8032 (digital signatures)
  • JCS — RFC 8785 (deterministic JSON canonicalization)
  • Cedar — AWS's open authorization policy language
  • IETF draftdraft-farley-acta-signed-receipts

Frequently asked questions about Review Agent Governance

Similar skills