Official agent skill

Babysit PR

by trailofbits in trailofbits/coop

Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging.

OfficialApache-2.0Auto-check passedTesting & QA

Install Babysit PR

skills CLI
$ npx skills add trailofbits/coop --skill babysit-pr -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install trailofbits/coop babysit-pr --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/trailofbits/coop.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/babysit-pr .claude/skills/babysit-pr && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
babysit-pr
GitHub stars
762
Token cost
~625 tokens
SKILL.md length
333 words
Files
1
Skills in repo
6
Repo updated
First seen
Licence
Apache-2.0

At a glance

Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging.

  • Works in 3 steps: Integrate a behind/conflicting base with… → Reproduce settled CI failures. Fix the… → Address unresolved review feedback only…
  • Asked to babysit
  • SKILL.md covers Preconditions and survey and Act in priority order
  • Calls gh

What it does

Babysit PR is an agent skill from trailofbits/coop, published by the product's own GitHub organization. Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging. Use when asked to babysit, monitor, or fix an existing PR.

Its SKILL.md is about 630 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Failing and flaky tests. The repository describes itself as: Isolated VM environment for running Claude Code and Codex. The licence is Apache-2.0.

When your agent uses it

  • Asked to babysit
  • Fix an existing PR

Example prompts

  • “/babysit-pr”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Integrate a behind/conflicting base with `git merge --no-ff
  2. Reproduce settled CI failures. Fix the underlying fmt, clippy, test, taplo,
  3. Address unresolved review feedback only when the ask is concrete and

What it can do on your machine

Read from SKILL.md and the folder at commit 0f1c4ef. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • gh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use gh, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Babysit PR loads about 625 tokens when it runs. Until then it costs about 50 tokens; SKILL.md has 333 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~50
When it runs · the whole SKILL.md, loaded when a task matches
~625

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from trailofbits/coop at commit 0f1c4ef, republished under its Apache-2.0 licence (© trailofbits). 333 words, ~625 tokens.

Download SKILL.mdSave it as .claude/skills/babysit-pr/SKILL.md (or your agent's skills folder).
name
babysit-pr
description
Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging. Use when asked to babysit, monitor, or fix an existing PR.

Babysit PR

One pass surveys the PR, acts on safe mechanical work, then reports. Do not merge, close, reopen, dismiss reviews, rewrite published history, or edit PR metadata unless asked.

Preconditions and survey

Require an authenticated gh, a clean git worktree, an open PR for the current branch, and ownership of its head branch. Stop on a dirty tree rather than stashing. A merged PR is terminal; a closed-unmerged PR needs user direction.

Gather once:

  • gh pr view with state, draft status, head/base refs and SHAs, merge state, review decision, check rollup, reviews, and comments;
  • base divergence after fetching the base;
  • inline reviewThreads through GraphQL (resolution is authoritative);
  • failed/pending run details and logs.

Treat top-level CHANGES_REQUESTED reviews and structured bot comments as feedback even when no inline thread exists. Do not assume a green-looking list is complete: compare the visible checks with the repository's required CI jobs.

Act in priority order

  1. Integrate a behind/conflicting base with git merge --no-ff origin/<base>. Never rebase. Resolve semantically; stop if the conflict crosses scope.
  2. Reproduce settled CI failures. Fix the underlying fmt, clippy, test, taplo, deny, or workflow issue. Surface flaky/environmental and real-VM integration failures rather than fabricating a code fix.
  3. Address unresolved review feedback only when the ask is concrete and mechanical. Verify the claim first. Escalate design, product, security, and scope decisions.

Make one logical change per commit. Run the narrow test plus format and clippy; run taplo for TOML. Never use --no-verify. Push normally and stop on rejection rather than force-pushing.

After a pushed commit fixes an inline thread, reply exactly Fixed in <sha>. or Fixed in <sha>: <one-line>., then resolve that thread. Do not reply or resolve before the fix is pushed, and never close a judgment-call thread for the user.

Re-survey once after changes. Report pushed commits, failed/pending checks, unresolved feedback, merge state, and absent integration/platform gates. If only an external signal remains, offer to monitor at an appropriate cadence; do not start recurring monitoring unless the user asked.

© trailofbits, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/babysit-pr of trailofbits/coop.

Open the folder on GitHubat commit 0f1c4ef

Compare with similar skills

Babysit PR next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Babysit PR compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Babysit PR this skilltrailofbits/coop762—~625Automated safety check: PassApache-2.0
Swig Testswig/swig6.3k—~2.3kAutomated safety check: PassCustom licence
Dynamo Jira TicketDynamoDS/Dynamo2k—~1.1kAutomated safety check: PassApache-2.0
Fix Ready PRsfastrepl/anarlog9.4k—~1.4kAutomated safety check: PassMIT
Trx Analysismicrosoft/vstest969—~1.8kAutomated safety check: PassMIT
Wioworkersio/skills180—~5.8kAutomated safety check: PassMIT

Similar skills

  • Swig Test

    swig/swig

    Run SWIG test suite for specific languages. An agent skill from swig/swig.

    6.3k GitHub stars~2.3k tokensUpdated today
    Testing & QAAuto-check passed
  • Dynamo Jira Ticket

    DynamoDS/Dynamo

    Create structured Jira tickets for Dynamo from bug reports, failing tests, or feature requests.

    2k GitHub stars~1.1k tokensUpdated today
    Testing & QAAuto-check passed
  • Fix Ready PRs

    fastrepl/anarlog

    Inspect every open non-draft PR for CI failures and unresolved Cursor Bugbot findings, then fix them on the existing PR branches.

    9.4k GitHub stars~1.4k tokensUpdated today
    Testing & QAAuto-check passed
  • Trx Analysis

    microsoft/vstest

    Official

    Parse and analyze Visual Studio TRX test result files. An agent skill from microsoft/vstest.

    969 GitHub stars~1.8k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Wio

    workersio/skills

    Testing workflow skill for finding high-value test candidates, writing focused tests, generating realistic workloads, reviewing test value, and diagnosing test-suite health.

    180 GitHub stars~5.8k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed
  • Quicksilver

    UditAkhourii/quicksilver

    Offload bulk judgment calls to Jev (TypeSafe's fast System One model) so Claude doesn't read, and pay for, content it only needs a verdict on.

    101 GitHub stars~2.2k tokensUpdated 12 days ago
    Testing & QAAuto-check: notes

More from trailofbits/coop

  • Closeout Review

    trailofbits/coop

    Official

    Run the final scope-controlled review before committing, pushing, or opening a PR.

    762 GitHub stars~700 tokensUpdated today
    Auto-check passed
  • Review

    trailofbits/coop

    Official

    Review a coop pull request or local diff with independent, self-validated correctness, design, convention, security, API, test, documentation, and comment lenses.

    762 GitHub stars~3.1k tokensUpdated today
    Auto-check passed
  • Babysit My PRs

    trailofbits/coop

    Official

    Triage and shepherd all open PRs owned by the current GitHub user, isolating each writable worker in its own worktree.

    762 GitHub stars~453 tokensUpdated today
    Auto-check passed
  • Integration

    trailofbits/coop

    Official

    Run and interpret coop's VM integration suite locally on Lima or remotely on Firecracker.

    762 GitHub stars~227 tokensUpdated today
    Auto-check passed
  • Mutation Check

    trailofbits/coop

    Official

    Run cargo-mutants for changed coop logic and keep .cargo/mutants.toml synchronized.

    762 GitHub stars~389 tokensUpdated today
    Auto-check passed

Categories

Questions about Babysit PR

What does Babysit PR do?

Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging. Babysit PR is an agent skill from trailofbits/coop, published by the product's own GitHub organization. Shepherd the current user's open PR through base updates, CI failures, and review feedback without rewriting history or merging.

When should I use Babysit PR?

Babysit PR fits situations like: asked to babysit; fix an existing PR.

How do I install Babysit PR in Claude Code?

Run `npx skills add trailofbits/coop --skill babysit-pr -a claude-code`. Or copy the skill folder (.agents/skills/babysit-pr in trailofbits/coop) into .claude/skills/babysit-pr in your project. Claude Code loads it when a task matches its description.

How do I install Babysit PR in Codex?

Run `npx skills add trailofbits/coop --skill babysit-pr -a codex`. Or copy the skill folder (.agents/skills/babysit-pr in trailofbits/coop) into .agents/skills/babysit-pr in your project. Codex loads it when a task matches its description.

Can I use Babysit PR in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add trailofbits/coop --skill babysit-pr -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/babysit-pr, .gemini/skills/babysit-pr, .github/skills/babysit-pr and .opencode/skills/babysit-pr in your project.

What does Babysit PR need to run?

Going by SKILL.md and its folder, Babysit PR needs the command-line tools its instructions call (gh).

Does Babysit PR access the network?

SKILL.md contains no URLs. Its commands use gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Babysit PR safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Babysit PR use?

Babysit PR is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Babysit PR use?

About 625 tokens (SKILL.md is roughly 2.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Babysit PR?

Skills that share tags, products or a category with Babysit PR: Swig Test (swig/swig, 6.3k stars), Dynamo Jira Ticket (DynamoDS/Dynamo, 2k stars), Fix Ready PRs (fastrepl/anarlog, 9.4k stars) and Trx Analysis (microsoft/vstest, 969 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Babysit PR?

trailofbits (a GitHub organization, an official publisher) maintains it in trailofbits/coop, which has 762 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on October 7, 2026.

Source: trailofbits/coop on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.