Agent skill

Babysit

by getlago in getlago/lago-front

A skill your agent uses when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge.

AGPL-3.0Auto-check passed

Install Babysit

skills CLI
$ npx skills add getlago/lago-front --skill babysit -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install getlago/lago-front babysit --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/getlago/lago-front.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/babysit .claude/skills/babysit && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
babysit
GitHub stars
163
Token cost
~5.2k tokens
SKILL.md length
2,505 words
Files
4 (incl. references)
Skills in repo
17
Repo updated
First seen
Licence
AGPL-3.0

At a glance

A skill your agent uses when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge.

  • Works in 7 steps: Identify or open the PR → 5 - Resolve the mode → Review, report, triage → …
  • Asked to babysit
  • SKILL.md covers Setup, The two durable-state rules, Phase 0 - Identify or open the… and Phase 0.5 - Resolve the mode, plus 6 more sections
  • Calls gh, git and pnpm

What it does

Babysit is an agent skill from getlago/lago-front. Use when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge. Reviews the PR first and reports findings, loops on CI and review feedback, announces the PR in

Its SKILL.md is about 5.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `references/ci-playbook.md`, `references/gh-cookbook.md` and `references/slack-format.md`).

It works with GitHub. The repository describes itself as: Open Source Metering and Usage Based Billing. The licence is AGPL-3.0.

When your agent uses it

  • Asked to babysit
  • Keep working on a GitHub pull request until it is green

Example prompts

  • “/babysit”

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Identify or open the PR
  2. 5 - Resolve the mode
  3. Review, report, triage
  4. The loop
  5. Comment triage
  6. Announce in #frontend
  7. Keep watching

What it can do on your machine

Read from SKILL.md and the folder at commit d4cb9fe. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • gh
    • git
    • pnpm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use gh, git and pnpm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Babysit loads about 5.2k tokens when it runs, and up to ~7.9k if it reads all its reference files. Until then it costs about 66 tokens; SKILL.md has 2,505 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~66
When it runs · the whole SKILL.md, loaded when a task matches
~5.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~7.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from getlago/lago-front at commit d4cb9fe, republished under its AGPL-3.0 licence (© getlago). 2,505 words, ~5,177 tokens.

Download SKILL.mdSave it as .claude/skills/babysit/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
babysit
description
Use when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge. Reviews the PR first and reports findings, loops on CI and review feedback, announces the PR in

Babysit

Take a pull request from wherever it is to ready-to-merge, and keep it there.

Opens the PR if the branch has none, reviews it, gets CI green, answers reviewer and bot comments, announces it in #frontend, then keeps watching until it is ready.

Never merge unless the user explicitly asks.

Setup

The Slack tools are deferred. Load them in one call before starting:

ToolSearch(query: "select:mcp__claude_ai_Slack__slack_search_public,mcp__claude_ai_Slack__slack_read_thread,mcp__claude_ai_Slack__slack_send_message,mcp__claude_ai_Slack__slack_search_channels")

References, read on demand rather than up front:

  • references/gh-cookbook.md - every gh and GraphQL call used below
  • references/ci-playbook.md - CI check to local pnpm command
  • references/slack-format.md - #frontend announcement format

These three describe babysit's own mechanics. No file in this skill restates a coding rule. The repo's conventions live in CLAUDE.md, .agents/docs/ and the lago-frontend-patterns skill, and the review points at those rather than copying them. A copy would drift, and a stale copy enforced during review is worse than no review at all.

The two durable-state rules

Babysit keeps no local state file. Both pieces of memory live in the systems of record, so any session on any machine resolves the same behaviour and teammates can see it.

  1. The decline ledger lives in the PR. Declined bot comments carry a hidden marker in the reply body.
  2. The announcement in #frontend is the mode switch. Its presence means the review and the announce already happened.

Never replace either with a file in the repo or in a scratch directory.


Phase 0 - Identify or open the PR

  1. Parse the argument. Strip any flags first, then treat what remains as the PR number or URL. The only flag is --review, which forces Phase 1 to run even in follow-up mode; record it and remove it before touching gh. No number left over -> find the current branch's PR with gh pr view --json number,url,title,body,headRefName,....
  2. Closed or merged -> report and stop.
  3. No PR for the branch -> open one, ready for review.
    • Refuse only in the degenerate cases: the branch is main, or there are no commits ahead of the base.
    • Push first if the branch has no upstream: git push -u origin HEAD.
    • Title: conventional-commit form, derived from the commits.
    • Body: .github/pull_request_template.md, with Fixes LAGO-XXX filled in from the branch name or commit trailers. Leave the placeholder alone when no ticket can be found.
    • Base: main, unless the branch is visibly stacked on another open PR's head, in which case base on that.
    • Not a draft. Do not prompt. Print the created PR and continue into Phase 1.
  4. Branch and head-ref mismatch. Compare the local branch name to headRefName. In a Conductor worktree they often differ. Every push in this run must then use git push origin HEAD:<headRefName>. Pushing the local branch name creates a stray branch and leaves the PR stale.

Phase 0.5 - Resolve the mode

Before reviewing anything, look for a prior announcement in #frontend. See references/slack-format.md for the query and the boundary check that stops pull/402 from matching pull/4020.

AnnouncementModeBehaviour
Not foundFirst runPhase 1 review, triage, loop, announce, then watch
Found (keep its ts)Follow-upSkip Phase 1. Skip the announce. Straight into the loop and the watch

Follow-up mode is the resume path: the earlier session was closed, ran out of context, or /loop started a fresh one. Skipping the review is deliberate. It already ran and the user already triaged it; re-running would re-litigate settled decisions. /babysit <n> --review forces Phase 1 again when the PR has changed substantially.

Phase 1 - Review, report, triage

First run only. Do not write a review from scratch here. Delegate to the built-in /review, which takes a PR number.

Point it at the repo's conventions; never restate them. CLAUDE.md, .agents/docs/ and the lago-frontend-patterns skill are the single source of truth, and they are what coding sessions already load. The review reads the same files, so a rule can never be enforced in review while being absent from the guidance the code was written against.

Work out which docs the diff touches, then pass their paths:

Diff touchesAlso read
anythingCLAUDE.md
tests, __tests__/, cypress/.agents/docs/testing-practices.md
.graphql, fragments, src/generated/.agents/docs/graphql-fragments.md
new files or directories under src/.agents/docs/folder-architecture.md
a new or unfamiliar library.agents/docs/documentation.md
a list, table or paginated query.agents/skills/lago-frontend-patterns/references/pagination.md
a drawer.agents/skills/lago-frontend-patterns/references/drawers.md
a dialog, modal or confirmation prompt.agents/skills/lago-frontend-patterns/references/dialogs.md
a form.agents/skills/lago-frontend-patterns/references/forms.md
an org id or slug, or an identifier embedding one.agents/skills/lago-frontend-patterns/references/organization-slug.md
a design-system or layout component.agents/skills/lago-frontend-patterns/SKILL.md (and the reference its index row names, if any)

CLAUDE.md already pulls in .agents/docs/typescript-conventions.md itself, and its own sections still cover router imports, MUI imports and translations. The pagination, drawer, dialog, form and organization-slug rules now live in lago-frontend-patterns, listed above. Nothing in it needs repeating here.

Skill(skill: "review", args: `<n>

Review against this repo's conventions. Read these files and treat them as the
authority, in this order:
  CLAUDE.md
  <the .agents/docs and .agents/skills paths selected above>

Weight violations of those documented rules above generic code-review findings.
Do not report anything CI already catches: formatting, type errors, failing tests,
lint. Do not report pre-existing issues on lines this PR did not touch.`)

If a listed path does not exist, say so in the run output rather than reviewing without it. A renamed doc should fail loudly, not silently narrow the review.

Then turn the findings into a triage list:

  1. Drop anything CI already covers (formatting, type errors, failing tests, lint). Those are the loop's job, not a decision for the user.
  2. Drop findings on lines the PR did not touch.
  3. Renumber the survivors and present them.
### Review - PR #4020 (title)

| # | Sev  | File:line            | Finding                                  |
|---|------|----------------------|------------------------------------------|
| 1 | high | usePlanDrawer.tsx:42 | ref-based drawer, lago-frontend-patterns forbids |
| 2 | med  | cache.ts             | new list field not in queryFieldPolicies |

CI: 2 failing (Run linters, Tests shard 3/4) | Reviews: none | Mergeable: clean

Fix which before the loop starts? (all / 1 / none)

Nothing is posted to GitHub in this phase. No findings -> say so and go straight to the loop. This is the one point in the run that waits for the user.

Phase 2 - The loop

Each round:

  1. Refresh. git fetch origin, gh pr view --json ..., and the reviewThreads GraphQL query.
  2. Rebuild the decline ledger. Scan every review thread, including resolved and outdated ones, for <!-- babysit:declined ... --> markers. Rebuilt from the PR each round, so it survives restarts.
  3. Work the highest-priority blocker: draft, then conflicts, then failing checks, then comments, then pending checks, then pending review.
Approval never blocks

The loop runs for hours. Halting on every decision would stall it on round one and leave the CI failure behind it undiscovered. So nothing in the loop is a blocking prompt. Each round sorts work into two piles.

Act now, no approval:

  • CI failures. Map the check to a local command via references/ci-playbook.md, reproduce locally, fix, validate, commit, push. Never push a speculative fix.
  • Bot comments judged DECLINE, and duplicate auto-resolves.
  • Mechanical bot fixes: provably no behaviour change. Typo, missing type annotation, extracted constant, unused import, renamed local, null check on a value already proven non-null on that path.

Queue and keep going:

  • Behavioural bot fixes: anything touching control flow, an API surface, a public prop, error-handling semantics, or falsy handling. || to ?? is behavioural, not a style nit.
  • All human feedback, however mechanical it looks. A human comment can carry intent the diff does not show.
  • Everything classified ESCALATE.

Queued items surface in every round summary and drain the moment the user answers, whether that is immediately or hours later. They never expire and are never silently dropped.

Round 4 | 14:22
  CI      : Run linters failed -> pnpm lint:fix, pushed 8f21ac
  Bots    : 1 declined (duplicate of a3f1c9), 1 mechanical applied (typo)
  Pending : 2 awaiting you
            [1] Copilot: || -> ?? in usePlanDrawer.tsx:42 (behavioural)
            [2] Allan (Slack): reuse useFeatureDrawer instead
  Next    : re-checking in 20 min. Answer any time: "apply 1", "skip 2".
Fixing rules
  • Read the code before changing it.
  • Keep fixes scoped to the blocker. Preserve unrelated worktree changes.
  • Run pnpm code:style once before the final push of a round, not after every edit.
  • Do not start a second copy of a validation command that is already running.
  • Same check failing twice for different reasons: keep going. Twice for the same unclear reason: summarise the evidence and queue it for the user.
  • Wait for a pending check rather than re-triggering it. Run Test E2E takes ~11 min.
  • Before every push, confirm the remote head sha still matches what the round started from. If it moved, abandon this round's push and start a fresh round. Never force.

Phase 3 - Comment triage

Take every unresolved review thread from the reviewThreads query, then split it by the login of its first comment. Both piles must be worked every round: a thread that matches neither rule below has been dropped, which is a bug.

First comment authorPileHandling
copilot-pull-request-reviewer[bot], any *[bot]BotSections A to C below: dedup, then APPLY / DECLINE / ESCALATE
Anyone elseHumanSection D. Always queued, never declined, never auto-resolved
A. Dedup against the ledger (bot threads only)

This is what stops Copilot re-posting a comment already settled.

Fingerprint: first 6 hex of sha1(path + "|" + normalised_body), where the body is lowercased, code fences and suggestion blocks stripped, and whitespace collapsed.

Do not strip digits. Line numbers live in the thread's own fields, not in the comment body, and path is the only positional value hashed, so drift cannot move the fingerprint anyway. Removing digits buys nothing and actively collides: "limit should be 20" and "limit should be 50" on one file normalise to the same string, and the second, genuinely new comment gets silently resolved as a duplicate without being read.

  • Exact fingerprint in the ledger -> resolve the thread immediately with a one-line reply linking the original decline. No re-analysis. One line in the round summary, nothing more.
  • No exact hit -> compare against the ledger's topic= slugs, of which there are only ever a handful. Same file and same topic, just reworded -> duplicate. Resolve it and record the new fingerprint as an alias so the next variant matches exactly.
Show full SKILL.md (1,063 more words)Show less
B. Decide, for genuinely new comments
VerdictWhenAction
APPLYReal bug, or a concrete CLAUDE.md violationMechanical: apply and push. Otherwise queue
DECLINEContradicts CLAUDE.md, pre-existing, a linter or typechecker concern, a nitpick, or wrong about the codeReply with reasoning and marker, resolve
ESCALATEProduct-sensitive, ambiguous, or a judgment call about intentQueue for the user, leave the thread open

DECLINE needs evidence, not an opinion: cite the CLAUDE.md rule, the file:line, or the git history that makes the comment wrong. A comment that is merely tedious to handle is an ESCALATE, not a DECLINE.

C. Post the decline
markdown
Not applying: <one or two sentences of concrete reasoning, citing a rule or file:line>.

<!-- babysit:declined v1 fp=a3f1c9 path=src/foo/useBar.tsx topic="ref-drawer-pattern" -->

Then resolve the thread. The marker does not render in GitHub's UI but is present in the API body, which is what makes the ledger work.

D. Human threads

Every unresolved thread from a non-bot author becomes a pending queue item, one per thread, carrying the reviewer's login, the path, and the comment text. Nothing here is ever auto-applied, auto-declined, or auto-resolved, however mechanical it looks: a human comment can carry intent the diff does not show.

  • Skip the ledger entirely. Fingerprints and decline markers are a bot-duplication defence and have no meaning for a person who wrote the comment once.
  • Disagreeing with a reviewer is an ESCALATE. Report the disagreement with reasoning and let the user answer the reviewer; babysit does not argue with humans on the PR.
  • A thread stays queued until the user answers. Only the user's answer closes it, and resolving the thread is the user's call, not babysit's.
  • Applied fixes get a reply with the commit sha, and nothing else.

A reviewer leaving "changes requested" must therefore always show up in the round summary. If a round reports an empty queue while reviewDecision is CHANGES_REQUESTED, the split above was not applied. Treat that as a bug in the run, not as a quiet PR.

Phase 4 - Announce in #frontend

Reaching this point means the babysitting worked. That is when the PR should reach the team.

Gate, all of which must hold:

  • PR open and not a draft.
  • All required checks passing, no merge conflicts.
  • No unresolved threads classified APPLY or ESCALATE. Threads that were DECLINEd and resolved do not block. That is the point of the ledger.
  • Not already announced.

Compose per references/slack-format.md and post to C04DJLU0KHD with slack_send_message. No confirmation prompt. Print the message and its permalink so the run's output shows exactly what went out. Keep the returned ts as the thread anchor.

Gate not met -> skip the announce, say which condition failed, and carry on watching. The gate is re-evaluated every Phase 5 round until the PR is announced, so the common case of arriving here while Run Test E2E is still pending resolves itself the moment it goes green. Announcing is not a one-shot checkpoint.

Phase 5 - Keep watching

Announcing is not the end. Keep looping until the PR is ready to merge, now on a 20 minute interval, because it is waiting on humans rather than CI.

Each round is a delta check, not a full re-read. Compare head sha, check conclusions, review-thread count, and the Slack thread's latest reply ts against the previous round. Nothing changed and nothing pending -> one line, back to waiting. If the PR is not yet announced, re-run the Phase 4 gate as part of every round.

Foreground sleep is blocked by the harness, so pick the waiting tool by how many notifications the wait produces. Do not mix the two idioms.

What wakes a round. Monitor runs bash, so it can poll GitHub but cannot call the Slack tools. A reply in the announcement thread therefore does not wake anything by itself. Two things close that gap:

  • The poll loop emits a heartbeat every third quiet cycle (about hourly). Read the Slack thread on every wake, heartbeat included, not just on GitHub deltas.
  • A reviewer replying in the thread has usually also touched the PR, which fires a real delta anyway.

So Slack feedback is always acted on, but a Slack-only reply during an active watch can sit up to an hour before it is seen. Follow-up mode has no such lag: a fresh /babysit <n> reads the thread immediately.

  • Monitor for a stream of one event per change. Give it a poll loop that prints a line only when something moved, plus the heartbeat, so quiet rounds stay cheap:

    bash
    prev=""; quiet=0
    while true; do
      cur=$(gh pr view <n> --json headRefOid,state,mergeStateStatus,reviewDecision \
              --jq '"\(.headRefOid[0:8]) \(.state) \(.mergeStateStatus) \(.reviewDecision)"' 2>/dev/null || true)
      if [ -n "$cur" ]; then
        if [ -z "$prev" ]; then echo "ARMED: $cur"
        elif [ "$cur" != "$prev" ]; then echo "CHANGED: $cur"; quiet=0
        else quiet=$((quiet+1)); [ $((quiet % 3)) -eq 0 ] && echo "QUIET x$quiet: $cur"
        fi
        prev=$cur
      fi
      sleep 1200
    done

    Set persistent: true; the watch is session-length. Guard the gh calls with || true so one flaky request does not kill the monitor.

  • Bash with run_in_background for a single wake, using an until loop that exits once the condition holds. One notification, then done.

An until loop handed to Monitor prints nothing and so notifies nothing: it sits armed until timeout while the round never fires. That is the failure to avoid.

gh pr checks <n> --watch --interval 30 still covers the short CI waits inside a round.

Slack thread as a second feedback source

Read the announcement thread with slack_read_thread(C04DJLU0KHD, ts) on every wake, heartbeat rounds included, since no Slack reply can wake the monitor on its own. Replies are humans, so they follow the human rules exactly: always queued, never auto-applied, and disagreement escalates. Anything newer than babysit's own last reply is new.

After fixes land, post one batched reply in the thread with the commit sha. One per round, not one per comment.

Feedback this round:
  GitHub threads : 1 new (copilot, duplicate of a3f1c9, resolved)
  Slack thread   : 2 new (Allan, Mimmo)
Exit the watch

Always print the resume command on the way out.

  • Ready to merge. Report and stop. Never merge unless explicitly asked.
  • 6 consecutive quiet rounds (~2 hours) with an empty queue. Reviewers are not looking today. Exit with /loop 20m /babysit <n>, which does the same watch unattended without holding a session open.
  • Context running low. Exit deliberately with a written handoff and the same /loop command, rather than degrading mid-round.
  • Protected action needed, CI unavailable long enough that waiting is pointless, or the user stops it.

Final report

  • PR URL and state.
  • What was fixed, and what was verified.
  • Declined comments, one line each.
  • Slack permalink, if announced.
  • Unanswered pending items.
  • Remaining human action.

© getlago, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (references) in .agents/skills/babysit of getlago/lago-front.

  • SKILL.md
  • references/ci-playbook.md
  • references/gh-cookbook.md
  • references/slack-format.md

Open the folder on GitHubat commit d4cb9fe

Compare with similar skills

Babysit next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Babysit compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Babysit this skillgetlago/lago-front163—~5.2kAutomated safety check: PassAGPL-3.0
PR Babysitteropeninterpreter/openinterpreter69k3 repos~4.2kAutomated safety check: PassApache-2.0
Diagnosing Superpowers Sessionsobra/superpowers296k3 repos~1.7kAutomated safety check: PassMIT
GitHub Deep Researchbytedance/deer-flow83k5 repos~1.3kAutomated safety check: PassMIT
Greplooponyx-dot-app/onyx32k4 repos~3.3kAutomated safety check: PassMIT
Update V8 Versionopeninterpreter/openinterpreter69k2 repos~845Automated safety check: PassApache-2.0

Similar skills

  • PR Babysitter

    openinterpreter/openinterpreter

    Watches an open GitHub pull request until it merges, handling review comments, diagnosing CI failures and retrying flaky checks along the way.

    69k GitHub starsUsed in 3 repos~4.2k tokens
    DevelopmentAuto-check passed
  • Investigates a session where Superpowers went wrong, reads the transcripts on disk and produces an evidence-cited report, optionally prepared as a bug report for the maintainers.

    296k GitHub starsUsed in 3 repos~1.7k tokens
    Agent WorkflowsAuto-check passed
  • GitHub Deep Research

    bytedance/deer-flow

    Researches a GitHub repository over four rounds using the GitHub API and web search, then writes a structured markdown report with timeline, metrics and Mermaid diagrams.

    83k GitHub starsUsed in 5 repos~1.3k tokens
    Research & ScienceAuto-check passed
  • Greploop

    onyx-dot-app/onyx

    Iteratively improves a PR (GitHub), MR (GitLab), or shelved changelist (Perforce) until Greptile gives it a 5/5 confidence score with zero unresolved comments.

    32k GitHub starsUsed in 4 repos~3.3k tokens
    DevelopmentAuto-check passed
  • Update V8 Version

    openinterpreter/openinterpreter

    Bumps the pinned v8 and rusty_v8 versions in Codex, validates the release-candidate path with the v8-canary check, and traces failures to upstream build changes.

    69k GitHub starsUsed in 2 repos~845 tokens
    DevOps & CloudAuto-check passed
  • Last30days

    mvanhorn/last30days-skill

    Research what people actually say about any topic in the last 30 days.

    64k GitHub stars~7.8k tokensUpdated today
    Research & ScienceAuto-check: notes

More from getlago/lago-front

All 17 skills in this repo
  • Cve Doctor

    getlago/lago-front

    Triage a CVE / Dependabot alert in a JS/TS project and recommend the least-invasive fix.

    163 GitHub stars~2.9k tokensUpdated today
    Auto-check passed
  • Extract Section To Drawer

    getlago/lago-front

    Extract a Formik form section into a TanStack Form drawer with Zod validation, following the plan form migration pattern.

    163 GitHub stars~4k tokensUpdated today
    Auto-check: notes
  • Loop Build

    getlago/lago-front

    Phase 2 of the loop pipeline for lago-front. An agent skill from getlago/lago-front.

    163 GitHub stars~3.5k tokensUpdated today
    Auto-check: notes
  • Loop Clean

    getlago/lago-front

    Cleanup phase of the loop pipeline for lago-front, for the worktree layout only.

    163 GitHub stars~816 tokensUpdated today
    Auto-check passed
  • Docker Expert

    getlago/lago-front

    You are an advanced Docker containerization expert with comprehensive, practical knowledge of container optimization, security hardening, multi-stage builds, orchestration patterns, and production…

    163 GitHub starsUsed in 10 repos~3.6k tokens
    Auto-check passed
  • Loop Flywheel

    getlago/lago-front

    Harvest phase of the loop pipeline for lago-front. An agent skill from getlago/lago-front.

    163 GitHub stars~1.9k tokensUpdated today
    Auto-check passed

Works with

Questions about Babysit

What does Babysit do?

A skill your agent uses when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge. Babysit is an agent skill from getlago/lago-front. Use when asked to babysit, monitor, shepherd, or keep working on a GitHub pull request until it is green, review-ready, approved, mergeable, or ready to merge.

When should I use Babysit?

Babysit fits situations like: asked to babysit; keep working on a GitHub pull request until it is green.

How do I install Babysit in Claude Code?

Run `npx skills add getlago/lago-front --skill babysit -a claude-code`. Or copy the skill folder (.agents/skills/babysit in getlago/lago-front) into .claude/skills/babysit in your project. Claude Code loads it when a task matches its description.

How do I install Babysit in Codex?

Run `npx skills add getlago/lago-front --skill babysit -a codex`. Or copy the skill folder (.agents/skills/babysit in getlago/lago-front) into .agents/skills/babysit in your project. Codex loads it when a task matches its description.

Can I use Babysit in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add getlago/lago-front --skill babysit -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/babysit, .gemini/skills/babysit, .github/skills/babysit and .opencode/skills/babysit in your project.

What does Babysit need to run?

Going by SKILL.md and its folder, Babysit needs the command-line tools its instructions call (gh, git and pnpm).

Does Babysit access the network?

SKILL.md contains no URLs. Its commands use gh and git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Babysit safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Babysit use?

Babysit is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Babysit use?

About 5.2k tokens (SKILL.md is roughly 21k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.7k tokens, read only when the agent opens those files.

What are the alternatives to Babysit?

Skills that share tags, products or a category with Babysit: PR Babysitter (openinterpreter/openinterpreter, 69k stars), Diagnosing Superpowers Sessions (obra/superpowers, 296k stars), GitHub Deep Research (bytedance/deer-flow, 83k stars) and Greploop (onyx-dot-app/onyx, 32k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Babysit?

getlago (a GitHub organization) maintains it in getlago/lago-front, which has 163 GitHub stars. The repository holds 17 skills in this directory. The repository was last updated on October 8, 2026.

Source: getlago/lago-front on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.