Agent skill

Babysit PR To Pass CI

by sgl-project in sgl-project/sglang

Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head.

Apache-2.0Auto-check passedDevelopment

Install Babysit PR To Pass CI

skills CLI
$ npx skills add sgl-project/sglang --skill babysit-pr-to-pass-ci -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install sgl-project/sglang babysit-pr-to-pass-ci --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/sgl-project/sglang.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/babysit-pr-to-pass-ci .claude/skills/babysit-pr-to-pass-ci && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
babysit-pr-to-pass-ci
GitHub stars
37k
Used in
2 other repos
Token cost
~3k tokens
SKILL.md length
1,667 words
Files
2
Skills in repo
32
Repo updated
First seen
Licence
Apache-2.0

At a glance

Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head.

  • Works in 6 steps: Resolve the PR and exact workflow set… → Inspect the current goal. → If no unfinished goal exists, create one… → …
  • Asked to monitor
  • SKILL.md covers Invocation and arguments, Start or continue the durable…, Preflight and Track only current-head runs, plus 5 more sections
  • Calls git and gh

What it does

Babysit PR To Pass CI is an agent skill from sgl-project/sglang. Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head. Use when asked to monitor, babysit, retry, or fix PR CI for lint.yml, pr-test.yml, pr-test-extra.yml, AMD, or other named workflows; classify failures as PR-related versus flaky or infrastructural, auto-fix and push only small clean fixes, rerun failed jobs only up to 10 times, and ignore unselected workflows.

Its SKILL.md is about 3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in Development, covering CI/CD, Linting and formatting and Pull requests. It works with SGLang and GitHub Actions. The repository describes itself as: SGLang is a high-performance serving framework for large language models and multimodal models. The licence is Apache-2.0.

When your agent uses it

  • Asked to monitor
  • Fix PR CI for lint.yml
  • Pr-test-extra.yml
  • Other named workflows

Example prompts

  • “/babysit-pr-to-pass-ci”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Resolve the PR and exact workflow set first.
  2. Inspect the current goal.
  3. If no unfinished goal exists, create one immediately using the product's goal mechanism. Include the resolved PR URL, selected workflow…
  4. Include these constraints in the goal objective
  5. If the active goal already represents this same PR and workflow set, continue it instead of creating another goal.
  6. If a different unfinished goal is active, do not replace it silently. Report the conflict and ask the user to pause or clear it.

What it can do on your machine

Read from SKILL.md and the folder at commit dab108b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git
    • gh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use git and gh, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Babysit PR To Pass CI loads about 3k tokens when it runs. Until then it costs about 118 tokens; SKILL.md has 1,667 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~118
When it runs · the whole SKILL.md, loaded when a task matches
~3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from sgl-project/sglang at commit dab108b, republished under its Apache-2.0 licence (© sgl-project). 1,667 words, ~2,967 tokens.

Download SKILL.mdSave it as .claude/skills/babysit-pr-to-pass-ci/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
babysit-pr-to-pass-ci
description
Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head. Use when asked to monitor, babysit, retry, or fix PR CI for lint.yml, pr-test.yml, pr-test-extra.yml, AMD, or other named workflows; classify failures as PR-related versus flaky or infrastructural, auto-fix and push only small clean fixes, rerun failed jobs only up to 10 times, and ignore unselected workflows.

Babysit a PR Until Selected CI Passes

Drive only the selected workflows to green on the PR's latest head commit. Keep working across turns by starting a durable goal; do not stop merely because a run was dispatched, a retry began, or a fix was pushed.

Invocation and arguments

Accept:

text
$babysit-pr-to-pass-ci [PR_NUMBER|PR_URL] [WORKFLOW ...]
$babysit-pr-to-pass-ci [PR_NUMBER|PR_URL] --only WORKFLOW [WORKFLOW ...]
  • If the PR is omitted, resolve the PR associated with the current branch.
  • If workflows are omitted, select lint.yml and pr-test.yml.
  • Without --only, add named workflows to the two defaults.
  • With --only, monitor exactly the named workflows.
  • Accept a workflow basename such as pr-test-extra.yml, a repository path such as .github/workflows/pr-test-extra.yml, or an unambiguous workflow display name.
  • Treat extra as pr-test-extra.yml.
  • Accept natural-language equivalents, such as “PR 12345 plus extra and AMD.” Resolve “AMD” to an exact workflow file from the user's wording or repository context; do not guess when multiple AMD workflows fit.

Examples:

text
$babysit-pr-to-pass-ci
$babysit-pr-to-pass-ci 12345
$babysit-pr-to-pass-ci https://github.com/sgl-project/sglang/pull/12345 pr-test-extra.yml
$babysit-pr-to-pass-ci 12345 --only pr-test-amd.yml

Start or continue the durable goal

Treat explicit invocation of this skill as authorization to create a goal and perform the scoped CI actions below.

  1. Resolve the PR and exact workflow set first.
  2. Inspect the current goal.
  3. If no unfinished goal exists, create one immediately using the product's goal mechanism. Include the resolved PR URL, selected workflow files, and this stopping condition: every selected workflow has a successful run for the PR's latest head SHA.
  4. Include these constraints in the goal objective:
    • Ignore every unselected workflow.
    • Diagnose each selected-workflow failure as PR-related, unrelated/infrastructure, or still uncertain.
    • Implement, validate, commit, and push a PR-related fix only when it is clean, non-tricky, and at most 100 changed lines.
    • Stop for user review before editing when the proposed fix is tricky, dirty, or over 100 changed lines.
    • Retry unrelated failures at most 10 times per workflow and head SHA, using failed-jobs-only reruns.
    • Never rerun a full workflow merely to recover failed tests.
  5. If the active goal already represents this same PR and workflow set, continue it instead of creating another goal.
  6. If a different unfinished goal is active, do not replace it silently. Report the conflict and ask the user to pause or clear it.

Do not mark the goal complete until the completion check at the end of this skill succeeds.

Preflight

  1. Verify GitHub authentication and resolve the repository without changing remote configuration.
  2. Resolve the PR and record at least:
    • PR number and URL
    • open/draft state
    • base branch and base SHA
    • head branch and head SHA
    • head repository owner/name, including whether it is a fork
    • labels
    • whether maintainers may modify the branch
  3. Require an open PR. If it is draft and a selected workflow is blocked by the draft gate, stop and ask the user to mark it ready; do not change draft state automatically.
  4. Normalize every selected workflow to an existing file under .github/workflows/. Fail fast on ambiguous or nonexistent names.
  5. Read each selected workflow's triggers and gates. Do not assume all optional workflows behave like pr-test.yml.
  6. Inspect git status. Preserve all user changes. If a fix becomes necessary and the current checkout is dirty, on a different branch, or otherwise unsafe, use an isolated worktree for the PR head.

Known SGLang gates:

  • pr-test.yml requires the run-ci label for relevant PR changes. If its existing gate job failed because the label was absent, add run-ci when authorized and rerun failed jobs in that existing run; adding the label alone does not necessarily create a new pr-test.yml run.
  • pr-test-extra.yml requires both run-ci and run-ci-extra. It listens for relevant label events, so adding a missing label may create a fresh run. Prefer that fresh current-head run; rerun failed jobs in an existing current run only when needed.
  • AMD-extra workflows may use the same run-ci plus run-ci-extra gate. Verify the actual selected file.
  • Treat draft, cooldown, maintenance, and missing-label failures as gate/infrastructure conditions, not code regressions. Fix only the safe gate condition in scope; wait for cooldown or maintenance windows rather than burning retries immediately.

Adding the known CI opt-in labels is within scope. Do not add unrelated labels, mark the PR ready, alter reviewers, merge, or mutate other PR metadata.

Track only current-head runs

Use both PR checks and workflow-run data so workflow names, run IDs, attempts, jobs, and logs can be correlated. Prefer workflow file IDs over display names when filtering runs.

For each selected workflow:

  1. Select the newest run associated with the current PR head SHA.
  2. Ignore runs for older head SHAs, including green runs made stale by a new push.
  3. Treat queued, requested, waiting, pending, and in_progress as active; keep monitoring.
  4. Treat a workflow as passing only when its current-head run concludes success.
  5. If no current-head run exists, wait briefly for event delivery, then inspect the workflow's path filters, event triggers, and label gates.
  6. Trigger only through a mechanism explicitly supported by that workflow. Never invent workflow_dispatch inputs or dispatch a workflow against the default branch when the intent is to test the PR head.
  7. If a selected workflow cannot safely be triggered for the PR head, report the exact blocker and request direction.

After any user push or skill-authored push, re-read the PR head SHA, discard stale run selections, and start tracking the new head. Reset retry counters for the new head SHA.

Monitor persistently

  • Maintain a compact ledger containing workflow file, head SHA, run ID, attempt, status, conclusion, skill-initiated retry count, and latest failure signature.
  • Poll at reasonable intervals, normally 30–60 seconds while runs are active. Give concise progress updates during long waits.
  • Re-check the PR head SHA on each monitoring cycle and before every rerun or push.
  • Ignore failures, cancellations, and pending checks from unselected workflows. Do not diagnose, retry, fix, or wait for them.
  • Do not declare success while a selected workflow is absent, stale, pending, skipped, cancelled, neutral, or failed.
Show full SKILL.md (699 more words)Show less

Diagnose a selected-workflow failure

  1. Inspect the failed jobs and failed-step logs. In pr-test.yml, distinguish the root failure from cascade failures caused by check-pr-test-health, wait jobs, or final aggregation.

  2. Record the exact failing test/check, error text, affected runner or hardware, and relevant log URL.

  3. Compare the failure against the PR's changed files and diff.

  4. Classify it:

    PR-related when evidence ties the failure to changed code, tests, dependencies, configuration, generated files, or behavior introduced by the PR. A reproducible failure on the PR that does not occur on the base is strong evidence.

    Unrelated/infrastructure when evidence points to a transient network or download error, runner loss, GPU/driver instability, service outage, pre-existing main-branch failure, known flaky threshold, unrelated test area, or another environmental cause.

    Uncertain when evidence is insufficient. Do not label a failure flaky merely because it might be intermittent. Narrowly reproduce it, compare recent main/scheduled CI, inspect relevant code, or otherwise gather enough evidence to choose a category.

  5. When several root failures exist, classify each one. Address PR-related failures before spending retries on unrelated cascades.

Estimate the proposed fix before editing.

A fix is eligible for automatic execution only when all are true:

  • The fix is well understood, localized, and clean.
  • The fix is not tricky, hacky, speculative, or a broad workaround.
  • The fix changes at most 100 lines, counting additions plus deletions in the fix patch, not the pre-existing PR diff.
  • The fix does not require a broad API or architecture redesign, risky dependency migration, large generated-file rewrite, or unrelated cleanup.
  • A focused validation is available.

If any condition fails, stop before editing. Report the diagnosis, proposed approach, expected files, estimated size, risks, and validation plan, then wait for user review.

For an eligible fix:

  1. Work from the latest PR head in a clean checkout or isolated worktree.
  2. Make only the necessary change and add or update focused tests when appropriate.
  3. Run the smallest meaningful local validation. For lint failures, reproduce the specific hook/check before considering broader lint validation.
  4. Review the resulting patch and count changed lines. If it becomes over 100 lines or turns tricky/dirty, do not commit or push it; stop for review and clearly identify the uncommitted skill-authored changes.
  5. Commit only skill-authored files. Never use git add ., rewrite user commits, amend without request, or include pre-existing local changes.
  6. Push a new commit non-forcibly to the exact PR head repository and branch. Never force-push. If permission or fork routing is unclear, stop and report the push command/target that needs authorization.
  7. Resolve the new PR head SHA and return to current-head monitoring. A push is progress, not completion.

Handle an unrelated or infrastructure failure

Retry only the failed jobs in the current selected-workflow run:

bash
gh run rerun <RUN_ID> --failed --repo <OWNER/REPO>

Rules:

  • Never use a full-run rerun as a substitute for --failed.
  • Never rerun a stale run after the PR head SHA changes or a newer current-head run exists.
  • Count only successfully dispatched skill-initiated reruns.
  • Allow at most 10 such reruns per selected workflow and head SHA.
  • Wait for each new attempt to finish, inspect its result, and update the failure signature before deciding on another retry.
  • If the failure changes or becomes plausibly PR-related, return to diagnosis instead of blindly consuming all retries.
  • If 10 retries are exhausted, stop. Report the attempts, recurring signatures, job/run links, and evidence that the failure is unrelated or still uncertain. Do not exceed the cap without a new explicit user instruction.

If GitHub cannot rerun only the failed portion of the selected run, stop and ask the user rather than rerunning the full workflow.

Completion check

Before completing the goal:

  1. Fetch the PR head SHA again.
  2. For every selected workflow, verify that its newest applicable run is for that exact SHA and concludes success.
  3. Verify no selected workflow is still pending and no success was invalidated by a newer push.
  4. Ignore all unselected workflows even if they are failing.
  5. Mark the goal complete only after all selected workflows satisfy these checks.

Report the final PR URL, head SHA, selected workflows, passing run links, any commits pushed, and retry counts. Keep the report scoped to the selected workflows.

© sgl-project, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .agents/skills/babysit-pr-to-pass-ci of sgl-project/sglang.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit dab108b

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in sgl-project/sglang, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Babysit PR To Pass CI next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Babysit PR To Pass CI compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Babysit PR To Pass CI this skillsgl-project/sglang37k2 repos~3kAutomated safety check: PassApache-2.0
ReviewdogAgentSecOps/SecOpsAgentKit2201 repos~3kAutomated safety check: PassCustom licence
GitHub Actions Hardeninggithub/awesome-copilot40k1 repos~2.4kAutomated safety check: PassMIT
Ghbubbuild/bub1.7k—~798Automated safety check: PassApache-2.0
Renovate Actions PR Reviewbacknotprop/plannotator9.3k—~640Automated safety check: PassApache-2.0
Update Dependenciesalorence/django-modern-rpc111—~1.3kAutomated safety check: PassMIT

Similar skills

  • Reviewdog

    AgentSecOps/SecOpsAgentKit

    Automated code review and security linting integration for CI/CD pipelines using reviewdog.

    220 GitHub starsUsed in 1 repo~3k tokens
    DevelopmentAuto-check passed
  • GitHub Actions Hardening

    github/awesome-copilot

    Official

    Security hardening reviewer for GitHub Actions workflow files (.github/workflows/.yml).

    40k GitHub starsUsed in 1 repo~2.4k tokens
    DevOps & CloudAuto-check passed
  • Gh

    bubbuild/bub

    GitHub CLI skill for interacting with GitHub via the gh command line tool.

    1.7k GitHub stars~798 tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Renovate Actions PR Review

    backnotprop/plannotator

    Reviews Renovate pull requests that bump GitHub Actions by checking pinned SHAs against upstream tags, scanning changelogs and confirming workflows stay compatible.

    9.3k GitHub stars~640 tokensUpdated today
    DevelopmentAuto-check passed
  • Update Dependencies

    alorence/django-modern-rpc

    Routine update of all project dependencies — uv itself, uv.lock (all groups), tool versions pinned in GitHub workflows and .pre-commit-config.yaml (uv, ruff, mypy...), and SHA-pinned GitHub Actions.

    111 GitHub stars~1.3k tokensUpdated 3 days ago
    DevelopmentAuto-check passed
  • Golang Continuous Integration

    samber/cc-skills-golang

    GitHub Actions CI/CD pipeline configuration for Golang projects — workflow files for test, lint, SAST, coverage and vulnerability-scan jobs, Dependabot and Renovate config files, GoReleaser release…

    3.4k GitHub stars~3.7k tokensUpdated 9 days ago
    DevelopmentAuto-check passed

More from sgl-project/sglang

All 32 skills in this repo
  • Sglang Prod Incident Triage

    sgl-project/sglang

    Replay-first debug flow for SGLang serving problems. An agent skill from sgl-project/sglang.

    37k GitHub starsUsed in 3 repos~2.1k tokens
    Auto-check passed
  • LLM Torch Profiler Analysis

    sgl-project/sglang

    Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed.

    37k GitHub starsUsed in 2 repos~6.4k tokens
    Auto-check passed
  • Compute Mamba Ratio

    sgl-project/sglang

    Compute the optimal --mamba-full-memory-ratio (or --max-mamba-cache-size pin) for a hybrid attention + linear-attention (Mamba / GDN / KDA) model's two serving memory pools, from the workload and…

    37k GitHub starsUsed in 2 repos~2.9k tokens
    Auto-check passed
  • Debug Distributed Hang

    sgl-project/sglang

    Debug hanging issues in SGLang distributed inference (TP/PP/DP/EP).

    37k GitHub starsUsed in 2 repos~2.4k tokens
    Auto-check passed
  • Env Var Conventions

    sgl-project/sglang

    Conventions for SGLang environment variables — where to define, how to access, how to name, and how to deprecate.

    37k GitHub starsUsed in 2 repos~2.9k tokens
    Auto-check passed
  • Kl Consistency Test

    sgl-project/sglang

    Write, calibrate, and debug the prefill-vs-decode logprob (KL) consistency tests in sglang -- the two independent conditions a zero requires (every operator batch-invariant, and the two paths…

    37k GitHub starsUsed in 2 repos~3.7k tokens
    Auto-check passed

Questions about Babysit PR To Pass CI

What does Babysit PR To Pass CI do?

Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head. Babysit PR To Pass CI is an agent skill from sgl-project/sglang. Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head.

When should I use Babysit PR To Pass CI?

Babysit PR To Pass CI fits situations like: asked to monitor; fix PR CI for lint.yml; pr-test-extra.yml; other named workflows.

How do I install Babysit PR To Pass CI in Claude Code?

Run `npx skills add sgl-project/sglang --skill babysit-pr-to-pass-ci -a claude-code`. Or copy the skill folder (.agents/skills/babysit-pr-to-pass-ci in sgl-project/sglang) into .claude/skills/babysit-pr-to-pass-ci in your project. Claude Code loads it when a task matches its description.

How do I install Babysit PR To Pass CI in Codex?

Run `npx skills add sgl-project/sglang --skill babysit-pr-to-pass-ci -a codex`. Or copy the skill folder (.agents/skills/babysit-pr-to-pass-ci in sgl-project/sglang) into .agents/skills/babysit-pr-to-pass-ci in your project. Codex loads it when a task matches its description.

Can I use Babysit PR To Pass CI in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add sgl-project/sglang --skill babysit-pr-to-pass-ci -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/babysit-pr-to-pass-ci, .gemini/skills/babysit-pr-to-pass-ci, .github/skills/babysit-pr-to-pass-ci and .opencode/skills/babysit-pr-to-pass-ci in your project.

What does Babysit PR To Pass CI need to run?

Going by SKILL.md and its folder, Babysit PR To Pass CI needs the command-line tools its instructions call (git and gh).

Does Babysit PR To Pass CI access the network?

SKILL.md contains no URLs. Its commands use git and gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Babysit PR To Pass CI safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Babysit PR To Pass CI use?

Babysit PR To Pass CI is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Babysit PR To Pass CI use?

About 3k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Babysit PR To Pass CI?

Skills that share tags, products or a category with Babysit PR To Pass CI: Reviewdog (AgentSecOps/SecOpsAgentKit, 220 stars), GitHub Actions Hardening (github/awesome-copilot, 40k stars), Gh (bubbuild/bub, 1.7k stars) and Renovate Actions PR Review (backnotprop/plannotator, 9.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Babysit PR To Pass CI?

sgl-project (a GitHub organization) maintains it in sgl-project/sglang, which has 36,973 GitHub stars. The repository holds 32 skills in this directory. The repository was last updated on October 11, 2026.

Source: sgl-project/sglang on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.