Kayba Pipeline
kayba-ai/agentic-context-engine
End-to-end agent evaluation and improvement pipeline. An agent skill from kayba-ai/agentic-context-engine.
bdg's shipping workflow: take verified issues from the queue to merged PRs with coding subagents (implementer, fresh reviewer, fresh-agent test before merge), with bdg's commands, gates, conventions…
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install szymdzum/browser-debugger-cli ship --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/ship .claude/skills/ship && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .claude/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/shipType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install szymdzum/browser-debugger-cli ship --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/ship .agents/skills/ship && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .agents/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install szymdzum/browser-debugger-cli ship --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/ship .cursor/skills/ship && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .cursor/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/szymdzum/browser-debugger-cli.git --path .claude/skills/ship--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install szymdzum/browser-debugger-cli ship --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/ship .gemini/skills/ship && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .gemini/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install szymdzum/browser-debugger-cli shipInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/ship .github/skills/ship && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .github/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add szymdzum/browser-debugger-cli --skill ship -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install szymdzum/browser-debugger-cli ship --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/szymdzum/browser-debugger-cli.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/ship .opencode/skills/ship && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "ship" agent skill from https://github.com/szymdzum/browser-debugger-cli/tree/main/.claude/skills/ship into .opencode/skills/ship/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ship", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
shipbdg's shipping workflow: take verified issues from the queue to merged PRs with coding subagents (implementer, fresh reviewer, fresh-agent test before merge), with bdg's commands, gates, conventions…
Ship is an agent skill from szymdzum/browser-debugger-cli. bdg's shipping workflow: take verified issues from the queue to merged PRs with coding subagents (implementer, fresh reviewer, fresh-agent test before merge), with bdg's commands, gates, conventions and fresh-agent test scenarios. Use when the user asks to work through issues, run a session or wave, ship a feature end to end, run a fresh-agent test round or exploratory sweep, or says 'jedziemy w cyklu'.
Its SKILL.md is about 4.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `briefs.md` and `scenarios.md`).
It sits in Agent Workflows, covering Agent evaluation and testing and Subagents. It works with Figma and Chrome DevTools. The repository describes itself as: Let Claude Code and other coding agents drive and debug Chrome from the shell. DOM, network, console and raw CDP as short commands, no screenshots and no MCP needed. The licence is MIT.
8 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit f82876a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
ghgitnpmnpxFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use gh, git, npm and npx, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Ship loads about 4.4k tokens when it runs. Until then it costs about 103 tokens; SKILL.md has 2,337 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from szymdzum/browser-debugger-cli at commit f82876a, republished under its MIT licence (© szymdzum). 2,337 words, ~4,351 tokens.
.claude/skills/ship/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.You are the orchestrator: you turn issues into briefs, run the subagents, hold the gates, merge and report. You don't write feature code yourself.
Each PR gets three independent looks before it lands: tests the implementer wrote first, a fresh reviewer who never saw the implementation, and a fresh-agent test in which a tester uses the change as a new user would. Skip none of them for behaviour changes.
Files in this skill:
main: a bdg invocation on a fixture page or real site, a test, or a script. It must be deterministic, or for timing bugs reproduce at a rate you can debug against (loop it, add load). Record the command and its output in the issue; it becomes the red commit's first test. Behaviour that can't be run (dead code, a wrong doc) may be confirmed with file:line instead. No command and no file:line: verify first or ask.needs repro comment on the issue.--json fields. Present options with a recommendation, then wait. A standing "pick the recommended option" never covers a contract change or a third fix round; each gets its own yes/no question.Copy one per issue into your notes and tick it off:
#<N> <slug>
- [ ] Brief sent (briefs.md → Implementer), worktree ../bdg-<N>
- [ ] Implementer report: red commit (SHA, failure line), before/after evidence, draft PR opened
- [ ] Fresh review (briefs.md → Reviewer), red commit checked → findings back to the same implementer
- [ ] Second review, only if the fix commit is large or risky
- [ ] Fresh-agent test on the branch (briefs.md → Tester), findings fixed in the PR
- [ ] Gates 1–8 hold on the final head SHA
- [ ] Merged, main CI green, worktree and branch removed, fixture server (PID file) killed, issue updated
- [ ] Every finding fixed or filedSpawn one implementer per issue with the implementer brief. Fill in scope, decisions and the files other agents are working on; the brief carries the rules (red commit first, no CHANGELOG.md edits, forbidden commands).
Implementers run in the background, so a finished one gets its review while the others still work (on the first plan, a wave of three held a ready PR for 40 minutes). Keep the user informed instead: one line at each step change (started, red commit, draft PR, report in), and git status --short of the worktree when asked. Reviewers and testers run in the foreground: they take 1–5 minutes and the next step waits on them. Every subagent gets a fresh context (general-purpose / code-reviewer, never a fork of this conversation).
An implementer's report ends with one fixed line: PR <url> HEAD <sha> CI <run id> <conclusion>. No such line means the work isn't done. If an agent stops without it, don't wait: read the PR and CI state yourself (gh pr view, gh pr checks) and resume the agent with what's missing.
When the implementer reports, spawn a fresh reviewer with the reviewer brief: the diff, the issue, and the risky areas of this change (races, failure and cleanup paths, leaks, security, contract changes, recently merged features it could break). Don't pass on the implementer's reasoning.
The highest-yield step. Run it for every behaviour change (tier 2 and 3 below). Use the tester brief with the scenarios from scenarios.md that touch the change's area.
| Tier | What | Review | Fresh-agent test | Red commit |
|---|---|---|---|---|
| 1 docs / tooling / tiny fix (no change to output, flags, exit codes or timing) | docs, CI, refactor under existing tests, a typo in a message | fresh reviewer | no | no (the PR says why) |
| 2 behaviour change | new or changed output, flag, hint, default | fresh reviewer | yes | yes |
| 3 high risk | session lifecycle, Chrome launch, files on disk, concurrency, security, contracts | fresh reviewer + second review of the fix commit | yes, plus the interrupt scenario (S08) when relevant | yes, plus repeat=10 for timing |
briefs.md) reproduces them all on current main in one pass; you group what survived into issues (one per root cause, with the command and output) in the right milestone. Don't verify 15 findings by hand, and don't file unverified ones. The verifier runs no git commands, so you prepare its checkout: git worktree add ../bdg-verify --detach origin/main && ln -s <main checkout>/node_modules ../bdg-verify/node_modules && npm --prefix ../bdg-verify run build, and remove it after the report (git worktree remove ../bdg-verify).main merged changes to files this PR touches since the branch was created (git -C ../bdg-<N> log --oneline HEAD..origin/main -- <files>), the implementer rebases before the fresh-agent test; otherwise the tester reports bugs main already fixed.npx tsx src/__testutils__/serveFixtures.ts in ../bdg-<N>), never from another worktree or the main checkout: a tester once got a page the branch had added from a server that didn't have it. Start it as npx tsx src/__testutils__/serveFixtures.ts & echo $! > /tmp/bdg-fixtures-<N>.pid and kill that PID when the tester reports (one server ran for 2.5 h after its round).When every gate below holds:
R=szymdzum/browser-debugger-cli
sha=$(gh pr view <n> --repo $R --json headRefOid -q .headRefOid) # the SHA the gates were checked on
gh pr ready <n> --repo $R
gh pr merge <n> --repo $R --merge --match-head-commit $sha --subject "<PR title> (#<n>)"
m=$(gh pr view <n> --repo $R --json mergeCommit -q .mergeCommit.oid)
gh run list --repo $R --branch main --json headSha,name,status,conclusion -q ".[]|select(.headSha==\"$m\")" # until done, macOS included
git -C ../bdg-<N> status --short # anything modified or untracked that matters? look before you remove
git worktree remove ../bdg-<N> && git branch -D <branch> # no --force; GitHub deletes the remote branchA red main is fixed before any new work. Close the issue, or comment with what's left, if the PR didn't.
Merge only when all hold. If one fails, fix it; don't negotiate it.
not implemented skeletons); the PR says which criteria are measurements or docs and how they were checked, and the reviewer ran them on it: they fail on an assertion about the issue, not a build or import error. Exempt: docs-only, pure refactors covered by existing tests, CI/tooling (the PR says which). Review-fix commits carry their test with the fix. Self-reported "it failed before" doesn't count.-f node=22); the full Node matrix only when the change touches the Node runtime or package.json engines.gh pr view <n> --json headRefOid matches what you checked; never --watch), in three classes:CI OK passed (it aggregates changes, build, quality, contract tests and Linux smoke; Security and macOS are outside it);main (gh pr view <n> --json mergeable; UNKNOWN right after a push means re-check), and tested against it: when main has merged changes to the PR's files since its last rebase, either the implementer rebases and re-runs, or you run a merge-check yourself: tree=$(git merge-tree --write-tree origin/main <branch>) (a conflict exits non-zero), then git worktree add /tmp/mc-<N> $tree, link node_modules, npm run check, the affected unit and smoke files, remove the worktree. After any rebase: typecheck and the affected tests re-run, no conflict markers.docs/CLI_REFERENCE.md, help text, .claude/skills/bdg/SKILL.md, option behaviors. CHANGELOG.md untouched; the PR description says what changed for users and marks changed defaults, contracts and breaking changes.~/.bdg or ~/Downloads (compare with the snapshot taken before the agents ran).docs/RELEASE_PROCESS.md), never at the end of a session by default.src/__testutils__/fixturePages/ exporting ROUTES, never an edit to fixtureServer.ts. Option behaviors: the area table in src/commands/optionBehaviors/<area>.ts.PR … HEAD … CI … line is not done. Read the PR and CI state yourself and resume the agent.SECRET): rename, don't suppress. An open CodeQL alert on main gets an issue the day it appears; six sat there without one until the session D retro.gh --repo szymdzum/browser-debugger-cli, default branch main.<PR title> (#N). Required check: CI OK.git push -u git@github.com:szymdzum/browser-debugger-cli.git <branch>. After a push to the URL, gh pr create needs --head <branch> (the branch isn't tracked under the origin name).export PATH="$HOME/.nvm/versions/node/v22.15.0/bin:$PATH" at the start of every shell call (no .nvmrc; the default Node may be too new).git worktree add ../bdg-<N> -b <type>/<slug> origin/main, then ln -s <main checkout>/node_modules ../bdg-<N>/node_modules. Each builds its own dist; never rebuild one other agents use.BDG_TEST_SESSION_PARENT=/tmp/bt-<N> and BDG_TEST_HOME_DIR=/tmp/bt-<N>-h, one pair per worktree: each test process gets its own session dir under the parent (removed on exit), in a short path an agent can tell apart from other agents'. Not BDG_TEST_SESSION_DIR: that is one fixed dir shared by every test process of the run (#576). Manual sessions: BDG_SESSION_DIR=/tmp/....node_modules is a symlink to the main checkout: never npm install through it; a PR that changes package-lock.json runs its own npm ci in the worktree instead.npm run check (release: npm run check:enhanced), unit npm test, build npm run build.npx tsx --test --test-concurrency=1 src/__tests__/smoke/<file>.smoke.test.ts. Integration: ./tests/run-all-tests.sh --integration.gh workflow run ci.yml --repo szymdzum/browser-debugger-cli --ref <branch> -f smoke_files='<space-separated paths>' -f repeat=10 -f node=22. node picks the one Node version of the smoke jobs (22, 24 or 26; push and nightly run all three). No brace globs, no debug=true (#510).timeout ≥ 1800000 ms). Wait for each expected workflow by name, because the PR workflows (CI, Security) are created at different moments and a loop over "all runs of the commit" can end when the first is done and the second doesn't exist yet:R=szymdzum/browser-debugger-cli; sha=<full sha> # PR head or merge commit
for wf in CI Security; do # a PR based on another PR's branch gets no Security run (it triggers on base main only): wait for CI alone
until [ "$(gh run list --repo $R --commit $sha --workflow $wf --json status -q '.[0].status')" = completed ]; do sleep 60; done
done
gh run list --repo $R --commit $sha --json name,conclusion -q '.[]|"\(.name) \(.conclusion)"'gh pr checks <n> once at the end (it lists the jobs; it doesn't wait).git stash; git commit --no-verify or any other hook bypass; broad pkill/killall/pkill -P (kill only PIDs you started).~/.bdg, other worktrees or the main checkout from an agent; leaving files in ~/Downloads.© szymdzum, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files in .claude/skills/ship of szymdzum/browser-debugger-cli.
Open the folder on GitHubat commit f82876a
Ship next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Ship this skillszymdzum/browser-debugger-cli | 173 | — | ~4.4k | Automated safety check: Pass | MIT | |
| Kayba Pipelinekayba-ai/agentic-context-engine | 2.6k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | |
| Eval Answermalloydata/publisher | 116 | — | ~4.3k | Automated safety check: Pass | MIT | |
| Skill Forge EvalAgriciDaniel/skill-forge | 179 | — | ~1.7k | Automated safety check: Pass | MIT | |
| Diagnosing Superpowers SessionsjnMetaCode/superpowers-zh | 8.3k | — | ~858 | Automated safety check: Pass | MIT | |
| Figma From Codebitovi/ai-enablement-prompts | 121 | — | ~7.4k | Automated safety check: Pass | MIT |
kayba-ai/agentic-context-engine
End-to-end agent evaluation and improvement pipeline. An agent skill from kayba-ai/agentic-context-engine.
malloydata/publisher
Score one analytical answer against a verified golden, and score which of the entities the golden depends on retrieval delivered to the answerer.
AgriciDaniel/skill-forge
Run evaluation pipelines on Claude Code skills to test triggering accuracy, workflow correctness, and output quality.
jnMetaCode/superpowers-zh
Investigates what went wrong in a superpowers session by reading its transcript, reports findings with path and line citations, and can draft a GitHub issue or redacted bundle.
bitovi/ai-enablement-prompts
Orchestrates the full code-to-Figma rebuild workflow for a web application.
bitovi/ai-enablement-prompts
Subagent for figma-from-code Phase 2.5. An agent skill from bitovi/ai-enablement-prompts.
szymdzum/browser-debugger-cli
Use bdg CLI to drive and debug a real Chrome via Chrome DevTools Protocol - navigate, click, fill and submit forms, check what an action changed (navigation, new messages, pending requests), inspect…
Works with
Categories
bdg's shipping workflow: take verified issues from the queue to merged PRs with coding subagents (implementer, fresh reviewer, fresh-agent test before merge), with bdg's commands, gates, conventions…. Ship is an agent skill from szymdzum/browser-debugger-cli. bdg's shipping workflow: take verified issues from the queue to merged PRs with coding subagents (implementer, fresh reviewer, fresh-agent test before merge), with bdg's commands, gates, conventions and fresh-agent test scenarios.
Ship fits situations like: the user asks to work through issues; ship a feature end to end; run a fresh-agent test round; exploratory sweep.
Run `npx skills add szymdzum/browser-debugger-cli --skill ship -a claude-code`. Or copy the skill folder (.claude/skills/ship in szymdzum/browser-debugger-cli) into .claude/skills/ship in your project. Claude Code loads it when a task matches its description.
Run `npx skills add szymdzum/browser-debugger-cli --skill ship -a codex`. Or copy the skill folder (.claude/skills/ship in szymdzum/browser-debugger-cli) into .agents/skills/ship in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add szymdzum/browser-debugger-cli --skill ship -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/ship, .gemini/skills/ship, .github/skills/ship and .opencode/skills/ship in your project.
Going by SKILL.md and its folder, Ship needs the command-line tools its instructions call (gh, git, npm and npx). Our summary lists: Node.js.
SKILL.md contains no URLs. Its commands use gh, git, npm and npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Ship is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 4.4k tokens (SKILL.md is roughly 17k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Ship: Kayba Pipeline (kayba-ai/agentic-context-engine, 2.6k stars), Eval Answer (malloydata/publisher, 116 stars), Skill Forge Eval (AgriciDaniel/skill-forge, 179 stars) and Diagnosing Superpowers Sessions (jnMetaCode/superpowers-zh, 8.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
szymdzum (a GitHub user) maintains it in szymdzum/browser-debugger-cli, which has 173 GitHub stars. The repository was last updated on October 10, 2026.
Source: szymdzum/browser-debugger-cli on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.