Agent skill

Build Loop Codex

by BuildGreatProducts in BuildGreatProducts/builder-os

A skill your agent uses when building features with Codex (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop.

MITAuto-check passedTesting & QA

Install Build Loop Codex

skills CLI
$ npx skills add BuildGreatProducts/builder-os --skill build-loop-codex -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install BuildGreatProducts/builder-os build-loop-codex --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/BuildGreatProducts/builder-os.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/build-loop-codex .claude/skills/build-loop-codex && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
build-loop-codex
GitHub stars
227
Token cost
~888 tokens
SKILL.md length
436 words
Files
1
Skills in repo
8
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when building features with Codex (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop.

  • Works in 6 steps: Build. Implement exactly what the task… → Review. Run /review and select "Review… → Test end to end. Run the task's… → …
  • Run the build loop
  • SKILL.md covers Source of work, The loop and Rules
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Build Loop Codex is an agent skill from BuildGreatProducts/builder-os. Use when building features with Codex (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop. Triggers on "run the build loop", "build the next task", "continue the plan", "build this feature properly", or any request to implement work from a plan file or a direct feature prompt. Builds from the plan (or the prompt if no plan exists), runs Codex's /review on uncommitted changes and fixes every issue found, tests and verifies the feature end to end, fixes…

Its SKILL.md is about 890 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering End-to-end testing. The repository describes itself as: BuilderOS is your operating system for building with AI. The licence is MIT.

When your agent uses it

  • Run the build loop
  • Build the next task
  • Continue the plan
  • Build this feature properly

Example prompts

  • “run the build loop”
  • “build the next task”
  • “continue the plan”
  • “/build-loop-codex”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Build. Implement exactly what the task specifies. Simplest implementation that satisfies it, surgical changes, no speculative scope. Match…
  2. Review. Run /review and select "Review uncommitted changes". If the change touches auth, payments, user input, or data access, run a…
  3. Test end to end. Run the task's verification step (or the success criteria). Run the full test suite — everything that passed before must…
  4. Fix. Anything testing finds goes back through the loop: fix → /review → re-test. Never mark a failing task complete; never start the next…
  5. Continue. Mark the task - [x], update any progress/status line in the plan, and loop to the next task until the requested scope is complete.
  6. Report. When done, tell the user: what was built and plan progress, review findings fixed and anything deferred, how it was verified…

What it can do on your machine

Read from SKILL.md and the folder at commit fb74cac. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Build Loop Codex loads about 888 tokens when it runs. Until then it costs about 161 tokens; SKILL.md has 436 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~161
When it runs · the whole SKILL.md, loaded when a task matches
~888

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from BuildGreatProducts/builder-os at commit fb74cac, republished under its MIT licence (© BuildGreatProducts). 436 words, ~888 tokens.

Download SKILL.mdSave it as .claude/skills/build-loop-codex/SKILL.md (or your agent's skills folder).
name
build-loop-codex
description
Use when building features with **Codex** (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop. Triggers on "run the build loop", "build the next task", "continue the plan", "build this feature properly", or any request to implement work from a plan file or a direct feature prompt. Builds from the plan (or the prompt if no plan exists), runs Codex's `/review` on uncommitted changes and fixes every issue found, tests and verifies the feature end to end, fixes anything testing surfaces, and reports back once complete. Repeats until all plan tasks are checked off.
license
MIT
metadata.author
BuilderOS
metadata.version
1.0

Codex Build Loop

Quality-gated feature work: nothing ships on "it compiles" — every increment is built, reviewed, tested end to end, and fixed before the user hears "done."

Source of work

  • A plan file exists (roadmap, refactor plan, or task list with - [ ] checkboxes — search the repo): work the first unchecked task. Tasks are ordered intentionally — never skip ahead. If the plan references spec docs, read only the sections relevant to the current task.
  • No plan (or the request is outside it): build from the user's prompt. Restate it as a verifiable goal with 2–4 success criteria and confirm scope in one message before building.

The loop

Run per task (or per prompted feature). Do not advance until every step passes.

  1. Build. Implement exactly what the task specifies. Simplest implementation that satisfies it, surgical changes, no speculative scope. Match existing project conventions.

  2. Review. Run /review and select "Review uncommitted changes". If the change touches auth, payments, user input, or data access, run a second pass via "Custom review instructions" (e.g. "Focus on security vulnerabilities and unvalidated input"). Fix all findings in scope — bugs, security issues, edge cases, performance, style in files you touched. If the project has a design system spec (design tokens file, DESIGN.md, theme config), check UI changes against it — no hardcoded colors, type, or spacing that bypass tokens. Note pre-existing issues in untouched code for the report instead of fixing silently. Re-run /review until clean. If a finding contradicts the task or spec, the spec wins — flag the disagreement.

  3. Test end to end. Run the task's verification step (or the success criteria). Run the full test suite — everything that passed before must still pass. Add tests for new logic. Then exercise the feature as a user would: run the app, walk the real flow including empty, loading, and error states.

  4. Fix. Anything testing finds goes back through the loop: fix → /review → re-test. Never mark a failing task complete; never start the next task with the app broken.

  5. Continue. Mark the task - [x], update any progress/status line in the plan, and loop to the next task until the requested scope is complete.

  6. Report. When done, tell the user: what was built and plan progress, review findings fixed and anything deferred, how it was verified (tests + flow walked), and what needs their attention next. Be honest about anything flaky or partially verified.

Show full SKILL.md (39 more words)Show less

Rules

  • Skipped review or untested work = unfinished work.
  • Don't relitigate plan decisions; if a task seems wrong, ask one specific question rather than guessing.
  • Discovered work no task covers? Surface it and propose a task — never silently expand scope.

© BuildGreatProducts, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/build-loop-codex of BuildGreatProducts/builder-os.

Open the folder on GitHubat commit fb74cac

Compare with similar skills

Build Loop Codex next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Build Loop Codex compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Build Loop Codex this skillBuildGreatProducts/builder-os227—~888Automated safety check: PassMIT
Web Application Testinganthropics/skills180k51 repos~966Automated safety check: PassApache-2.0
TDD WorkflowhellangleZ/burn-in-cceverywhere-ralph11211 repos~2.4kAutomated safety check: PassNone
Uloop Replay Inputkurotu/VRCQuestTools3733 repos~615Automated safety check: PassMIT
Ui4 Convert Testspayloadcms/payload45k—~3.5kAutomated safety check: PassMIT
E2Estackia/rtp2httpd2.2k—~517Automated safety check: PassGPL-2.0

Similar skills

  • Web Application Testing

    anthropics/skills

    Official

    Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

    180k GitHub starsUsed in 51 repos~966 tokens
    Testing & QAAuto-check passed
  • TDD Workflow

    hellangleZ/burn-in-cceverywhere-ralph

    A skill your agent uses when writing new features, fixing bugs, or refactoring code.

    112 GitHub starsUsed in 11 repos~2.4k tokens
    Testing & QAAuto-check passed
  • Uloop Replay Input

    kurotu/VRCQuestTools

    Replay recorded PlayMode keyboard and mouse input. An agent skill from kurotu/VRCQuestTools.

    373 GitHub starsUsed in 3 repos~615 tokens
    Testing & QAAuto-check passed
  • Ui4 Convert Tests

    payloadcms/payload

    A skill your agent uses when UI changes are complete and e2e tests need updating.

    45k GitHub stars~3.5k tokensUpdated today
    Testing & QAAuto-check passed
  • E2E

    stackia/rtp2httpd

    Write, run, review, or debug rtp2httpd E2E tests and their harness in e2e/ and scripts/run-e2e.sh.

    2.2k GitHub stars~517 tokensUpdated 5 days ago
    Testing & QAAuto-check passed
  • Moav E2E

    MotherofallVPNs/MoaV

    Run and debug MoaV's end-to-end tests — real protocol connectivity (client-test.sh) and the moav CLI smoke test — against a LIVE server, via the self-hosted e2e workflow or a local test VPS.

    448 GitHub stars~1.9k tokensUpdated today
    Testing & QAAuto-check: notes

More from BuildGreatProducts/builder-os

All 8 skills in this repo
  • Design System

    BuildGreatProducts/builder-os

    Translates an image (or a set of image references — screenshots, mockups, Figma URLs, live websites) into two mirrored design-system artifacts: docs/design.md (YAML tokens + prose, following…

    227 GitHub stars~4.8k tokensUpdated 3 mo ago
    Auto-check passed
  • Idea Generator

    BuildGreatProducts/builder-os

    Guided discovery of a product idea by mining what the founder already knows or already does — covers source selection (business vs.

    227 GitHub stars~3.1k tokensUpdated 3 mo ago
    Auto-check passed
  • Idea Validator

    BuildGreatProducts/builder-os

    Pressure-tests a product idea before the founder invests in planning, building, or launching.

    227 GitHub stars~3.8k tokensUpdated 3 mo ago
    Auto-check passed
  • Product Planner

    BuildGreatProducts/builder-os

    Vision intake conversation followed by generation of three product documents — docs/product-vision.md (strategy and brand), docs/prd.md (technical spec for coding agents), and…

    227 GitHub stars~4.6k tokensUpdated 3 mo ago
    Auto-check passed
  • Build Loop Cursor

    BuildGreatProducts/builder-os

    A skill your agent uses when building features with Cursor in any codebase and the work should go through a disciplined build → review → test → fix loop.

    227 GitHub stars~827 tokensUpdated 3 mo ago
    Auto-check passed
  • Build Mvp

    BuildGreatProducts/builder-os

    Use inside a product repository when the user wants the full MVP built from their BuilderOS spec documents.

    227 GitHub stars~1.2k tokensUpdated 3 mo ago
    Auto-check: notes

Categories

Questions about Build Loop Codex

What does Build Loop Codex do?

A skill your agent uses when building features with Codex (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop. Build Loop Codex is an agent skill from BuildGreatProducts/builder-os. Use when building features with Codex (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop.

When should I use Build Loop Codex?

Build Loop Codex fits situations like: run the build loop; build the next task; continue the plan; build this feature properly.

How do I install Build Loop Codex in Claude Code?

Run `npx skills add BuildGreatProducts/builder-os --skill build-loop-codex -a claude-code`. Or copy the skill folder (skills/build-loop-codex in BuildGreatProducts/builder-os) into .claude/skills/build-loop-codex in your project. Claude Code loads it when a task matches its description.

How do I install Build Loop Codex in Codex?

Run `npx skills add BuildGreatProducts/builder-os --skill build-loop-codex -a codex`. Or copy the skill folder (skills/build-loop-codex in BuildGreatProducts/builder-os) into .agents/skills/build-loop-codex in your project. Codex loads it when a task matches its description.

Can I use Build Loop Codex in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add BuildGreatProducts/builder-os --skill build-loop-codex -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/build-loop-codex, .gemini/skills/build-loop-codex, .github/skills/build-loop-codex and .opencode/skills/build-loop-codex in your project.

What does Build Loop Codex need to run?

SKILL.md names no scripts, command-line tools or credentials: Build Loop Codex is instructions for the agent only.

Does Build Loop Codex access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Build Loop Codex safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Build Loop Codex use?

Build Loop Codex is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Build Loop Codex use?

About 888 tokens (SKILL.md is roughly 3.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Build Loop Codex?

Skills that share tags, products or a category with Build Loop Codex: Web Application Testing (anthropics/skills, 180k stars), TDD Workflow (hellangleZ/burn-in-cceverywhere-ralph, 112 stars), Uloop Replay Input (kurotu/VRCQuestTools, 373 stars) and Ui4 Convert Tests (payloadcms/payload, 45k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Build Loop Codex?

BuildGreatProducts (a GitHub user) maintains it in BuildGreatProducts/builder-os, which has 227 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on July 7, 2026.

Source: BuildGreatProducts/builder-os on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.