Agent skill

Skill Factory

by nyldn in nyldn/claude-octopus

Run a full build-and-ship pipeline from a spec — use for hands-off project generation

MITAuto-check passedAgent Workflows

Install Skill Factory

skills CLI
$ npx skills add nyldn/claude-octopus --skill skill-factory -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nyldn/claude-octopus skill-factory --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nyldn/claude-octopus.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/skill-factory .claude/skills/skill-factory && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
skill-factory
GitHub stars
4.2k
Used in
1 other repo
Token cost
~2k tokens
SKILL.md length
699 words
Files
2
Skills in repo
62
Repo updated
First seen
Licence
MIT

At a glance

Run a full build-and-ship pipeline from a spec — use for hands-off project generation

  • Works in 9 steps: Clarifying Questions (MANDATORY) → Display Visual Indicators (MANDATORY —… → Validate Spec File (MANDATORY —… → …
  • Hands-off project generation
  • SKILL.md covers EXECUTION CONTRACT (MANDATORY…, Error Handling (by step) and Prohibited Actions
  • Calls claude and codex

What it does

Skill Factory is an agent skill from nyldn/claude-octopus. Run a full build-and-ship pipeline from a spec — use for hands-off project generation

Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in Agent Workflows. The repository describes itself as: Run multiple AI models against the same research, design, or coding task. Surface disagreements before you ship. The licence is MIT.

When your agent uses it

  • Hands-off project generation

Example prompts

  • “/skill-factory”

Workflow steps

9 steps, taken from the step headings in SKILL.md.

  1. Clarifying Questions (MANDATORY)
  2. Display Visual Indicators (MANDATORY — BLOCKING)
  3. Validate Spec File (MANDATORY — Validation Gate)
  4. Read Prior State (OPTIONAL)
  5. 5: Adversarial Scenario Coverage Gate (RECOMMENDED)
  6. Execute orchestrate.sh factory (MANDATORY — Bash Tool)
  7. Verify Factory Report (MANDATORY — Validation Gate)
  8. Read Scores and Present Results (MANDATORY)
  9. Update State & Next Steps (MANDATORY)

What it can do on your machine

Read from SKILL.md and the folder at commit b34780d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • claude
    • codex

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Skill Factory loads about 2k tokens when it runs. Until then it costs about 25 tokens; SKILL.md has 699 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~25
When it runs · the whole SKILL.md, loaded when a task matches
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from nyldn/claude-octopus at commit b34780d, republished under its MIT licence (© nyldn). 699 words, ~2,018 tokens.

Download SKILL.mdSave it as .claude/skills/skill-factory/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
skill-factory
description
Run a full build-and-ship pipeline from a spec — use for hands-off project generation
disable-model-invocation
true

Host: Codex CLI — This skill was designed for Claude Code and adapted for Codex. Cross-reference commands use installed skill names in Codex rather than /octo:* slash commands. Use the active Codex shell and subagent tools. Do not claim a provider, model, or host subagent is available until the current session exposes it. For host tool equivalents, see skills/blocks/codex-host-adapter.md.

STOP - SKILL ALREADY LOADED

DO NOT call Skill() again. Execute directly.

EXECUTION CONTRACT (MANDATORY — 8-step sequence, CANNOT SKIP)

STEP 1: Clarifying Questions (MANDATORY)

Ask via AskUserQuestion BEFORE any other action:

  1. Spec location — path to NLSpec file, or paste inline
  2. Satisfaction target override (Use spec default / Custom 0.80-0.99)
  3. Cost approval — ~0.50-2.00 USD for ~20-30 agent calls (Approve / Approve --ci / Decline)

If spec path provided inline with the command, use it but still ask remaining questions. If user says "skip", use defaults and proceed. DO NOT proceed to Step 2 until answered.

STEP 2: Display Visual Indicators (MANDATORY — BLOCKING)

Check providers:

bash
command -v codex &> /dev/null && codex_status="Available" || codex_status="Not installed"
command -v agy &> /dev/null && agy_status="Available" || agy_status="Not installed"

Display banner:

🐙 CLAUDE OCTOPUS ACTIVATED - Dark Factory Mode
Pipeline: Parse → Scenarios → Embrace → Holdout → Score → Report

Providers:
  Codex CLI - <status> — Scenario generation + holdout evaluation
  Antigravity CLI - <status> — Additional external-model challenge
  Claude - Orchestration + satisfaction scoring

Spec: <path>
Estimated: 0.50-2.00 USD / 15-45 min

Validation: All external providers unavailable → continue with Claude-only (warn user about reduced diversity). At least one available → proceed normally.

STEP 3: Validate Spec File (MANDATORY — Validation Gate)
bash
if [[ ! -f "<spec_path>" ]]; then
    echo "ERROR: Spec file not found at <spec_path>"
    exit 1
fi
# Check minimum content
word_count=$(wc -w < "<spec_path>")
if [[ $word_count -lt 20 ]]; then
    echo "WARNING: Spec is very thin ($word_count words). Results may be limited."
fi

If spec file missing → STOP, ask user for correct path. If spec is thin (< 50 words) → WARN but proceed. DO NOT proceed past this gate if file does not exist.

STEP 4: Read Prior State (OPTIONAL)
bash
"${HOME}/.claude-octopus/plugin/scripts/state-manager.sh" init_state 2>/dev/null || true
"${HOME}/.claude-octopus/plugin/scripts/state-manager.sh" set_current_workflow "factory" "factory" 2>/dev/null || true

Failure → continue without state, warn user.

Before committing to the expensive embrace phase (~0.50-2.00 USD), verify that generated scenarios actually cover the spec's edge cases. A quick cross-provider challenge here can prevent wasting an entire factory run on incomplete scenario coverage.

After orchestrate.sh parses the spec (Phase 1-2) and before embrace execution (Phase 4), dispatch a scenario coverage review:

If a second provider is available, dispatch the challenge through Octopus routing:

bash
# Read the spec to extract behaviors and constraints
SPEC_CONTENT=$(<"<spec_path>")
review_provider=""
command -v codex >/dev/null 2>&1 && review_provider="codex"
[[ -z "$review_provider" ]] && command -v agy >/dev/null 2>&1 && review_provider="agy"

if [[ -n "$review_provider" ]]; then
  "${HOME}/.claude-octopus/plugin/scripts/orchestrate.sh" spawn "$review_provider" \
    "You are a QA adversary. Given this specification and generated scenarios, identify coverage gaps.

SPECIFICATION:
${SPEC_CONTENT}

Answer:
1. Which spec BEHAVIORS have no scenario testing them?
2. Which EDGE CASES from the spec constraints are untested?
3. Which FAILURE MODES are not covered (network failures, invalid input, concurrent access)?
4. Rate overall coverage: SUFFICIENT / GAPS-FOUND / CRITICAL-GAPS

Be specific — cite the behavior or constraint ID that lacks coverage." 2>/dev/null || true
fi

After receiving the challenge response:

  • If CRITICAL-GAPS found: warn the user and suggest refining the spec with /octo:spec before proceeding
  • If GAPS-FOUND: note the gaps but proceed — the holdout phase (Phase 5) will catch some of these
  • If SUFFICIENT: proceed with confidence

This is a lightweight gate — it adds ~30 seconds but can save a 2.00 USD factory run on a spec with poor scenario coverage.

Skip with --fast or when user explicitly requests speed over thoroughness.

STEP 5: Execute orchestrate.sh factory (MANDATORY — Bash Tool)
bash
${HOME}/.claude-octopus/plugin/scripts/orchestrate.sh factory --spec "<spec_path>"

With optional flags based on Step 1 answers:

  • --holdout-ratio <value> if custom ratio requested
  • --max-retries <value> if custom retry count
  • --ci if user approved non-interactive mode
<HARD-GATE>
PROHIBITED from:
- Running embrace directly (MUST use factory command which wraps it)
- Simulating or faking holdout testing
- Substituting direct Claude analysis for multi-provider scoring
- Skipping the factory pipeline
- Creating working/progress files in plugin directory
</HARD-GATE>
Show full SKILL.md (249 more words)Show less
STEP 6: Verify Factory Report (MANDATORY — Validation Gate)
bash
REPORT_FILE=$(find .octo/factory -name "factory-report.md" -mmin -60 2>/dev/null | sort -r | head -1)
if [[ -z "$REPORT_FILE" ]]; then
    echo "ERROR: Factory report not found"
    exit 1
fi
cat "$REPORT_FILE"

If validation fails: report error, show logs from ~/.claude-octopus/logs/, DO NOT proceed, DO NOT substitute.

STEP 7: Read Scores and Present Results (MANDATORY)
bash
SCORES_FILE=$(find .octo/factory -name "satisfaction-scores.json" -mmin -60 2>/dev/null | sort -r | head -1)
if [[ -n "$SCORES_FILE" ]]; then
    cat "$SCORES_FILE"
fi

Present to user:

  1. Verdict with emoji (PASS/WARN/FAIL)
  2. Composite score vs satisfaction target
  3. Dimension breakdown (behavior, constraints, holdout, quality)
  4. Holdout highlights — which blind scenarios passed/failed
  5. Artifact directory path for deep review
STEP 8: Update State & Next Steps (MANDATORY)
bash
"${HOME}/.claude-octopus/plugin/scripts/state-manager.sh" record_decision "factory" "Factory run completed: <verdict> (<score>/<target>)" 2>/dev/null || true

Present next-step suggestions based on verdict:

  • PASS: Implementation meets spec. Review artifacts, run manual testing, ship.
  • WARN: Close to target. Review holdout failures, consider targeted fixes.
  • FAIL: Below target. Review holdout-results.md for specific failures. Consider:
    • Refining the NLSpec with /octo:spec
    • Manual fixes + re-run with --max-retries 2
    • Breaking spec into smaller, clearer pieces

Include attribution footer:

Dark Factory Mode powered by Claude Octopus v8.25.0
Pipeline: Spec → Scenarios → Embrace → Holdout → Score → Report
Providers: Codex | Antigravity | Claude

Error Handling (by step)

  • Step 1 (Questions): If user declines all, proceed with defaults
  • Step 2 (Providers): Both external unavailable → Claude-only mode (warn)
  • Step 3 (Spec): File missing → STOP. Thin spec → WARN and proceed.
  • Step 4 (State): Failure → continue without state
  • Step 5 (orchestrate.sh): Show bash error, check logs — DO NOT substitute
  • Step 6 (Report): Missing → show logs, DO NOT proceed
  • Step 7 (Scores): Missing JSON → extract from report markdown
  • Step 8 (State): Failure → skip state update, still present results

Prohibited Actions

  • CANNOT skip orchestrate.sh factory execution
  • CANNOT simulate or fake the factory pipeline
  • CANNOT substitute direct Claude analysis for multi-provider scoring
  • CANNOT skip spec validation gate
  • CANNOT proceed past a failed validation gate
  • CANNOT create working/progress files in plugin directory

© nyldn, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/skill-factory of nyldn/claude-octopus.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit b34780d

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in nyldn/claude-octopus, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Skill Factory next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Skill Factory compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Skill Factory this skillnyldn/claude-octopus4.2k1 repos~2kAutomated safety check: PassMIT
MCP Server Builderanthropics/skills180k63 repos~2.3kAutomated safety check: PassApache-2.0
Hook Development for Claude Code Pluginsanthropics/claude-plugins-official38k10 repos~4.1kAutomated safety check: NotesApache-2.0
Using Superpowersfarm-fe/farm5.6k35 repos~1.4kAutomated safety check: PassMIT
Executing Plans Inlineobra/superpowers297k2 repos~5.1kAutomated safety check: PassMIT
Skill CreatorAzure/azqr79689 repos~8.2kAutomated safety check: PassApache-2.0

Similar skills

  • MCP Server Builder

    anthropics/skills

    Official

    Guides the design and implementation of Model Context Protocol servers in TypeScript or Python, from tool naming and error messages to evaluation.

    180k GitHub starsUsed in 63 repos~2.3k tokens
    Agent WorkflowsAuto-check passed
  • Hook Development for Claude Code Plugins

    anthropics/claude-plugins-official

    Official

    Explains how to write Claude Code plugin hooks, both prompt-based checks and bash commands, for events such as PreToolUse, Stop and SessionStart.

    38k GitHub starsUsed in 10 repos~4.1k tokens
    Agent WorkflowsAuto-check: notes
  • Using Superpowers

    farm-fe/farm

    A skill your agent uses when starting any conversation - establishes how to find and use skills, requiring Skill tool invocation before ANY response including clarifying questions

    5.6k GitHub starsUsed in 35 repos~1.4k tokens
    Agent WorkflowsAuto-check passed
  • Executing Plans Inline

    obra/superpowers

    Has the agent carry out an implementation plan itself, task by task in the current session, keeping a ledger, proving each step with a test and ending with one whole-branch review.

    297k GitHub starsUsed in 2 repos~5.1k tokens
    Agent WorkflowsAuto-check passed
  • Skill Creator

    Azure/azqr

    Official

    Create new skills, modify and improve existing skills, and measure skill performance.

    796 GitHub starsUsed in 89 repos~8.2k tokens
    Agent WorkflowsAuto-check passed
  • Claude Code Agent Development

    anthropics/claude-plugins-official

    Official

    Explains how to write agents for Claude Code plugins: the markdown file with YAML frontmatter, trigger descriptions, model and color settings, and system prompt design.

    38k GitHub starsUsed in 7 repos~2.8k tokens
    Agent WorkflowsAuto-check passed

More from nyldn/claude-octopus

All 62 skills in this repo
  • Octopus Quick

    nyldn/claude-octopus

    Quick execution for ad-hoc tasks without full workflow overhead — use for small, self-contained requests

    4.2k GitHub starsUsed in 1 repo~2.2k tokens
    Auto-check passed
  • Octopus Research

    nyldn/claude-octopus

    Thorough research across multiple sources — use for complex topics needing broad synthesis

    4.2k GitHub starsUsed in 1 repo~1.9k tokens
    Auto-check passed
  • Octopus Security Audit

    nyldn/claude-octopus

    OWASP compliance, vulnerability scanning, and adversarial red team testing — use for security reviews

    4.2k GitHub starsUsed in 1 repo~2.3k tokens
    Auto-check passed
  • Skill Audit

    nyldn/claude-octopus

    Audit codebases for quality, consistency, and broken patterns — use for pre-release or tech debt review

    4.2k GitHub starsUsed in 1 repo~3.2k tokens
    Auto-check passed
  • Skill Content Pipeline

    nyldn/claude-octopus

    Extract patterns and anatomy from URLs — use to reverse-engineer content strategies from live pages

    4.2k GitHub starsUsed in 1 repo~3.9k tokens
    Auto-check passed
  • Skill Context Detection

    nyldn/claude-octopus

    Auto-detect work context (Dev vs Knowledge) — use to tailor workflows based on current task type

    4.2k GitHub starsUsed in 1 repo~2.6k tokens
    Auto-check passed

Categories

Questions about Skill Factory

What does Skill Factory do?

Run a full build-and-ship pipeline from a spec — use for hands-off project generation. Skill Factory is an agent skill from nyldn/claude-octopus.

When should I use Skill Factory?

Skill Factory fits situations like: hands-off project generation.

How do I install Skill Factory in Claude Code?

Run `npx skills add nyldn/claude-octopus --skill skill-factory -a claude-code`. Or copy the skill folder (skills/skill-factory in nyldn/claude-octopus) into .claude/skills/skill-factory in your project. Claude Code loads it when a task matches its description.

How do I install Skill Factory in Codex?

Run `npx skills add nyldn/claude-octopus --skill skill-factory -a codex`. Or copy the skill folder (skills/skill-factory in nyldn/claude-octopus) into .agents/skills/skill-factory in your project. Codex loads it when a task matches its description.

Can I use Skill Factory in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nyldn/claude-octopus --skill skill-factory -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/skill-factory, .gemini/skills/skill-factory, .github/skills/skill-factory and .opencode/skills/skill-factory in your project.

What does Skill Factory need to run?

Going by SKILL.md and its folder, Skill Factory needs the command-line tools its instructions call (claude and codex).

Does Skill Factory access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Skill Factory safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Skill Factory use?

Skill Factory is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Skill Factory use?

About 2k tokens (SKILL.md is roughly 8.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Skill Factory?

Skills that share tags, products or a category with Skill Factory: MCP Server Builder (anthropics/skills, 180k stars), Hook Development for Claude Code Plugins (anthropics/claude-plugins-official, 38k stars), Using Superpowers (farm-fe/farm, 5.6k stars) and Executing Plans Inline (obra/superpowers, 297k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Skill Factory?

nyldn (a GitHub user) maintains it in nyldn/claude-octopus, which has 4,198 GitHub stars. The repository holds 62 skills in this directory. The repository was last updated on October 9, 2026.

Source: nyldn/claude-octopus on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.