Agent skill

Test Generation Execution

by protect-my-hair in protect-my-hair/nucleus-marketplace

Nucleus V2 Test Generation and Execution baseline. An agent skill from protect-my-hair/nucleus-marketplace.

MITAuto-check passedTesting & QA

Install Test Generation Execution

skills CLI
$ npx skills add protect-my-hair/nucleus-marketplace --skill test-generation-execution -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install protect-my-hair/nucleus-marketplace test-generation-execution --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/protect-my-hair/nucleus-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/nucleus/skills/test-generation-execution .claude/skills/test-generation-execution && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-generation-execution
GitHub stars
164
Token cost
~961 tokens
SKILL.md length
214 words
Files
10 (incl. scripts, references, assets)
Skills in repo
18
Repo updated
First seen
Licence
MIT

At a glance

Nucleus V2 Test Generation and Execution baseline. An agent skill from protect-my-hair/nucleus-marketplace.

  • Works in 8 steps: 读取测试上下文:读取… → 创建宿主任务包:把本 checklist 同步成当前会话的宿主… → 生成测试资产候选:只生成… → …
  • Tasks that involve Test generation
  • SKILL.md covers 核心原则, Checklist, 边界 and 生成测试资产, plus 2 more sections
  • Runs Python scripts from its folder; calls python3

What it does

Test Generation Execution is an agent skill from protect-my-hair/nucleus-marketplace. Nucleus V2 Test Generation and Execution baseline. Generates reviewed unit/api/integration test assets and executes approved baseline test commands with .nucleus/tests/ reports.

Its SKILL.md is about 960 tokens, which your agent loads only when the skill is triggered. The skill folder holds 13 other files, including scripts, reference files and assets (for example `agents/openai.yaml`, `assets/nucleus-test-asset-candidate.schema.json` and `assets/nucleus-test-run-report.schema.json`).

It sits in Testing & QA, covering Test generation and Third-party API integration. The repository describes itself as: Nucleus — Claude Code / Codex 工作流插件,确保 AI 编码产出结构可信、边界可审计、过程可追溯. The licence is MIT.

When your agent uses it

  • Tasks that involve Test generation
  • Tasks that involve Third-party API integration

Example prompts

  • “/test-generation-execution”

Requirements

  • Python 3

Workflow steps

8 steps, taken from the first numbered list in SKILL.md.

  1. 读取测试上下文:读取 .nucleus/context/.json、primary feature 和 feature development plan;缺关键输入时 ALERT_AND_BLOCK。
  2. 创建宿主任务包:把本 checklist 同步成当前会话的宿主 todo/task;未同步前不得生成测试资产或执行命令。
  3. 生成测试资产候选:只生成 test-asset-candidate.*,不得直接写正式 tests/**。
  4. 呈现候选并等待写入审批:向用户或 PMS 呈现 planned test paths、归属特性和变更范围;缺 test-asset-write-approval 时 STOP。
  5. 校验测试归属:确认候选和审批覆盖所有 planned test paths,并校验 Nucleus.featurePath 归属。
  6. 呈现测试命令并等待审批:说明 unit / api / integration baseline 命令、cwd 和 argv;缺 test-command 时 STOP。
  7. 执行 baseline 测试:只执行已审批且命中 allowlist 的测试命令,并写 .nucleus/tests//** report。
  8. 写入结果证据:构建 .nucleus/runs//result.json;测试报告和 result 不能替代宿主任务、人工审批或质量放行。

What it can do on your machine

Read from SKILL.md and the folder at commit 3c82085. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Generation Execution loads about 961 tokens when it runs, and up to ~1.6k if it reads all its reference files. Until then it costs about 51 tokens; SKILL.md has 214 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~51
When it runs · the whole SKILL.md, loaded when a task matches
~961
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from protect-my-hair/nucleus-marketplace at commit 3c82085, republished under its MIT licence (© protect-my-hair). 214 words, ~961 tokens.

Download SKILL.mdSave it as .claude/skills/test-generation-execution/SKILL.md (or your agent's skills folder). This skill also uses 9 other files; get the full folder from GitHub.
name
test-generation-execution
description
Nucleus V2 Test Generation and Execution baseline. Generates reviewed unit/api/integration test assets and executes approved baseline test commands with .nucleus/tests/** reports.

Test Generation and Execution

前置:使用本 Skill 前,先按 using-nucleus 完成 Nucleus 入口识别(Claude Code 会话由插件 SessionStart hook 自动注入该纪律)。

本 Skill 用于 Nucleus V2 的测试资产生成与测试执行 baseline。它安装在 Claude Code、Codex 等 coding agent 内部使用,不是 PMS 后端,也不是独立服务。

核心原则

测试资产写入和测试命令执行必须先有明确人工审批事实;approval JSON、测试报告和 result package 只能记录证据,不能替代宿主任务、人工审批或质量放行。

<HARD-GATE>
未创建宿主 todo/task 任务包,不得执行 `generate`、`run` 或进入 result package。

缺少 test-asset-write-approval 明确审批事实,不得写正式 tests/**。

缺少 test-command 明确审批事实,不得执行任何测试命令。

请求测试资产写入审批或测试命令审批前,必须先按对应 subagentPreReview 调度独立 reviewer 子代理;未取得“材料可提交人工审查”预审结论时,不得请求人工审批。

测试报告、summary 或 result 不能替代测试资产审批、测试命令审批、人工 review、缺陷创建、PR/MR 审查、合入或发布事实。 </HARD-GATE>

Checklist

启动本 Skill 后,必须先为以下每一项创建宿主 todo/task,并按顺序执行;Codex 使用计划 / 任务工具,Claude Code 使用 TodoWrite 或等价宿主 todo。每完成、阻塞、等待审批或需要复核一项,都必须逐项更新状态。

  1. 读取测试上下文:读取 .nucleus/context/<workflowRunId>.json、primary feature 和 feature development plan;缺关键输入时 ALERT_AND_BLOCK。
  2. 创建宿主任务包:把本 checklist 同步成当前会话的宿主 todo/task;未同步前不得生成测试资产或执行命令。
  3. 生成测试资产候选:只生成 test-asset-candidate.*,不得直接写正式 tests/**。
  4. 呈现候选并等待写入审批:向用户或 PMS 呈现 planned test paths、归属特性和变更范围;缺 test-asset-write-approval 时 STOP。
  5. 校验测试归属:确认候选和审批覆盖所有 planned test paths,并校验 Nucleus.featurePath 归属。
  6. 呈现测试命令并等待审批:说明 unit / api / integration baseline 命令、cwd 和 argv;缺 test-command 时 STOP。
  7. 执行 baseline 测试:只执行已审批且命中 allowlist 的测试命令,并写 .nucleus/tests/<workflowRunId>/** report。
  8. 写入结果证据:构建 .nucleus/runs/<workflowRunId>/result.json;测试报告和 result 不能替代宿主任务、人工审批或质量放行。

边界

  • 只支持 unit、api、integration 三类 baseline 测试。
  • generate 写正式 tests/** 前必须读取 test-asset-write-approval.json,且 approval 覆盖所有 planned test paths。
  • run 只读取 test-command.json 作为 approval fact,命令必须命中 baseline test runner allowlist,并使用 shell=False 执行。
  • 测试报告只写 .nucleus/tests/<workflowRunId>/**。
  • result package 只写 .nucleus/runs/<workflowRunId>/result.json。
  • 失败或阻塞时 defectArtifacts=[],不得创建缺陷或写 docs/requirement/**/defects/**。
  • 不 commit、push、创建 PR/MR、merge、release。
  • UI/E2E/Playwright 真实执行不属于 baseline。

生成测试资产

bash
python3 skills/test-generation-execution/scripts/test_generation_execution.py generate \
  --repo-root <target-repo> \
  --context <target-repo>/.nucleus/context/<workflowRunId>.json \
  --feature-development-plan <target-repo>/.nucleus/runs/<workflowRunId>/feature-development-plan.json \
  --approval <target-repo>/.nucleus/runs/<workflowRunId>/test-asset-write-approval.json \
  --candidate <target-repo>/.nucleus/runs/<workflowRunId>/test-asset-candidate.json

generate 会校验唯一 primary feature、feature development plan、approval fact 和 plannedWrites[],只写已声明且已审批的 tests/** 文件。每个生成或更新的测试文件前 40 行必须包含:

text
Nucleus.featurePath: docs/features/.../feature.md

执行测试命令

bash
python3 skills/test-generation-execution/scripts/test_generation_execution.py run \
  --repo-root <target-repo> \
  --context <target-repo>/.nucleus/context/<workflowRunId>.json \
  --level unit \
  --command-json <target-repo>/.nucleus/runs/<workflowRunId>/test-command.json

test-command.json 必须包含:

  • approved=true
  • approvalType=test-command
  • approvalFactId
  • workflowRunId
  • featurePath
  • level
  • sourceFeatureDevelopmentPlanPath
  • argv
  • cwd

输出

  • .nucleus/runs/<workflowRunId>/test-asset-candidate.json
  • .nucleus/runs/<workflowRunId>/test-asset-candidate.md
  • .nucleus/runs/<workflowRunId>/test-attribution-gate.json
  • .nucleus/tests/<workflowRunId>/<level>-report.json
  • .nucleus/runs/<workflowRunId>/result.json

成功态为 NEEDS_HUMAN_REVIEW。失败或证据不足为 FAILED_BLOCKED。

© protect-my-hair, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 9 other files (scripts, references, assets) in plugins/nucleus/skills/test-generation-execution of protect-my-hair/nucleus-marketplace.

  • SKILL.md
  • agents/openai.yaml
  • assets/nucleus-test-asset-candidate.schema.json
  • assets/nucleus-test-run-report.schema.json
  • assets/test-asset-candidate-template.json
  • assets/test-asset-candidate-template.md
  • references/blocker-matrix.md
  • references/report-schema.md
  • references/test-boundary.md
  • scripts/test_generation_execution.py

Open the folder on GitHubat commit 3c82085

Compare with similar skills

Test Generation Execution next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Generation Execution compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Generation Execution this skillprotect-my-hair/nucleus-marketplace164—~961Automated safety check: PassMIT
Write Testsryokun6/ryos1.3k—~1.7kAutomated safety check: PassAGPL-3.0
Emcaklofas/kicad-happy1.4k1 repos~2.8kAutomated safety check: PassMIT
Swig Testswig/swig6.3k—~2.3kAutomated safety check: PassCustom licence
Generate Test Cases342164796/generate-test-cases1201 repos~2.9kAutomated safety check: PassNone
Wioworkersio/skills204—~5.8kAutomated safety check: PassMIT

Similar skills

  • Write Tests

    ryokun6/ryos

    Write and run ryOS tests with Bun's native test runner (bun:test).

    1.3k GitHub stars~1.7k tokensUpdated today
    Testing & QAAuto-check passed
  • Emc

    aklofas/kicad-happy

    EMC pre-compliance risk analysis for KiCad PCB designs — 18 check categories, 44 rule IDs covering ground planes, decoupling, I/O filtering, switching harmonics, clock routing, differential pair…

    1.4k GitHub starsUsed in 1 repo~2.8k tokens
    Testing & QAAuto-check passed
  • Swig Test

    swig/swig

    Run SWIG test suite for specific languages. An agent skill from swig/swig.

    6.3k GitHub stars~2.3k tokensUpdated 4 days ago
    Testing & QAAuto-check passed
  • Generate Test Cases

    342164796/generate-test-cases

    自主学习型测试文档生成器。从需求文档(Markdown)生成测试用例 XMind 文件,支持持久化记忆和持续学习。当用户提到"生成测试用例"、"根据需求生成测试"时触发。

    120 GitHub starsUsed in 1 repo~2.9k tokens
    Testing & QAAuto-check passed
  • Wio

    workersio/skills

    Testing workflow skill for finding high-value test candidates, writing focused tests, generating realistic workloads, reviewing test value, and diagnosing test-suite health.

    204 GitHub stars~5.8k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed
  • Verify Cc Safety Net

    kenryu42/cc-safety-net

    Launch and drive the real cc-safety-net CLI — the hook decision path, explain, status/doctor, logs, and the local policy GUI — against an isolated home, capturing evidence.

    1.6k GitHub stars~2k tokensUpdated today
    Testing & QAAuto-check passed

More from protect-my-hair/nucleus-marketplace

All 18 skills in this repo
  • Nucleus Delivery Evidence Closure

    protect-my-hair/nucleus-marketplace

    Collects prior Nucleus task evidence into a candidate-only delivery closure package and PR/MR draft text, blocking if any feature's design, review or verification evidence is missing.

    164 GitHub stars~1k tokensUpdated 2 mo ago
    Auto-check passed
  • Nucleus Context Validator

    protect-my-hair/nucleus-marketplace

    Validates the context file for a Nucleus V2 workflow run, discovers standalone context from explicit input, and writes a blocked result when required context is missing.

    164 GitHub stars~576 tokensUpdated 2 mo ago
    Auto-check passed
  • Nucleus Session Binding Report

    protect-my-hair/nucleus-marketplace

    Collects, freezes and validates a Nucleus V2 session binding report that records the native Claude Code or Codex session ID, transcript path and hash.

    164 GitHub stars~477 tokensUpdated 2 mo ago
    Auto-check passed
  • Nucleus Result Package Builder

    protect-my-hair/nucleus-marketplace

    Builds the result.json file for a Nucleus V2 workflow run from its context, gates, blockers and step status, and exits non-zero when the run is blocked.

    164 GitHub stars~414 tokensUpdated 2 mo ago
    Auto-check passed
  • Defect Intake and RCA Candidates

    protect-my-hair/nucleus-marketplace

    Checks whether a defect report has enough context to be accepted, then drafts candidate-only root-cause evidence for human review without confirming causes or changing code.

    164 GitHub stars~745 tokensUpdated 2 mo ago
    Auto-check passed
  • Nucleus Delivery Evidence Finalizer

    protect-my-hair/nucleus-marketplace

    Combines the evidence from completed Nucleus workflow steps into a final evidence report and result package, without scheduling steps, updating trackers or opening pull requests.

    164 GitHub stars~454 tokensUpdated 2 mo ago
    Auto-check passed

Categories

Questions about Test Generation Execution

What does Test Generation Execution do?

Nucleus V2 Test Generation and Execution baseline. An agent skill from protect-my-hair/nucleus-marketplace. Test Generation Execution is an agent skill from protect-my-hair/nucleus-marketplace. Nucleus V2 Test Generation and Execution baseline.

When should I use Test Generation Execution?

Test Generation Execution fits situations like: tasks that involve Test generation; tasks that involve Third-party API integration.

How do I install Test Generation Execution in Claude Code?

Run `npx skills add protect-my-hair/nucleus-marketplace --skill test-generation-execution -a claude-code`. Or copy the skill folder (plugins/nucleus/skills/test-generation-execution in protect-my-hair/nucleus-marketplace) into .claude/skills/test-generation-execution in your project. Claude Code loads it when a task matches its description.

How do I install Test Generation Execution in Codex?

Run `npx skills add protect-my-hair/nucleus-marketplace --skill test-generation-execution -a codex`. Or copy the skill folder (plugins/nucleus/skills/test-generation-execution in protect-my-hair/nucleus-marketplace) into .agents/skills/test-generation-execution in your project. Codex loads it when a task matches its description.

Can I use Test Generation Execution in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add protect-my-hair/nucleus-marketplace --skill test-generation-execution -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-generation-execution, .gemini/skills/test-generation-execution, .github/skills/test-generation-execution and .opencode/skills/test-generation-execution in your project.

What does Test Generation Execution need to run?

Going by SKILL.md and its folder, Test Generation Execution needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Test Generation Execution access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Test Generation Execution safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Test Generation Execution use?

Test Generation Execution is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Generation Execution use?

About 961 tokens (SKILL.md is roughly 3.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 612 tokens, read only when the agent opens those files.

What are the alternatives to Test Generation Execution?

Skills that share tags, products or a category with Test Generation Execution: Write Tests (ryokun6/ryos, 1.3k stars), Emc (aklofas/kicad-happy, 1.4k stars), Swig Test (swig/swig, 6.3k stars) and Generate Test Cases (342164796/generate-test-cases, 120 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Generation Execution?

protect-my-hair (a GitHub user) maintains it in protect-my-hair/nucleus-marketplace, which has 164 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on July 23, 2026.

Source: protect-my-hair/nucleus-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.