Execute implementation plans from .artifacts/plan/. An agent skill from alchemiststudiosDOTai/harness-engineering.

MITAuto-check passedAgent Workflows

Install Execute Phase

skills CLI
$ npx skills add alchemiststudiosDOTai/harness-engineering --skill execute-phase -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install alchemiststudiosDOTai/harness-engineering execute-phase --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/alchemiststudiosDOTai/harness-engineering.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/execute-phase .claude/skills/execute-phase && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
execute-phase
GitHub stars
105
Token cost
~1.8k tokens
SKILL.md length
400 words
Files
1
Skills in repo
8
Repo updated
First seen
Licence
MIT

At a glance

Execute implementation plans from .artifacts/plan/. An agent skill from alchemiststudiosDOTai/harness-engineering.

  • Works in 7 steps: Input → Read Plan & Lock Context → Pre-Flight Snapshot → …
  • The user says execute this plan
  • SKILL.md covers Overview, North Star Rule, When to Use and What This Skill Does NOT Do, plus 7 more sections
  • Calls git, pytest and mypy

What it does

Execute Phase is an agent skill from alchemiststudiosDOTai/harness-engineering. Execute implementation plans from .artifacts/plan/. Focus on EXECUTING ONLY - no planning, no fixes outside plan scope. Uses gated checks, atomic commits, and maintains a single execution log in .artifacts/execute/. Use when the user says "execute this plan" or provides a plan path.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Agent Workflows, covering Planning. The repository describes itself as: harness-engineering discussion of shortcuts, automation, hacks and overall productivity with code agents like claude code, codex, and other harness. The licence is MIT.

When your agent uses it

  • The user says execute this plan
  • Provides a plan path

Example prompts

  • “execute this plan”
  • “/execute-phase”

Requirements

  • Docker
  • Pre-approved tools (allowed-tools): Edit, Read, Write, Bash(git:*), Bash(python:*), Bash(pytest:*), Bash(mypy:*), Bash(black:*), Bash(coverage:*), Bash(docker:*), Bash(trivy:*), Bash(hadolint:*), Bash(jq:*), Bash(curl:*), Bash(gh:*)

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Input
  2. Read Plan & Lock Context
  3. Pre-Flight Snapshot
  4. Task-by-Task Execution
  5. Quality Gates
  6. Permalinks & Artifacts
  7. Post-Deploy Verification (if applicable)

What it can do on your machine

Read from SKILL.md and the folder at commit 7a9fa15. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Edit
    • Read
    • Write
    • Bash(git:*)
    • Bash(python:*)
    • Bash(pytest:*)
    • Bash(mypy:*)
    • Bash(black:*)
    • Bash(coverage:*)
    • Bash(docker:*)

    …and 5 more on the same allowed-tools line.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git
    • pytest
    • mypy
    • black
    • gh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use git and gh, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Execute Phase loads about 1.8k tokens when it runs. Until then it costs about 74 tokens; SKILL.md has 400 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~74
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from alchemiststudiosDOTai/harness-engineering at commit 7a9fa15, republished under its MIT licence (© alchemiststudiosDOTai). 400 words, ~1,798 tokens.

Download SKILL.mdSave it as .claude/skills/execute-phase/SKILL.md (or your agent's skills folder).
name
execute-phase
description
Execute implementation plans from .artifacts/plan/. Focus on EXECUTING ONLY - no planning, no fixes outside plan scope. Uses gated checks, atomic commits, and maintains a single execution log in .artifacts/execute/. Use when the user says "execute this plan" or provides a plan path.
allowed-tools
Edit, Read, Write, Bash(git:*), Bash(python:*), Bash(pytest:*), Bash(mypy:*), Bash(black:*), Bash(coverage:*), Bash(docker:*), Bash(trivy:*), Bash(hadolint:*), Bash(jq:*), Bash(curl:*), Bash(gh:*)
writes-to
.artifacts/execute/
hard-guards
NO planning - follow the plan exactly, NO fixes outside plan scope, Create rollback point before starting, Keep ONE execution log updated as you work, Commit…

Execute Phase

Overview

Execute implementation plans from .artifacts/plan/ with strict discipline: gated checks, atomic commits, and a single living execution log.

North Star Rule

Follow the plan exactly. Do not improvise. Do not fix what isn't in the plan.

If something is ambiguous or missing, stop and ask the user.

When to Use

  • User provides a plan path in .artifacts/plan/
  • User says "execute this plan" or "run the plan"
  • User references a plan document for implementation

What This Skill Does NOT Do

❌ DON'T✅ DO INSTEAD
Re-plan the workExecute tasks as written
Fix unrelated issuesFollow plan scope only
Skip quality gatesRun all gates, document failures
Ignore the planPlan is the source of truth

Execute Phase Workflow

0. Input

User provides: $ARGUMENTS = path to plan in .artifacts/plan/

1. Read Plan & Lock Context

Read the FULL plan. Extract:

  • Milestones
  • Task IDs and order
  • Acceptance tests
  • Quality gates
  • Success criteria
2. Pre-Flight Snapshot

Before any code changes:

bash
# Capture git state
BRANCH=$(git branch --show-current)
SHA=$(git rev-parse --short HEAD)
STATUS=$(git status --short)

Create rollback point:

bash
git add -A
git commit -m "rollback: before executing plan <topic>"

Create .artifacts/execute/YYYY-MM-DD_HH-MM-SS_<topic>.md:

markdown
---
title: "<topic> execution log"
link: "<topic>-execute"
type: debug_history
ontological_relations:
  - relates_to: [[<plan-link>]]
tags: [execute, <topic>]
uuid: "<uuid>"
created_at: "<ISO-8601 timestamp>"
owner: "{{user}}"
plan_path: ".artifacts/plan/<file>.md"
start_commit: "<short_sha>"
env: {target: "local|staging|prod", notes: ""}
---

## Pre-Flight Checks
- Branch: <branch>
- Rollback commit: <sha>
- DoR satisfied: yes/no
- Access/secrets: present/missing
- Fixtures/data: ready/not ready

[If any NO → abort and add Blockers section]
3. Task-by-Task Execution

For EACH task in plan order:

1. Read task requirements
2. Implement minimal slice aligned with acceptance test
3. Run local validation
4. Commit atomic change with Task ID in message
5. Update execution log

Commit message format:

T<NNN>: <task summary>

<brief description of change>

Refs: plan/<file>.md
4. Quality Gates

Run gates in order. Document ALL results.

Gate C - Code Quality
bash
# Run tests
pytest

# Type check
mypy src/

# Lint
black --check src/

# Coverage
coverage report

Document in the execution log:

### Gate Results
- Tests: pass/fail + evidence
- Coverage: X% (target Y%)
- Type checks: pass/fail
- Linters: pass/fail

If gate FAILS:

  • Record failure + remediation attempted
  • STOP and ask user for next steps
  • Do NOT roll back without user confirmation

If commits pushed:

bash
# Get repo info for permalinks
gh repo view --json owner,name

Attach permalinks to:

  • PRs/commits
  • Build logs
  • Coverage reports
  • Security scans

Persist artifact pointers in the execution log.

6. Post-Deploy Verification (if applicable)
### Post-Deploy Verification
- Error rates: <metrics>
- Latencies: <metrics>
- Dashboard links: <URLs>
- Smoke/E2E results: <pass/fail>
Show full SKILL.md (158 more words)Show less

Execution Log Template

Keep ONE document. Update as you work.

markdown
---
title: "<topic> execution log"
link: "<topic>-execute"
type: debug_history
ontological_relations:
  - relates_to: [[<plan-link>]]
tags: [execute, <topic>]
uuid: "<uuid>"
created_at: "<ISO-8601 timestamp>"
plan_path: ".artifacts/plan/<file>.md"
start_commit: "<sha>"
end_commit: "<sha>"
env: {target: "...", notes: "..."}
---

## Pre-Flight Checks
- Branch: <branch>
- Rollback: <commit_sha>
- DoR: satisfied/not
- Ready: yes/no

## Task Execution

### T001 – <Summary>
- Status: completed/skipped/failed
- Commit: <sha>
- Files: <list>
- Commands: <cmd> → <output>
- Tests: pass/fail
- Coverage delta: +X%
- Notes: <decisions made>

### T002 – <Summary>
[... repeat for each task ...]

## Gate Results
- Tests: X/Y passed
- Coverage: X% (target Y%)
- Type checks: pass/fail
- Security: # issues
- Linters: pass/fail

## Deployment (if applicable)
- Staging: success/fail
- Prod: success/fail
- Timestamps: <start> → <end>

## Issues & Resolutions
- T<NNN> – <issue> → <resolution|rollback|asked user>

## Success Criteria
- [ ] All planned gates passed
- [ ] Rollout completed or rolled back
- [ ] KPIs/SLOs within thresholds
- [ ] Execution log saved

## Next Steps
- Follow-ups, tech debt, docs

Final Report

After the Execute phase, summarize:

markdown
# Execution Report – <topic>

**Date:** {{date}}
**Plan:** <plan_file>
**Log:** <log_file>

## Overview
- Environment: <env>
- Start: <sha>
- End: <sha>
- Duration: Xh Ym
- Branch: <branch>

## Outcomes
- Tasks attempted: N
- Tasks completed: N
- Final status: Success | Failure | Blocked

## Gate Results
- Tests: pass/fail
- Coverage: X% (target Y%)
- Type checks: pass/fail
- Security: # issues

## What Was Touched
[List all files modified]

## Next Steps
- [ ] Item 1
- [ ] Item 2

Strict Rules

  1. ONE execution log - Create once, update as you work. Do not create multiple docs.
  2. Atomic commits - One commit per task. Task ID in commit message.
  3. Rollback first - Create rollback commit before any code changes.
  4. Gates are mandatory - Run all gates. Document failures. Stop on failure.
  5. Plan is source of truth - Do not add, remove, or change tasks.
  6. Ask when blocked - If gates fail or plan is ambiguous, stop and ask.

Validation Questions

Before proceeding with each task:

  1. Clarity: Do I understand exactly what this task requires?
  2. Scope: Is this within the plan's scope?
  3. Dependencies: Are prerequisite tasks completed?
  4. Rollback: Can I revert to the safe state?

Output Format

Start:

Executing plan: .artifacts/plan/<file>.md
Branch: <branch>
Rollback point: <commit_sha>
Tasks: N
Milestones: M

End:

Execution complete: Success | Failure | Blocked
Tasks completed: N/N
Log: .artifacts/execute/<file>.md
Next step: QA from execute using the generated execution log path

Handoff

After writing the execution log to .artifacts/execute/, proceed to qa-from-execute if the next step is the QA phase.

© alchemiststudiosDOTai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/execute-phase of alchemiststudiosDOTai/harness-engineering.

Open the folder on GitHubat commit 7a9fa15

Compare with similar skills

Execute Phase next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Execute Phase compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Execute Phase this skillalchemiststudiosDOTai/harness-engineering105—~1.8kAutomated safety check: PassMIT
Improvefossasia/eventyay-interpretation1.6k10 repos~3.7kAutomated safety check: WarnMIT
Workflow Orchestrationvxcozy/workflow-orchestration116—~1kAutomated safety check: PassMIT
Deep Planpiercelamb/deep-plan101—~4.8kAutomated safety check: PassMIT
Plan Py4vaspvasp-dev/py4vasp100—~2.3kAutomated safety check: PassApache-2.0
Test-First Implementation Plangittower/git-flow-next458—~1.3kAutomated safety check: NotesCustom licence

Similar skills

  • Improve

    fossasia/eventyay-interpretation

    Survey any codebase as a senior advisor and produce prioritized, self-contained implementation plans for OTHER models/agents to execute.

    1.6k GitHub starsUsed in 10 repos~3.7k tokens
    Agent WorkflowsAuto-check: warnings
  • Workflow Orchestration

    vxcozy/workflow-orchestration

    Disciplined task execution with planning, verification, and self-improvement loops.

    116 GitHub stars~1k tokensUpdated 5 mo ago
    Agent WorkflowsAuto-check passed
  • Deep Plan

    piercelamb/deep-plan

    Creates detailed, sectionized, TDD-oriented implementation plans through research, stakeholder interviews, and multi-LLM review.

    101 GitHub stars~4.8k tokensUpdated 3 mo ago
    Agent WorkflowsAuto-check passed
  • Plan Py4vasp

    vasp-dev/py4vasp

    Plan a py4vasp change as an ordered list of test-first chunks — that chunk list is the plan.

    100 GitHub stars~2.3k tokensUpdated today
    Agent WorkflowsAuto-check passed
  • Test-First Implementation Plan

    gittower/git-flow-next

    Builds a two-phase implementation plan from a spec issue, analysis or concept, writing a detailed test plan first and the implementation outline second.

    458 GitHub stars~1.3k tokensUpdated 1 mo ago
    Agent WorkflowsAuto-check: notes
  • Implementation Plan

    POW-Software/ByteSync

    Produce a structured implementation plan (no code) from an existing specification or deep analysis.

    172 GitHub stars~391 tokensUpdated 5 mo ago
    Agent WorkflowsAuto-check passed

More from alchemiststudiosDOTai/harness-engineering

All 8 skills in this repo
  • Ast Grep Setup

    alchemiststudiosDOTai/harness-engineering

    Set up ast-grep for a codebase with common TypeScript rules for detecting anti-patterns, enforcing best practices, and preventing bugs.

    105 GitHub stars~4.7k tokensUpdated 6 mo ago
    Auto-check passed
  • Research Phase

    alchemiststudiosDOTai/harness-engineering

    This skill should be used when mapping or researching a codebase to understand its structure, patterns, and architecture.

    105 GitHub stars~1.4k tokensUpdated 6 mo ago
    Auto-check: notes
  • Harness Map

    alchemiststudiosDOTai/harness-engineering

    Map a repository's mechanical harness layers: canonical check command, local and CI gates, architecture boundaries, structural rules, behavioral verification, docs ratchets, evidence workflows, and…

    105 GitHub stars~1.8k tokensUpdated 6 mo ago
    Auto-check passed
  • Agents Md Mapper

    alchemiststudiosDOTai/harness-engineering

    This skill should be used when creating, refreshing, or validating a repository AGENTS.md so it stays concise, current, and grounded in repository evidence.

    105 GitHub stars~1.8k tokensUpdated 6 mo ago
    Auto-check passed
  • Differential Session Runner

    alchemiststudiosDOTai/harness-engineering

    Run or continue a differential debugging session between two implementations, traces, captures, or outputs.

    105 GitHub stars~1.7k tokensUpdated 6 mo ago
    Auto-check: notes
  • Plan Phase

    alchemiststudiosDOTai/harness-engineering

    Generate execution-ready implementation plans from research docs - planning ONLY, no fixing or verifying.

    105 GitHub stars~1.8k tokensUpdated 6 mo ago
    Auto-check passed

Questions about Execute Phase

What does Execute Phase do?

Execute implementation plans from .artifacts/plan/. An agent skill from alchemiststudiosDOTai/harness-engineering. Execute Phase is an agent skill from alchemiststudiosDOTai/harness-engineering.artifacts/plan/.

When should I use Execute Phase?

Execute Phase fits situations like: the user says execute this plan; provides a plan path.

How do I install Execute Phase in Claude Code?

Run `npx skills add alchemiststudiosDOTai/harness-engineering --skill execute-phase -a claude-code`. Or copy the skill folder (skills/execute-phase in alchemiststudiosDOTai/harness-engineering) into .claude/skills/execute-phase in your project. Claude Code loads it when a task matches its description.

How do I install Execute Phase in Codex?

Run `npx skills add alchemiststudiosDOTai/harness-engineering --skill execute-phase -a codex`. Or copy the skill folder (skills/execute-phase in alchemiststudiosDOTai/harness-engineering) into .agents/skills/execute-phase in your project. Codex loads it when a task matches its description.

Can I use Execute Phase in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add alchemiststudiosDOTai/harness-engineering --skill execute-phase -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/execute-phase, .gemini/skills/execute-phase, .github/skills/execute-phase and .opencode/skills/execute-phase in your project.

What does Execute Phase need to run?

Going by SKILL.md and its folder, Execute Phase needs the command-line tools its instructions call (git, pytest, mypy, black and gh). Our summary lists: Docker. Its frontmatter pre-approves these tools: Edit, Read, Write, Bash(git:*), Bash(python:*), Bash(pytest:*), Bash(mypy:*), Bash(black:*), Bash(coverage:*), Bash(docker:*), Bash(trivy:*), Bash(hadolint:*), Bash(jq:*), Bash(curl:*), Bash(gh:*).

Does Execute Phase access the network?

SKILL.md contains no URLs. Its commands use git and gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Execute Phase safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Execute Phase use?

Execute Phase is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Execute Phase use?

About 1.8k tokens (SKILL.md is roughly 7.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Execute Phase?

Skills that share tags, products or a category with Execute Phase: Improve (fossasia/eventyay-interpretation, 1.6k stars), Workflow Orchestration (vxcozy/workflow-orchestration, 116 stars), Deep Plan (piercelamb/deep-plan, 101 stars) and Plan Py4vasp (vasp-dev/py4vasp, 100 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Execute Phase?

alchemiststudiosDOTai (a GitHub organization) maintains it in alchemiststudiosDOTai/harness-engineering, which has 105 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on March 17, 2026.

Source: alchemiststudiosDOTai/harness-engineering on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.