Agent skill

Loop Execution Evaluator

by Ibrahim-3d in Ibrahim-3d/orchestrator-supaconductor

Evaluate-Loop Step 4: EVALUATE EXECUTION. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

AGPL-3.0Auto-check passedDevelopment

Install Loop Execution Evaluator

skills CLI
$ npx skills add Ibrahim-3d/orchestrator-supaconductor --skill loop-execution-evaluator -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Ibrahim-3d/orchestrator-supaconductor loop-execution-evaluator --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Ibrahim-3d/orchestrator-supaconductor.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/loop-execution-evaluator .claude/skills/loop-execution-evaluator && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
loop-execution-evaluator
GitHub stars
380
Token cost
~1.7k tokens
SKILL.md length
451 words
Files
1
Skills in repo
27
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Evaluate-Loop Step 4: EVALUATE EXECUTION. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

  • Works in 3 steps: The executor's summary includes Business… → Affected documents are listed → This flags the Conductor to run Step 5.5…
  • Tasks that involve UI design
  • SKILL.md covers Why Specialized Evaluators?, Dispatch Logic, Dispatch Workflow and Structural Checks (Always Run), plus 3 more sections
  • Calls npm

What it does

Loop Execution Evaluator is an agent skill from Ibrahim-3d/orchestrator-supaconductor. Evaluate-Loop Step 4: EVALUATE EXECUTION. This is the dispatcher agent — it determines the track type and invokes the correct specialized evaluator. Does NOT run a generic checklist. Instead dispatches to: eval-ui-ux (screens/design), eval-code-quality (features/infrastructure), eval-integration (APIs/auth/payments), eval-business-logic (generator/rules/state). Triggered by: 'evaluate execution', 'review implementation', 'check build', '/phase-review'. Always runs after loop-executor.

Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Development, covering UI design and Code quality. The repository describes itself as: Multi-agent orchestration system for Claude Code with parallel execution, automated quality gates, Board of Directors, and bundled Superpowers skills. The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve UI design
  • Tasks that involve Code quality

Example prompts

  • “evaluate execution”
  • “review implementation”
  • “check build”
  • “/loop-execution-evaluator”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. The executor's summary includes Business Doc Sync Required: Yes
  2. Affected documents are listed
  3. This flags the Conductor to run Step 5.5 (Business Doc Sync) before marking complete

What it can do on your machine

Read from SKILL.md and the folder at commit 76c9b10. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Loop Execution Evaluator loads about 1.7k tokens when it runs. Until then it costs about 129 tokens; SKILL.md has 451 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~129
When it runs · the whole SKILL.md, loaded when a task matches
~1.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Ibrahim-3d/orchestrator-supaconductor at commit 76c9b10, republished under its AGPL-3.0 licence (© Ibrahim-3d). 451 words, ~1,737 tokens.

Download SKILL.mdSave it as .claude/skills/loop-execution-evaluator/SKILL.md (or your agent's skills folder).
name
loop-execution-evaluator
description
Evaluate-Loop Step 4: EVALUATE EXECUTION. This is the dispatcher agent — it determines the track type and invokes the correct specialized evaluator. Does NOT run a generic checklist. Instead dispatches to: eval-ui-ux (screens/design), eval-code-quality (features/infrastructure), eval-integration (APIs/auth/payments), eval-business-logic (generator/rules/state). Triggered by: 'evaluate execution', 'review implementation', 'check build', '/phase-review'. Always runs after loop-executor.

Loop Execution Evaluator — Step 4: Dispatcher

This agent does NOT evaluate directly. It determines the track type and dispatches the correct specialized evaluator.

Why Specialized Evaluators?

Different track types need fundamentally different checks:

  • A UI track needs design system adherence, visual consistency, responsive checks
  • A feature track needs build integrity, type safety, code patterns
  • An integration track needs API contracts, auth flows, error recovery
  • A business logic track needs product rules, edge cases, state transitions

A generic checklist misses critical issues specific to each type.

Dispatch Logic

read_file the track's metadata.json and spec.md to determine the track type, then dispatch:

Track TypeKeywords in spec/metadataEvaluator
UI / Design"screen", "component", "design system", "layout", "visual", "UI shell"eval-ui-ux
Feature / Code"implement", "feature", "refactor", "infrastructure", "hook", "store"eval-code-quality
Integration"Supabase", "Stripe", "Gemini", "API", "auth", "database", "webhook"eval-integration
Business Logic"generation", "lock", "dependency", "pricing", "tier", "pipeline", "download"eval-business-logic
Multi-Type Tracks

Some tracks need multiple evaluators. For example:

  • A generator logic track → eval-business-logic + eval-code-quality
  • An auth/DB integration track → eval-integration + eval-code-quality
  • A UI shell track → eval-ui-ux only

When multiple evaluators apply, run them all. The track passes only if ALL evaluators pass.

Dispatch Workflow

1. read_file track metadata.json + spec.md
2. Determine track type(s)
3. Dispatch evaluator(s):
   → eval-ui-ux         (if UI track)
   → eval-code-quality   (if code/feature track)
   → eval-integration    (if integration track)
   → eval-business-logic (if logic track)
4. Collect results from all dispatched evaluators
5. Aggregate into final verdict

Structural Checks (Always Run)

Regardless of track type, always verify these baseline checks:

CheckMethod
plan.md updatedAll completed tasks marked [x] with commit SHA and summary
Scope alignmentNo unplanned work added without documentation
No skipped tasksAll [ ] tasks either completed or documented as intentionally deferred
Build passesnpm run build exits 0
Business docs in syncIf track made pricing/model/business decisions, verify docs are flagged for Step 5.5 sync
Show full SKILL.md (191 more words)Show less
Business Doc Sync Check

If the track made any business-impacting changes, verify:

  1. The executor's summary includes Business Doc Sync Required: Yes
  2. Affected documents are listed
  3. This flags the Conductor to run Step 5.5 (Business Doc Sync) before marking complete

What counts as business-impacting:

  • Pricing tier, price point, or feature list changes
  • AI model, SDK, or cost structure changes
  • New package or product tier additions
  • Asset pipeline changes (add/remove/modify assets)
  • Persona, GTM, or revenue assumption changes

See ${CLAUDE_PLUGIN_ROOT}/skills/business-docs-sync/SKILL.md for the full registry.

Aggregated Verdict

markdown
## Execution Evaluation Report

**Track**: [track-id]
**Evaluator**: loop-execution-evaluator (dispatcher)
**Date**: [YYYY-MM-DD]

### Evaluators Dispatched
| Evaluator | Reason | Verdict |
|-----------|--------|---------|
| eval-ui-ux | Track builds P0 screens | PASS ✅ / FAIL ❌ |
| eval-code-quality | Track implements features | PASS ✅ / FAIL ❌ |

### Structural Checks
- plan.md updated: YES / NO
- Scope alignment: YES / NO
- Build passes: YES / NO
- Business doc sync needed: YES / NO (if YES, list affected docs)

### Final Verdict: PASS ✅ / FAIL ❌
All evaluators must PASS for the track to pass.

[If FAIL, aggregate all fix actions from all evaluators]

Metadata Checkpoint Updates

The execution evaluator MUST update the track's metadata.json at key points:

On Start
json
{
  "loop_state": {
    "current_step": "EVALUATE_EXECUTION",
    "step_status": "IN_PROGRESS",
    "step_started_at": "[ISO timestamp]",
    "checkpoints": {
      "EVALUATE_EXECUTION": {
        "status": "IN_PROGRESS",
        "started_at": "[ISO timestamp]",
        "agent": "loop-execution-evaluator"
      }
    }
  }
}
On PASS
json
{
  "loop_state": {
    "current_step": "BUSINESS_SYNC",
    "step_status": "NOT_STARTED",
    "checkpoints": {
      "EVALUATE_EXECUTION": {
        "status": "PASSED",
        "completed_at": "[ISO timestamp]",
        "verdict": "PASS",
        "evaluators_run": [
          { "evaluator": "eval-code-quality", "verdict": "PASS", "issues": [] },
          { "evaluator": "eval-business-logic", "verdict": "PASS", "issues": [] }
        ],
        "business_sync_required": true
      },
      "BUSINESS_SYNC": {
        "status": "NOT_STARTED",
        "required": true
      }
    }
  }
}
On FAIL
json
{
  "loop_state": {
    "current_step": "FIX",
    "step_status": "NOT_STARTED",
    "checkpoints": {
      "EVALUATE_EXECUTION": {
        "status": "FAILED",
        "completed_at": "[ISO timestamp]",
        "verdict": "FAIL",
        "evaluators_run": [
          { "evaluator": "eval-code-quality", "verdict": "PASS", "issues": [] },
          { "evaluator": "eval-business-logic", "verdict": "FAIL", "issues": ["Business rule violation found"] }
        ],
        "failure_items": [
          "Fix business rule enforcement in resolver",
          "Add test coverage for edge case"
        ]
      },
      "FIX": {
        "status": "NOT_STARTED",
        "cycle": 1
      }
    }
  }
}
Update Protocol
  1. read_file current metadata.json
  2. Update loop_state.checkpoints.EVALUATE_EXECUTION with results
  3. If PASS + business sync needed: Set current_step to BUSINESS_SYNC
  4. If PASS + no sync needed: Set current_step to COMPLETE
  5. If FAIL: Set current_step to FIX, increment fix_cycle_count in loop_state
  6. write_file back to metadata.json

Handoff

  • ALL PASS + No Business Doc Sync → Conductor marks track complete (Step 5)
  • ALL PASS + Business Doc Sync Needed → Conductor runs Step 5.5 (Business Doc Sync) before marking complete
  • ANY FAIL → Conductor dispatches loop-fixer with combined fix list

© Ibrahim-3d, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/loop-execution-evaluator of Ibrahim-3d/orchestrator-supaconductor.

Open the folder on GitHubat commit 76c9b10

Compare with similar skills

Loop Execution Evaluator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Loop Execution Evaluator compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Loop Execution Evaluator this skillIbrahim-3d/orchestrator-supaconductor380—~1.7kAutomated safety check: PassAGPL-3.0
Warp GUI UI Guidelineswarpdotdev/warp65k1 repos~947Automated safety check: PassAGPL-3.0
Prototype First UIDejavuMoe/Smoji114—~4.2kAutomated safety check: PassApache-2.0
Spec App Consistency Auditleo-kuang-ai/spec-first107—~4.6kAutomated safety check: PassMIT
Extension Reviewervicinaehq/extensions155—~1.1kAutomated safety check: PassNone
HTML Diff Refactor Checktabler/tabler42k—~1.1kAutomated safety check: PassMIT

Similar skills

  • Warp GUI UI Guidelines

    warpdotdev/warp

    Guidelines for writing UI code in Warp's GUI desktop client, centered on reusing shared components and themes instead of adding one-off styling.

    65k GitHub starsUsed in 1 repo~947 tokens
    Frontend & DesignAuto-check passed
  • Prototype First UI

    DejavuMoe/Smoji

    Run a safe prototype-first UI/UX workflow across web, desktop, mobile, extensions, and multimodal design inputs.

    114 GitHub stars~4.2k tokensUpdated 4 days ago
    DevelopmentAuto-check passed
  • Spec App Consistency Audit

    leo-kuang-ai/spec-first

    Audit mobile App PRD/Figma/local-source consistency across page routes, KMP/Clean Architecture, components, analytics, i18n, engineering quality, and industry lenses before runtime validation; use…

    107 GitHub stars~4.6k tokensUpdated 13 days ago
    DevelopmentAuto-check passed
  • Extension Reviewer

    vicinaehq/extensions

    Review Vicinae extensions for publication in the official store, or prepare an extension for submission.

    155 GitHub stars~1.1k tokensUpdated 2 days ago
    DevelopmentAuto-check passed
  • Proves that a refactor leaves every rendered preview page unchanged by comparing fresh HTML output against a baseline captured beforehand.

    42k GitHub stars~1.1k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Om Auto Write Spec

    go-musicfox/go-musicfox

    Autonomously turn a brief or FR issue into a spec landed on a ready PR — runs om-spec-writing --autonomous (defaults posted for override), attaches UI mockups and current-app screenshots as PR…

    2.6k GitHub starsUsed in 1 repo~2.3k tokens
    DevelopmentAuto-check: notes

More from Ibrahim-3d/orchestrator-supaconductor

All 27 skills in this repo
  • Cto Advisor

    Ibrahim-3d/orchestrator-supaconductor

    Technical leadership guidance for engineering teams, architecture decisions, and technology strategy.

    380 GitHub starsUsed in 4 repos~2.4k tokens
    Auto-check passed
  • Context Driven Development

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when working with Conductor's context-driven development methodology, managing project context artifacts, or understanding the relationship between product.md, tech-stack.md…

    380 GitHub starsUsed in 8 repos~2.9k tokens
    Auto-check passed
  • Agent Factory

    Ibrahim-3d/orchestrator-supaconductor

    Creates specialized worker agents dynamically from templates.

    380 GitHub stars~2.9k tokensUpdated 10 days ago
    Auto-check passed
  • Board Of Directors

    Ibrahim-3d/orchestrator-supaconductor

    Simulate a 5-member expert board deliberation for major decisions.

    380 GitHub stars~1.9k tokensUpdated 10 days ago
    Auto-check passed
  • Business Docs Sync

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when completing a track that changes pricing, AI models, product features, or asset pipelines — syncs business context documents across all tiers.

    380 GitHub stars~2.1k tokensUpdated 10 days ago
    Auto-check passed
  • Context Loader

    Ibrahim-3d/orchestrator-supaconductor

    Load project context efficiently for Conductor workflows. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

    380 GitHub stars~830 tokensUpdated 10 days ago
    Auto-check passed

Questions about Loop Execution Evaluator

What does Loop Execution Evaluator do?

Evaluate-Loop Step 4: EVALUATE EXECUTION. An agent skill from Ibrahim-3d/orchestrator-supaconductor. Loop Execution Evaluator is an agent skill from Ibrahim-3d/orchestrator-supaconductor. Evaluate-Loop Step 4: EVALUATE EXECUTION.

When should I use Loop Execution Evaluator?

Loop Execution Evaluator fits situations like: tasks that involve UI design; tasks that involve Code quality.

How do I install Loop Execution Evaluator in Claude Code?

Run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill loop-execution-evaluator -a claude-code`. Or copy the skill folder (skills/loop-execution-evaluator in Ibrahim-3d/orchestrator-supaconductor) into .claude/skills/loop-execution-evaluator in your project. Claude Code loads it when a task matches its description.

How do I install Loop Execution Evaluator in Codex?

Run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill loop-execution-evaluator -a codex`. Or copy the skill folder (skills/loop-execution-evaluator in Ibrahim-3d/orchestrator-supaconductor) into .agents/skills/loop-execution-evaluator in your project. Codex loads it when a task matches its description.

Can I use Loop Execution Evaluator in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill loop-execution-evaluator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/loop-execution-evaluator, .gemini/skills/loop-execution-evaluator, .github/skills/loop-execution-evaluator and .opencode/skills/loop-execution-evaluator in your project.

What does Loop Execution Evaluator need to run?

Going by SKILL.md and its folder, Loop Execution Evaluator needs the command-line tools its instructions call (npm).

Does Loop Execution Evaluator access the network?

SKILL.md contains no URLs. Its commands use npm, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Loop Execution Evaluator safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Loop Execution Evaluator use?

Loop Execution Evaluator is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Loop Execution Evaluator use?

About 1.7k tokens (SKILL.md is roughly 6.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Loop Execution Evaluator?

Skills that share tags, products or a category with Loop Execution Evaluator: Warp GUI UI Guidelines (warpdotdev/warp, 65k stars), Prototype First UI (DejavuMoe/Smoji, 114 stars), Spec App Consistency Audit (leo-kuang-ai/spec-first, 107 stars) and Extension Reviewer (vicinaehq/extensions, 155 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Loop Execution Evaluator?

Ibrahim-3d (a GitHub user) maintains it in Ibrahim-3d/orchestrator-supaconductor, which has 380 GitHub stars. The repository holds 27 skills in this directory. The repository was last updated on September 27, 2026.

Source: Ibrahim-3d/orchestrator-supaconductor on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.