Agent skill

Eval Integration

by Ibrahim-3d in Ibrahim-3d/orchestrator-supaconductor

Specialized integration evaluator for the Evaluate-Loop. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

AGPL-3.0Auto-check passedBackend & APIs

Install Eval Integration

skills CLI
$ npx skills add Ibrahim-3d/orchestrator-supaconductor --skill eval-integration -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Ibrahim-3d/orchestrator-supaconductor eval-integration --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Ibrahim-3d/orchestrator-supaconductor.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/eval-integration .claude/skills/eval-integration && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
eval-integration
GitHub stars
380
Token cost
~1.8k tokens
SKILL.md length
473 words
Files
1
Skills in repo
27
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Specialized integration evaluator for the Evaluate-Loop. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

  • Works in 5 steps: Track's spec.md and plan.md → Environment config (.env.example, env… → API client code (src/lib/) → …
  • Tasks that involve Authentication
  • SKILL.md covers When This Evaluator Is Used, Inputs Required, Evaluation Passes (6 checks) and Verdict Template, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Eval Integration is an agent skill from Ibrahim-3d/orchestrator-supaconductor. Specialized integration evaluator for the Evaluate-Loop. Use this for evaluating tracks that integrate external services — Supabase auth/DB, Stripe payments, Gemini API, third-party APIs. Checks API contracts, auth flows, data persistence, error recovery, environment config, and end-to-end flow integrity. Dispatched by loop-execution-evaluator when track type is 'integration', 'auth', 'payments', or 'api'. Triggered by: 'evaluate integration', 'test auth flow', 'check API', 'verify payments'.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Backend & APIs, covering Authentication, API design and Third-party API integration. It works with Supabase, Stripe and Google Gemini. The repository describes itself as: Multi-agent orchestration system for Claude Code with parallel execution, automated quality gates, Board of Directors, and bundled Superpowers skills. The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve Authentication
  • Tasks that involve API design
  • Tasks that involve Third-party API integration

Example prompts

  • “integration”
  • “payments”
  • “. Triggered by:”
  • “/eval-integration”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Track's spec.md and plan.md
  2. Environment config (.env.example, env variable documentation)
  3. API client code (src/lib/)
  4. Database schema (if Supabase)
  5. Webhook handlers (if Stripe)

What it can do on your machine

Read from SKILL.md and the folder at commit 76c9b10. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown and sql).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Eval Integration loads about 1.8k tokens when it runs. Until then it costs about 129 tokens; SKILL.md has 473 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~129
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Ibrahim-3d/orchestrator-supaconductor at commit 76c9b10, republished under its AGPL-3.0 licence (© Ibrahim-3d). 473 words, ~1,769 tokens.

Download SKILL.mdSave it as .claude/skills/eval-integration/SKILL.md (or your agent's skills folder).
name
eval-integration
description
Specialized integration evaluator for the Evaluate-Loop. Use this for evaluating tracks that integrate external services — Supabase auth/DB, Stripe payments, Gemini API, third-party APIs. Checks API contracts, auth flows, data persistence, error recovery, environment config, and end-to-end flow integrity. Dispatched by loop-execution-evaluator when track type is 'integration', 'auth', 'payments', or 'api'. Triggered by: 'evaluate integration', 'test auth flow', 'check API', 'verify payments'.

Integration Evaluator Agent

Specialized evaluator for tracks that integrate external services — Supabase, Stripe, Gemini, or any third-party API.

When This Evaluator Is Used

Dispatched by loop-execution-evaluator when the track involves:

  • Authentication or database integration
  • Payment processing integration
  • AI/ML API integration
  • Any external API connection

Inputs Required

  1. Track's spec.md and plan.md
  2. Environment config (.env.example, env variable documentation)
  3. API client code (src/lib/)
  4. Database schema (if Supabase)
  5. Webhook handlers (if Stripe)

Evaluation Passes (6 checks)

Pass 1: API Contract Verification
CheckWhat to Look For
Request shapesAPI calls send correct payload structure
Response handlingResponses parsed with correct types
Error responses4xx/5xx errors handled with user-friendly messaging
Rate limitsRate limit handling present (retry, backoff, queue)
TimeoutReasonable timeout set on API calls
Auth headersBearer token / API key sent correctly
markdown
### API Contracts: PASS ✅ / FAIL ❌
- Endpoints verified: [count]
- Missing error handling: [list]
- Type mismatches: [list]
Pass 2: Authentication Flow
CheckWhat to Look For
Sign upCreates user, stores token, redirects to dashboard
Sign inValidates credentials, stores token, redirects
Sign outClears token, redirects to home
Token refreshHandles expired tokens (refresh or re-auth)
Protected routesUnauthenticated users redirected to login
OAuthThird-party login flow (if applicable)
markdown
### Auth Flow: PASS ✅ / FAIL ❌
- Flows tested: [sign up / sign in / sign out / token refresh]
- Broken flows: [list]
- Token handling: [correct / issues]
Pass 3: Data Persistence & Schema Hygiene

CRUD Operations:

CheckWhat to Look For
CreateData saved correctly to database/storage
read_fileData retrieved and rendered correctly
UpdateChanges persisted on save
DeleteRecords removed, UI reflects deletion
RelationshipsForeign keys / joins working correctly
StorageFile uploads stored and retrievable (if applicable)

Database Schema Quality (MANDATORY for all new tables/migrations):

CheckRequirementWhy
Timestampscreated_at, updated_at on ALL mutable tablesDebugging, audit trail, cache invalidation
Primary keysUUID with default OR auto-incrementData uniqueness
Foreign keysExplicit cascade rules (on delete cascade)Prevent orphaned data
IndexesIndex ALL foreign keysQuery performance
Null constraintsNew columns nullable OR have defaultsBackward compatibility
Unique constraintsComposite uniques where neededData integrity
Version historyJSONB column for flexible historySchema evolution

Schema Anti-Patterns to Flag:

sql
-- ❌ BAD: No timestamps
create table brands (
  id uuid primary key,
  name text
);

-- ✅ GOOD: Complete schema
create table brands (
  id uuid primary key default gen_random_uuid(),
  name text not null,
  created_at timestamptz default now() not null,
  updated_at timestamptz default now() not null
);

-- ❌ BAD: Foreign key without cascade
brand_id uuid references brands(id)

-- ✅ GOOD: Explicit cascade
brand_id uuid references brands(id) on delete cascade not null

-- ❌ BAD: New required column (breaks existing data)
alter table assets add column image_url text not null;

-- ✅ GOOD: Nullable or has default
alter table assets add column locked boolean default false;
markdown
### Data Persistence & Schema: PASS ✅ / FAIL ❌
- CRUD operations: [which work / which fail]
- Data integrity: [any corruption or loss]
- Storage: [files accessible / issues]
- **Tables missing timestamps: [count] — [list]**
- **Foreign keys without indexes: [count] — [list]**
- **Migrations without defaults: [count] — [list]**
- **Orphaned data risk: [YES/NO — describe]**
Show full SKILL.md (155 more words)Show less
Pass 4: Error Recovery
CheckWhat to Look For
Network failureOffline/timeout → user sees error, can retry
Invalid dataMalformed responses → graceful fallback
Auth failureExpired token → redirect to login, not crash
Payment failureDeclined card → clear message, can retry
API downService unavailable → error state, not blank screen
Partial failureOne API fails, others still work
markdown
### Error Recovery: PASS ✅ / FAIL ❌
- Scenarios tested: [list]
- Unhandled failures: [list]
- User messaging: [clear / missing]
Pass 5: Environment Configuration
CheckWhat to Look For
.env.exampleAll required variables documented
No secrets in codeNo API keys, tokens, or passwords in source files
Environment switchingDev/staging/prod configs separate
Missing varsApp handles missing env vars gracefully (error, not crash)
markdown
### Environment: PASS ✅ / FAIL ❌
- Variables documented: [YES/NO]
- Secrets in code: [NONE / list files with exposed secrets]
- Missing var handling: [graceful / crashes]
Pass 6: End-to-End Flow

Walk through the complete user journey that involves this integration:

FlowSteps to Verify
Auth flowLanding → Sign Up → Verify → Dashboard
Payment flowSelect plan → Checkout → Payment → Confirmation
Generation flowForm → Generate → View → Download
markdown
### E2E Flow: PASS ✅ / FAIL ❌
- Flow tested: [describe]
- Steps completed: [X]/[Y]
- Broken at step: [which step, if any]

Verdict Template

markdown
## Integration Evaluation Report

**Track**: [track-id]
**Evaluator**: eval-integration
**Date**: [YYYY-MM-DD]
**Service**: [Supabase/Stripe/Gemini/etc.]

### Results
| Pass | Status | Issues |
|------|--------|--------|
| 1. API Contracts | PASS/FAIL | [details] |
| 2. Auth Flow | PASS/FAIL | [details] |
| 3. Data Persistence | PASS/FAIL | [details] |
| 4. Error Recovery | PASS/FAIL | [details] |
| 5. Environment | PASS/FAIL | [details] |
| 6. E2E Flow | PASS/FAIL | [details] |

### Verdict: PASS ✅ / FAIL ❌
[If FAIL, list specific fix actions for loop-fixer]

Handoff

  • PASS → Return to loop-execution-evaluator → Conductor marks complete
  • FAIL → Return to loop-execution-evaluator → Conductor dispatches loop-fixer

© Ibrahim-3d, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/eval-integration of Ibrahim-3d/orchestrator-supaconductor.

Open the folder on GitHubat commit 76c9b10

Compare with similar skills

Eval Integration next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Eval Integration compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Eval Integration this skillIbrahim-3d/orchestrator-supaconductor380—~1.8kAutomated safety check: PassAGPL-3.0
Gemini Live API Devgoogle-gemini/gemini-skills4.3k—~4.6kAutomated safety check: PassApache-2.0
API Contract ReviewTencentCloudBase/CloudBase-AI-Toolkit1.1k—~1.5kAutomated safety check: PassMIT
API ForgeEliasOulkadi/shokunin114—~2.9kAutomated safety check: PassMIT
Payment Testingpetrkindlmann/qa-skills165—~4.9kAutomated safety check: PassMIT
DeepChat Provider IntegrationThinkInAIXYZ/deepchat6.4k1 repos~1.2kAutomated safety check: PassApache-2.0

Similar skills

  • Gemini Live API Dev

    google-gemini/gemini-skills

    Official

    A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live.

    4.3k GitHub stars~4.6k tokensUpdated 2 days ago
    Backend & APIsAuto-check passed
  • API Contract Review

    TencentCloudBase/CloudBase-AI-Toolkit

    A skill your agent uses when auditing CloudBase cloud API wrappers, MCP tools, generated action metadata, or related docs for outdated or incorrect action names, parameters, casing, request shapes…

    1.1k GitHub stars~1.5k tokensUpdated yesterday
    Backend & APIsAuto-check passed
  • API Forge

    EliasOulkadi/shokunin

    Design REST/GraphQL APIs with OpenAPI 3.1, error handling, pagination, rate limiting, webhooks, and idempotency.

    114 GitHub stars~2.9k tokensUpdated 3 days ago
    Backend & APIsAuto-check passed
  • Payment Testing

    petrkindlmann/qa-skills

    Test payment and checkout flows end to end against PSP sandboxes — Stripe first, with the general pattern for Adyen/Braintree/PayPal.

    165 GitHub stars~4.9k tokensUpdated 4 mo ago
    Testing & QAAuto-check passed
  • DeepChat Provider Integration

    ThinkInAIXYZ/deepchat

    Guides adding an LLM provider to DeepChat through explicit source changes: collect the provider details, pick a transport path and add registry entries and tests.

    6.4k GitHub starsUsed in 1 repo~1.2k tokens
    DevelopmentAuto-check passed
  • Firecrawl Build Onboarding

    firecrawl/firecrawl

    Gets Firecrawl working in a project: signs you in through the browser, saves FIRECRAWL_API_KEY to .env and picks the first SDK or REST path.

    190k GitHub starsUsed in 1 repo~1.4k tokens
    Backend & APIsAuto-check: notes

More from Ibrahim-3d/orchestrator-supaconductor

All 27 skills in this repo
  • Cto Advisor

    Ibrahim-3d/orchestrator-supaconductor

    Technical leadership guidance for engineering teams, architecture decisions, and technology strategy.

    380 GitHub starsUsed in 4 repos~2.4k tokens
    Auto-check passed
  • Context Driven Development

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when working with Conductor's context-driven development methodology, managing project context artifacts, or understanding the relationship between product.md, tech-stack.md…

    380 GitHub starsUsed in 9 repos~2.9k tokens
    Auto-check passed
  • Agent Factory

    Ibrahim-3d/orchestrator-supaconductor

    Creates specialized worker agents dynamically from templates.

    380 GitHub stars~2.9k tokensUpdated 11 days ago
    Auto-check passed
  • Board Of Directors

    Ibrahim-3d/orchestrator-supaconductor

    Simulate a 5-member expert board deliberation for major decisions.

    380 GitHub stars~1.9k tokensUpdated 11 days ago
    Auto-check passed
  • Business Docs Sync

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when completing a track that changes pricing, AI models, product features, or asset pipelines — syncs business context documents across all tiers.

    380 GitHub stars~2.1k tokensUpdated 11 days ago
    Auto-check passed
  • Context Loader

    Ibrahim-3d/orchestrator-supaconductor

    Load project context efficiently for Conductor workflows. An agent skill from Ibrahim-3d/orchestrator-supaconductor.

    380 GitHub stars~830 tokensUpdated 11 days ago
    Auto-check passed

Categories

Questions about Eval Integration

What does Eval Integration do?

Specialized integration evaluator for the Evaluate-Loop. An agent skill from Ibrahim-3d/orchestrator-supaconductor. Eval Integration is an agent skill from Ibrahim-3d/orchestrator-supaconductor. Specialized integration evaluator for the Evaluate-Loop.

When should I use Eval Integration?

Eval Integration fits situations like: tasks that involve Authentication; tasks that involve API design; tasks that involve Third-party API integration.

How do I install Eval Integration in Claude Code?

Run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill eval-integration -a claude-code`. Or copy the skill folder (skills/eval-integration in Ibrahim-3d/orchestrator-supaconductor) into .claude/skills/eval-integration in your project. Claude Code loads it when a task matches its description.

How do I install Eval Integration in Codex?

Run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill eval-integration -a codex`. Or copy the skill folder (skills/eval-integration in Ibrahim-3d/orchestrator-supaconductor) into .agents/skills/eval-integration in your project. Codex loads it when a task matches its description.

Can I use Eval Integration in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Ibrahim-3d/orchestrator-supaconductor --skill eval-integration -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/eval-integration, .gemini/skills/eval-integration, .github/skills/eval-integration and .opencode/skills/eval-integration in your project.

What does Eval Integration need to run?

SKILL.md names no scripts, command-line tools or credentials: Eval Integration is instructions for the agent only.

Does Eval Integration access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Eval Integration safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Eval Integration use?

Eval Integration is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Eval Integration use?

About 1.8k tokens (SKILL.md is roughly 7.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Eval Integration?

Skills that share tags, products or a category with Eval Integration: Gemini Live API Dev (google-gemini/gemini-skills, 4.3k stars), API Contract Review (TencentCloudBase/CloudBase-AI-Toolkit, 1.1k stars), API Forge (EliasOulkadi/shokunin, 114 stars) and Payment Testing (petrkindlmann/qa-skills, 165 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Eval Integration?

Ibrahim-3d (a GitHub user) maintains it in Ibrahim-3d/orchestrator-supaconductor, which has 380 GitHub stars. The repository holds 27 skills in this directory. The repository was last updated on September 27, 2026.

Source: Ibrahim-3d/orchestrator-supaconductor on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.