Skill Forge
AgriciDaniel/skill-forge
Ultimate Claude Code skill creator and architect. An agent skill from AgriciDaniel/skill-forge.
Walks you through creating, running and reading waza evals for an agent skill, then proposes concrete fixes when tasks fail or the score is low.
$ npx skills add microsoft/waza --skill waza-interactive -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install microsoft/waza waza-interactive --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/waza-interactive .claude/skills/waza-interactive && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .claude/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/microsoft/waza/tree/main/skills/waza-interactiveType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add microsoft/waza --skill waza-interactive -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install microsoft/waza waza-interactive --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/waza-interactive .agents/skills/waza-interactive && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .agents/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add microsoft/waza --skill waza-interactive -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install microsoft/waza waza-interactive --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/waza-interactive .cursor/skills/waza-interactive && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .cursor/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/microsoft/waza.git --path skills/waza-interactive--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add microsoft/waza --skill waza-interactive -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install microsoft/waza waza-interactive --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/waza-interactive .gemini/skills/waza-interactive && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .gemini/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install microsoft/waza waza-interactiveInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add microsoft/waza --skill waza-interactive -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/waza-interactive .github/skills/waza-interactive && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .github/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add microsoft/waza --skill waza-interactive -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install microsoft/waza waza-interactive --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/microsoft/waza.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/waza-interactive .opencode/skills/waza-interactive && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "waza-interactive" agent skill from https://github.com/microsoft/waza/tree/main/skills/waza-interactive into .opencode/skills/waza-interactive/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "waza-interactive", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
waza-interactiveWalks you through creating, running and reading waza evals for an agent skill, then proposes concrete fixes when tasks fail or the score is low.
The agent acts as a conversational partner for waza, a framework for measuring how well agent skills perform. It works through waza's MCP tools, which list eval suites, fetch an eval spec, validate the YAML, start a benchmark run, poll or cancel it, summarize scores, show per-task results and check a skill for compliance.
For a new eval it asks which skill to test, checks for existing suites, runs `waza init` in the terminal if none exist, explains the generated `eval.yaml` and helps write three to five tasks covering the happy path, an edge case and error handling. After a run it reports pass rate, weighted score and duration, treats 90% and above as strong and anything under 70% as a serious problem, and digs into failed tasks when the pass rate drops below 80%. The excerpt also begins a model-comparison scenario but is cut off there.
8 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 774df00. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Waza Interactive loads about 1.3k tokens when it runs. Until then it costs about 102 tokens; SKILL.md has 636 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from microsoft/waza at commit 774df00, republished under its MIT licence (© microsoft). 636 words, ~1,323 tokens.
.claude/skills/waza-interactive/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.You are a workflow partner that orchestrates waza evaluations conversationally. Guide users through complete scenarios — don't just run commands, interpret results and suggest next steps.
Call these tools to execute waza operations:
| Tool | Purpose |
|---|---|
waza_eval_list | List available eval suites |
waza_eval_get | Get eval spec details |
waza_eval_validate | Validate eval YAML syntax |
waza_eval_run | Execute an eval benchmark |
waza_task_list | List tasks in an eval |
waza_run_status | Poll running eval status |
waza_run_cancel | Cancel a running eval |
waza_results_summary | Get aggregate scores |
waza_results_runs | Get per-task run details |
waza_skill_check | Check skill compliance |
When user wants to create an eval suite for their skill:
waza_eval_list to check for existing evals for this skillwaza init <directory> via terminal to scaffoldeval.yaml structure — name, skill, executor, taskscode, regex)waza_eval_validate to confirm the YAML is validwaza_eval_run to verify the first task passesKey guidance: Start with 3–5 tasks covering happy path, edge case, and error handling.
When user wants to run evals and understand scores:
waza_eval_run with the eval spec path and context dirwaza_run_status until complete (check every 10s)waza_results_summary to get aggregate scoreswaza_results_runs for per-task details on failuresThresholds: ≥90% pass rate = strong, 70–89% = needs work, <70% = significant issues.
When user wants to compare model performance:
waza_eval_run with model A — save resultswaza_eval_run with model B — save resultsGuidance: Run each model 2–3 times to account for variance before drawing conclusions.
When user's skill is failing evals or behaving unexpectedly:
waza_skill_check to verify skill compliance (frontmatter, triggers, token count)waza_eval_run with --verbose and --transcript-dir flagswaza_results_runs to get per-task failure detailswaza_eval_run to verify the fixWhen user asks "is my skill ready?" or wants a pre-ship checklist:
waza_skill_check — verify compliance score ≥ medium-highwaza_eval_validate — confirm eval YAML is validwaza_eval_run — execute full eval suitewaza_results_summary — check aggregate scoresSHIP READINESS CHECKLIST:
☐ Skill compliance: [score] (need: medium-high+)
☐ Eval YAML valid: [yes/no]
☐ Pass rate: [X]% (need: ≥90%)
☐ Weighted score: [X.XX] (need: ≥0.85)
☐ No task timeouts
☐ Consistent across 2+ runs
VERDICT: [READY / NOT READY — fix items marked ✗]© microsoft, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file in skills/waza-interactive of microsoft/waza.
Open the folder on GitHubat commit 774df00
Waza Interactive next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Waza Interactive this skillmicrosoft/waza | 1.4k | — | ~1.3k | Automated safety check: Pass | MIT | |
| Skill ForgeAgriciDaniel/skill-forge | 177 | — | ~1.9k | Automated safety check: Notes | MIT | |
| Skill CreatorZS520L/HanakoPro | 102 | — | ~7.4k | Automated safety check: Pass | Apache-2.0 | |
| Octocode Graph Eval Loopbgauryy/octocode | 946 | — | ~1.6k | Automated safety check: Pass | MIT | |
| Create Skilldavekilleen/Dex | 493 | — | ~2.9k | Automated safety check: Pass | Custom licence | |
| Skill Creatorhimself65/finance-skills | 3.4k | — | ~3.8k | Automated safety check: Pass | MIT |
AgriciDaniel/skill-forge
Ultimate Claude Code skill creator and architect. An agent skill from AgriciDaniel/skill-forge.
ZS520L/HanakoPro
Create new skills, modify and improve existing skills, and measure skill performance.
bgauryy/octocode
Runs a measurable keep-or-discard improvement loop against a runnable sensor, from framing a goal and KPI through baseline, judging and held-out verification.
davekilleen/Dex
Author a new Dex skill that actually fires and passes the quality bar.
himself65/finance-skills
Create, improve, and evaluate agent skills (SKILL.md plus reference files).
vercel/vercel-plugin
Advanced AI agent benchmark scenarios that push Vercel's cutting-edge platform features — Workflow SDK, AI Gateway, MCP, Chat SDK, Queues, Flags, Sandbox, and multi-agent orchestration.
microsoft/waza
Shows a categorized, interactive menu of common Squad operations, such as install, upgrade and team management, and collects arguments before running anything.
microsoft/waza
Shared collaboration rules for a team of squad agents covering worktree awareness, writing decisions to an inbox, cross-agent requests and reviewer lockout.
microsoft/waza
Walks through releasing a new version of the waza azd extension: changelog from commits, semver bump with your confirmation, and a release PR.
microsoft/waza
Dev-first branching model for the Squad project: feature work branches from dev, issue branches follow a naming rule and parallel issues use git worktrees.
microsoft/waza
Evaluates agent skills with a Go CLI that runs YAML-defined benchmarks, compares runs and scores the quality of SKILL.md frontmatter.
microsoft/waza
Reviewer rejection workflow and strict lockout semantics. An agent skill from microsoft/waza.
Works with
Categories
Walks you through creating, running and reading waza evals for an agent skill, then proposes concrete fixes when tasks fail or the score is low. The agent acts as a conversational partner for waza, a framework for measuring how well agent skills perform. It works through waza's MCP tools, which list eval suites, fetch an eval spec, validate the YAML, start a benchmark run, poll or cancel it, summarize scores, show per-task results and check a skill for compliance.
Waza Interactive fits situations like: setting up an eval suite for a skill you are building; running waza benchmarks and making sense of the pass rate and weighted score; finding out why particular eval tasks keep failing; comparing how two models score on the same skill eval.
Run `npx skills add microsoft/waza --skill waza-interactive -a claude-code`. Or copy the skill folder (skills/waza-interactive in microsoft/waza) into .claude/skills/waza-interactive in your project. Claude Code loads it when a task matches its description.
Run `npx skills add microsoft/waza --skill waza-interactive -a codex`. Or copy the skill folder (skills/waza-interactive in microsoft/waza) into .agents/skills/waza-interactive in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add microsoft/waza --skill waza-interactive -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/waza-interactive, .gemini/skills/waza-interactive, .github/skills/waza-interactive and .opencode/skills/waza-interactive in your project.
SKILL.md names no scripts, command-line tools or credentials: Waza Interactive is instructions for the agent only. Our summary lists: The waza CLI; Access to the waza MCP tools.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Waza Interactive is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.3k tokens (SKILL.md is roughly 5.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Waza Interactive: Skill Forge (AgriciDaniel/skill-forge, 177 stars), Skill Creator (ZS520L/HanakoPro, 102 stars), Octocode Graph Eval Loop (bgauryy/octocode, 946 stars) and Create Skill (davekilleen/Dex, 493 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
microsoft (a GitHub organization, an official publisher) maintains it in microsoft/waza, which has 1,400 GitHub stars. The repository holds 16 skills in this directory. The repository was last updated on October 6, 2026.
Source: microsoft/waza on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.