Pydanticai
magnus919/agent-skills
Build type-safe AI agents and graph-based workflows with PydanticAI and PydanticGraph.
Add a new evaluator to the amp-evaluation Python library. An agent skill from wso2/agent-manager.
$ npx skills add wso2/agent-manager --skill add-evaluator -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install wso2/agent-manager add-evaluator --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/add-evaluator .claude/skills/add-evaluator && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .claude/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluatorType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add wso2/agent-manager --skill add-evaluator -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install wso2/agent-manager add-evaluator --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/add-evaluator .agents/skills/add-evaluator && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .agents/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add wso2/agent-manager --skill add-evaluator -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install wso2/agent-manager add-evaluator --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/add-evaluator .cursor/skills/add-evaluator && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .cursor/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/wso2/agent-manager.git --path .claude/skills/add-evaluator--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add wso2/agent-manager --skill add-evaluator -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install wso2/agent-manager add-evaluator --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/add-evaluator .gemini/skills/add-evaluator && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .gemini/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install wso2/agent-manager add-evaluatorInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add wso2/agent-manager --skill add-evaluator -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/add-evaluator .github/skills/add-evaluator && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .github/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add wso2/agent-manager --skill add-evaluator -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install wso2/agent-manager add-evaluator --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wso2/agent-manager.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/add-evaluator .opencode/skills/add-evaluator && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "add-evaluator" agent skill from https://github.com/wso2/agent-manager/tree/main/.claude/skills/add-evaluator into .opencode/skills/add-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "add-evaluator", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
add-evaluatorAdd a new evaluator to the amp-evaluation Python library. An agent skill from wso2/agent-manager.
Add Evaluator is an agent skill from wso2/agent-manager. Add a new evaluator to the amp-evaluation Python library. Use when the user asks to add, write, or register an evaluator, LLM-as-judge, or scoring check for agent traces in libs/amp-evaluation. Covers the decorator API and the type-hint-driven level/mode detection that determines whether the evaluator runs at trace/agent/LLM level and in experiment vs monitor mode.
Its SKILL.md is about 710 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in AI & LLM Engineering, covering Type safety and LLM evaluation. It works with Python. The repository describes itself as: WSO2 AI Agent Manager is an open control plane designed for enterprises to deploy, manage, and govern AI agents at scale. The licence is Apache-2.0.
7 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 6d4af26. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
ruffpippytestblackmypyFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Add Evaluator loads about 710 tokens when it runs. Until then it costs about 95 tokens; SKILL.md has 228 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from wso2/agent-manager at commit 6d4af26, republished under its Apache-2.0 licence (© wso2). 228 words, ~710 tokens.
.claude/skills/add-evaluator/SKILL.md (or your agent's skills folder).Read first: libs/amp-evaluation/AGENTS.md → "Defining an evaluator". This skill is the executable checklist. The non-obvious part is that level and mode are inferred from type hints, not declared.
src/amp_evaluation/evaluators/builtin/ (standard.py for rule-based, llm_judge.py for judges, deepeval.py for DeepEval wrappers). A user-defined one can live anywhere and be picked up by discover_evaluators(module).from amp_evaluation import evaluator, Trace, Task, EvalResult
@evaluator("my-check", description="…", tags=["rule-based", "quality"])
def evaluate(trace: Trace) -> EvalResult:
...Trace → TRACE, AgentTrace → AGENT, LLMSpan → LLM.task parameter:task: Task → EXPERIMENT only;task: Optional[Task] = None → both experiment and monitor;task param → both.build_prompt() (not evaluate()) with the same level/mode detection; tag ["llm-judge", <aspect>]. Needs LLM config via the any-llm extra.Param descriptor: max_latency_ms: float = Param(default=5000, description="…").runner.run().typing.get_type_hints() — keep annotations importable (avoid forward refs that can't resolve).semantic_similarity-style judges need expected_output on the task, so they're EXPERIMENT-only by nature.EvalResult; aggregations (mean/stddev) are computed per-evaluator by the runner from its scores.libs/amp-evaluation/)pip install -e '.[dev]'
pytest # runs with coverage (see pyproject)
ruff check src/ # lint (line-length 120)
black src/ # format (project configures Black; don't also run `ruff format`)
mypy src/ # type-checktask param.tags set (rule-based / llm-judge + aspect).tests/; pytest passes.ruff check + mypy clean.© wso2, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .claude/skills/add-evaluator of wso2/agent-manager.
Open the folder on GitHubat commit 6d4af26
Add Evaluator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Add Evaluator this skillwso2/agent-manager | 105 | — | ~710 | Automated safety check: Pass | Apache-2.0 | |
| Pydanticaimagnus919/agent-skills | 113 | — | ~4k | Automated safety check: Pass | MIT | |
| Azure AI Projects Python SDKmicrosoft/skills | 3.1k | 6 repos | ~2.8k | Automated safety check: Pass | MIT | |
| Formattingbrendanhasz/probflow | 175 | — | ~381 | Automated safety check: Pass | MIT | |
| Pydantic AIdavila7/claude-code-templates | 32k | 4 repos | ~2.9k | Automated safety check: Pass | MIT | |
| Pydantic AIdiegosouzapw/awesome-omni-skills | 159 | — | ~3.2k | Automated safety check: Pass | MIT |
magnus919/agent-skills
Build type-safe AI agents and graph-based workflows with PydanticAI and PydanticGraph.
microsoft/skills
Reference for building on Microsoft Foundry with the azure-ai-projects Python SDK: project clients, versioned agents, evaluations, connections, datasets and indexes.
brendanhasz/probflow
Ensure consistent code formatting using the uv package manager and pre-commit.
davila7/claude-code-templates
Build production-ready AI agents with PydanticAI — type-safe tool use, structured outputs, dependency injection, and multi-model support.
diegosouzapw/awesome-omni-skills
PydanticAI — Typed AI Agents in Python workflow skill. An agent skill from diegosouzapw/awesome-omni-skills.
FailproofAI/failproofai
Helps instrument a custom Python or TypeScript agent to record events for Failproof AI, verify what gets written, and run an evaluator worker that scores the runs.
wso2/agent-manager
Add a semantic audit event to agent-manager-service (the Go control plane).
wso2/agent-manager
Add an API-backed feature to the console (React/TypeScript web UI).
wso2/agent-manager
Write a service-layer unit test in agent-manager-service (the Go control plane).
wso2/agent-manager
Add or change a REST API resource in agent-manager-service (the Go control plane).
Works with
Categories
Add a new evaluator to the amp-evaluation Python library. An agent skill from wso2/agent-manager. Add Evaluator is an agent skill from wso2/agent-manager. Add a new evaluator to the amp-evaluation Python library.
Add Evaluator fits situations like: the user asks to add; register an evaluator; scoring check for agent traces in libs/amp-evaluation.
Run `npx skills add wso2/agent-manager --skill add-evaluator -a claude-code`. Or copy the skill folder (.claude/skills/add-evaluator in wso2/agent-manager) into .claude/skills/add-evaluator in your project. Claude Code loads it when a task matches its description.
Run `npx skills add wso2/agent-manager --skill add-evaluator -a codex`. Or copy the skill folder (.claude/skills/add-evaluator in wso2/agent-manager) into .agents/skills/add-evaluator in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add wso2/agent-manager --skill add-evaluator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/add-evaluator, .gemini/skills/add-evaluator, .github/skills/add-evaluator and .opencode/skills/add-evaluator in your project.
Going by SKILL.md and its folder, Add Evaluator needs the command-line tools its instructions call (ruff, pip, pytest, black and mypy). Our summary lists: Python 3.
SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Add Evaluator is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 710 tokens (SKILL.md is roughly 2.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Add Evaluator: Pydanticai (magnus919/agent-skills, 113 stars), Azure AI Projects Python SDK (microsoft/skills, 3.1k stars), Formatting (brendanhasz/probflow, 175 stars) and Pydantic AI (davila7/claude-code-templates, 32k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
wso2 (a GitHub organization) maintains it in wso2/agent-manager, which has 105 GitHub stars. The repository holds 5 skills in this directory. The repository was last updated on October 7, 2026.
Source: wso2/agent-manager on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.