Agent skill

Check

by davepoon in davepoon/buildwithclaude

Run CIAgent regression checks after changing an AI agent's code, prompts, or knowledge base in a repo that has agentcispec.yaml, and interpret the results.

MITAuto-check passedKnowledge Management

Install Check

skills CLI
$ npx skills add davepoon/buildwithclaude --skill check -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install davepoon/buildwithclaude check --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/davepoon/buildwithclaude.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/ciagent/skills/check .claude/skills/check && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
check
GitHub stars
3.6k
Token cost
~680 tokens
SKILL.md length
350 words
Files
1
Skills in repo
246
Repo updated
First seen
Licence
MIT

At a glance

Run CIAgent regression checks after changing an AI agent's code, prompts, or knowledge base in a repo that has agentcispec.yaml, and interpret the results.

  • Asks whether the agent still works
  • SKILL.md covers Which command, Reading results and Rules
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Tasks that involve Knowledge bases

What it does

Check is an agent skill from davepoon/buildwithclaude. Run CIAgent regression checks after changing an AI agent's code, prompts, or knowledge base in a repo that has agentcispec.yaml, and interpret the results. Use after editing agent logic, before committing agent changes, or when the user asks whether the agent still works.

Its SKILL.md is about 680 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Knowledge Management, covering Knowledge bases. The repository describes itself as: A single hub to find Claude Skills, Agents, Commands, Hooks, Plugins, and Marketplace collections to extend Claude Code, Claude Desktop, Agent SDK and OpenClaw. The licence is MIT.

When your agent uses it

  • Asks whether the agent still works
  • Tasks that involve Knowledge bases

Example prompts

  • “/check”

Requirements

  • Pre-approved tools (allowed-tools): Bash(ciagent *)

What it can do on your machine

Read from SKILL.md and the folder at commit 10bfc43. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(ciagent *)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Check loads about 680 tokens when it runs. Until then it costs about 70 tokens; SKILL.md has 350 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~70
When it runs · the whole SKILL.md, loaded when a task matches
~680

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from davepoon/buildwithclaude at commit 10bfc43, republished under its MIT licence (© davepoon). 350 words, ~680 tokens.

Download SKILL.mdSave it as .claude/skills/check/SKILL.md (or your agent's skills folder).
name
check
description
Run CIAgent regression checks after changing an AI agent's code, prompts, or knowledge base in a repo that has agentci_spec.yaml, and interpret the results. Use after editing agent logic, before committing agent changes, or when the user asks whether the agent still works.
allowed-tools
Bash(ciagent *)

Run CIAgent checks on this repo's agent

The repo has agentci_spec.yaml (if it does not, use the onboard skill instead). Your job: run the right check for the change that was just made, read the result correctly, and never paper over a failure.

Which command

SituationCommand
Spec or wiring changed, or no API keysciagent test --mock
Agent code / prompt / retrieval changedciagent test --yes --format json
Result differs from last run, or flakiness suspectedciagent test --runs 3 --yes
Knowledge base changedciagent generate-checks --dry-run, review, then apply
The LLM judge's verdicts look wrongciagent judge-audit

Live runs (test without --mock, judge-audit, generate-checks) call model APIs on the user's keys. Mock mode is free. If the user has not already approved live runs in this session, prefer --mock or ask.

Reading results

Exit codes: 0 pass (including flaky-but-passing), 1 correctness failure (with --runs N: failed in every run), 2 infra or config error — fix the setup, not the agent.

With --format json: per-query entries carry layer results (correctness / path / cost) and the answer text; with --runs N a top-level stability block lists flipped queries with flip_source.

Flip sources route the work:

  • agent-variance — the agent's answer changed between runs → fix the agent (prompt, retrieval, temperature).
  • judge-flake — same answer, the LLM judge changed its verdict → fix the eval (tighten the rubric or replace with a deterministic check).
  • infra-error — a judge API call failed → retry; fix nothing.
  • mixed — ambiguous; look at the answers yourself.
Show full SKILL.md (107 more words)Show less

Rules

  • A correctness failure means the agent lost a fact it used to state. Fix the agent, or — only if the check itself is factually wrong — fix the check. Never weaken or delete a correct check or baseline to make a run green; report the failure to the user instead.
  • After intentionally changing agent behavior, re-record the affected golden: delete its baseline file and rerun ciagent bootstrap --runner <runner> --queries <file> --yes for that query, or update the spec's expectations — with the user's confirmation.
  • Report results in one or two sentences: score, what failed and in which layer, flip sources if any, and the command you ran.

© davepoon, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in plugins/ciagent/skills/check of davepoon/buildwithclaude.

Open the folder on GitHubat commit 10bfc43

Compare with similar skills

Check next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Check compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Check this skilldavepoon/buildwithclaude3.6k—~680Automated safety check: PassMIT
LLM Wiki Knowledge GraphEgonex-AI/Understand-Anything85k1 repos~1.5kAutomated safety check: PassMIT
Xhs Virtual Productchenjin-cmd/xhs-virtual-product726—~862Automated safety check: PassMIT
OpenkbVectifyAI/OpenKB4.7k1 repos~2kAutomated safety check: WarnApache-2.0
Learn From Materialsdmoshehun-prog/learn-from-materials902—~7.9kAutomated safety check: PassMIT
Onyx CLIonyx-dot-app/onyx32k1 repos~2.6kAutomated safety check: PassCustom licence

Similar skills

  • LLM Wiki Knowledge Graph

    Egonex-AI/Understand-Anything

    Detects a Karpathy-pattern LLM wiki and builds an interactive knowledge graph with entities, implicit relationships and topic clusters.

    85k GitHub starsUsed in 1 repo~1.5k tokens
    Knowledge ManagementAuto-check passed
  • Xhs Virtual Product

    chenjin-cmd/xhs-virtual-product

    This skill helps plan, select, produce, and market Xiaohongshu (RED) virtual/digital products — templates, knowledge bases, test tools, study materials.

    726 GitHub stars~862 tokensUpdated 6 days ago
    Knowledge ManagementAuto-check passed
  • Openkb

    VectifyAI/OpenKB

    A skill your agent uses when the user asks about content in their OpenKB knowledge base — research topics, concepts compiled from their documents, cross-document synthesis — or mentions openkb, an…

    4.7k GitHub starsUsed in 1 repo~2k tokens
    Knowledge ManagementAuto-check: warnings
  • Learn From Materials

    dmoshehun-prog/learn-from-materials

    Turns books, PDFs, slides and web pages into a source-grounded knowledge base and an interactive learning page in English or Chinese, with quizzes, relationship maps and reusable methodology notes.

    902 GitHub stars~7.9k tokensUpdated 6 days ago
    Knowledge ManagementAuto-check passed
  • Onyx CLI

    onyx-dot-app/onyx

    Query the Onyx knowledge base using the onyx-cli command. An agent skill from onyx-dot-app/onyx.

    32k GitHub starsUsed in 1 repo~2.6k tokens
    Knowledge ManagementAuto-check passed
  • LLM Wiki

    lewislulu/llm-wiki-skill

    Build and maintain a Karpathy-style LLM knowledge base — a self-compiling Obsidian markdown wiki where an Agent ingests raw sources, compiles cross-linked concept/entity/summary pages, answers…

    655 GitHub stars~3.7k tokensUpdated 5 mo ago
    Knowledge ManagementAuto-check passed

More from davepoon/buildwithclaude

All 246 skills in this repo
  • Qwen Vision

    davepoon/buildwithclaude

    A skill your agent uses when the user asks to "analyze video", "watch this video", "what happens in this video", "describe this clip", "review this footage", "classify these videos", "compare…

    3.6k GitHub starsUsed in 1 repo~1.2k tokens
    Auto-check passed
  • Hard Predict Future

    davepoon/buildwithclaude

    Activate this agent for any future-oriented question that requires deep quantitative analysis, historical precedents, and structured scenario planning.

    3.6k GitHub starsUsed in 1 repo~4.2k tokens
    Auto-check passed
  • iOS Hig Design Guide

    davepoon/buildwithclaude

    Build, update, and apply iOS design specifications using Apple Human Interface Guidelines (HIG) source data.

    3.6k GitHub stars~735 tokensUpdated yesterday
    Auto-check passed
  • Video Downloader

    davepoon/buildwithclaude

    Download YouTube videos with customizable quality and format options.

    3.6k GitHub starsUsed in 1 repo~871 tokens
    Auto-check passed
  • Atlas Cloud Media

    davepoon/buildwithclaude

    Discover Atlas Cloud image and video models, inspect their live schemas, and submit one confirmed media generation request with bounded GET polling.

    3.6k GitHub stars~852 tokensUpdated yesterday
    Auto-check passed
  • Slack Gif Creator

    davepoon/buildwithclaude

    Toolkit for creating animated GIFs optimized for Slack, with validators for size constraints and composable animation primitives.

    3.6k GitHub starsUsed in 12 repos~4.3k tokens
    Auto-check passed

Questions about Check

What does Check do?

Run CIAgent regression checks after changing an AI agent's code, prompts, or knowledge base in a repo that has agentcispec.yaml, and interpret the results. Check is an agent skill from davepoon/buildwithclaude.yaml, and interpret the results.

When should I use Check?

Check fits situations like: asks whether the agent still works; tasks that involve Knowledge bases.

How do I install Check in Claude Code?

Run `npx skills add davepoon/buildwithclaude --skill check -a claude-code`. Or copy the skill folder (plugins/ciagent/skills/check in davepoon/buildwithclaude) into .claude/skills/check in your project. Claude Code loads it when a task matches its description.

How do I install Check in Codex?

Run `npx skills add davepoon/buildwithclaude --skill check -a codex`. Or copy the skill folder (plugins/ciagent/skills/check in davepoon/buildwithclaude) into .agents/skills/check in your project. Codex loads it when a task matches its description.

Can I use Check in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add davepoon/buildwithclaude --skill check -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/check, .gemini/skills/check, .github/skills/check and .opencode/skills/check in your project.

What does Check need to run?

SKILL.md names no scripts, command-line tools or credentials: Check is instructions for the agent only. Its frontmatter pre-approves these tools: Bash(ciagent *).

Does Check access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Check safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Check use?

Check is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Check use?

About 680 tokens (SKILL.md is roughly 2.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Check?

Skills that share tags, products or a category with Check: LLM Wiki Knowledge Graph (Egonex-AI/Understand-Anything, 85k stars), Xhs Virtual Product (chenjin-cmd/xhs-virtual-product, 726 stars), Openkb (VectifyAI/OpenKB, 4.7k stars) and Learn From Materials (dmoshehun-prog/learn-from-materials, 902 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Check?

davepoon (a GitHub user) maintains it in davepoon/buildwithclaude, which has 3,601 GitHub stars. The repository holds 246 skills in this directory. The repository was last updated on October 6, 2026.

Source: davepoon/buildwithclaude on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.