Adk Verify Snippets
google/adk-python
Checks that every Python code block in a Markdown file actually compiles and runs, by extracting each block to a temporary file, executing it in an isolated subprocess, and writing a pass/fail…
Run the confirmed plan's behavioral tests agent-side after assign-to-workforce merges its waves and before summarize-delivery closes the loop, then file what was found — evidence for what passed…
$ npx skills add agentculture/culture --skill validate-delivery -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install agentculture/culture validate-delivery --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/validate-delivery .claude/skills/validate-delivery && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .claude/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/agentculture/culture/tree/main/.claude/skills/validate-deliveryType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add agentculture/culture --skill validate-delivery -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install agentculture/culture validate-delivery --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/validate-delivery .agents/skills/validate-delivery && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .agents/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add agentculture/culture --skill validate-delivery -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install agentculture/culture validate-delivery --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/validate-delivery .cursor/skills/validate-delivery && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .cursor/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/agentculture/culture.git --path .claude/skills/validate-delivery--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add agentculture/culture --skill validate-delivery -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install agentculture/culture validate-delivery --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/validate-delivery .gemini/skills/validate-delivery && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .gemini/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install agentculture/culture validate-deliveryInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add agentculture/culture --skill validate-delivery -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/validate-delivery .github/skills/validate-delivery && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .github/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add agentculture/culture --skill validate-delivery -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install agentculture/culture validate-delivery --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agentculture/culture.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/validate-delivery .opencode/skills/validate-delivery && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "validate-delivery" agent skill from https://github.com/agentculture/culture/tree/main/.claude/skills/validate-delivery into .opencode/skills/validate-delivery/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "validate-delivery", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
validate-deliveryRun the confirmed plan's behavioral tests agent-side after assign-to-workforce merges its waves and before summarize-delivery closes the loop, then file what was found — evidence for what passed…
Validate Delivery is an agent skill from agentculture/culture. Run the confirmed plan's behavioral tests agent-side after assign-to-workforce merges its waves and before summarize-delivery closes the loop, then file what was found — evidence for what passed, behavioral deltas for what the run added, amended, or removed — as first-class, record-only entries via the devague CLI. Never runs the tests inside the CLI (issue 20); never suppresses a failing or partial outcome. Use when the user says "validate delivery", "run behavioral tests", "check what actually behaves", "file…
Its SKILL.md is about 3.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Testing & QA. It works with pytest. The repository describes itself as: Culture turns isolated stochastic agents into cooperative, inspectable, improvable artificial colleagues. The licence is Apache-2.0.
7 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 5d12851. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pytestFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
github.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Validate Delivery loads about 3.3k tokens when it runs. Until then it costs about 225 tokens; SKILL.md has 1,489 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from agentculture/culture at commit 5d12851, republished under its Apache-2.0 licence (© agentculture). 1,489 words, ~3,338 tokens.
.claude/skills/validate-delivery/SKILL.md (or your agent's skills folder).The skill is named validate-delivery; it is the execution-to-evidence
leg of the devague method — the seventh leg in flow order (the eighth
origin skill, chronologically), sitting between the two closing execution
skills:
scope -> think -> challenge -> spec-to-plan -> assign-to-workforce ->
deviate -> validate-delivery -> summarize-deliveryWhere /assign-to-workforce fans out a converged plan's waves and
/summarize-delivery closes the loop afterward, /validate-delivery runs
after waves merge and before the delivery summary is written. It is the
gap that used to be filled by memory: the confirmed plan's claims are
obligations, and until this skill existed nothing forced the run to check
whether the merged code actually behaves as claimed before the summary
asserted it did.
Run this skill once a wave (or the whole plan) has merged and there is
behavior to check against a claim or an approved deviation — always before
/summarize-delivery, never as a substitute for it. It is not gated on a
complete run: a partial or failed fan-out is still worth validating for
whatever did merge.
/deviate records to see what the run promised —
an announcement, an after-state, a success signal, an acceptance
criterion. Each one that has a behavioral test backing it is an
obligation this leg checks.@pytest.mark.behavioral — run with
pytest -m behavioral;tests/behavioral/ or
behavioral-tests/ — run that path directly.pass or fail. A failing outcome is filed
exactly like a passing one; it is never omitted or reworded into
something softer.added / amended / removed — with provenance back
to the claim or approved deviation that motivated it, and forward to the
evidence record(s) that back it./summarize-delivery. The filed evidence and deltas feed
directly into devague summary's Delivery Claims table: evidence
strength (coverage / fidelity / execution / sensitivity) is the
confidence vocabulary there, and any approved lapse on a claim caps its
confidence the same way it always has.Record-only. The devague CLI never runs a test itself (issue
#20) — it only records
what the agent already ran and found. The exact verb shapes below are
minimal placeholders while the underlying schema lands in a parallel task;
treat the verb names as stable and the flags as illustrative, and reconcile
against devague explain <move> once that task merges.
| Move | What it records |
|---|---|
devague oblige <cN> --seam "<seam>" --behavior "<behavior>" | Files a behavioral obligation against a claim, naming the seam to test and the behavior to assert (snapshots the claim text at filing). |
devague evidence --obligation <oN> --test "<ref>" --behavior "<asserted>" --contract "<claim text>" --type <type> --strength <level> --basis "<basis>" --outcome pass|fail [--run-commit <sha> --run-timestamp <ts>] | Files an evidence record: obligation met by this test, asserting this behavior, outcome pass or fail (a run reference is required at execution strength and above). llm-origin filings land proposed; the human adjudicates. |
devague delta --kind added|amended|removed --behavior "<what changed>" --caused-by <cN|dN> [--evidence <eN> ...] | Files a behavioral delta: provenance back to the claim/deviation it diverges from (--caused-by), forward to the evidence that backs it. |
devague summary [--pr] [--json] | Reads the filed evidence and deltas back into the Delivery Claims table (/summarize-delivery's starting point). |
--origin llm on oblige / evidence / delta lands the record proposed
— exactly the same anti-fabrication contract as deviate and lapse: an
agent's own filing never self-confirms, and only the human's --confirm /
--reject moves a proposed record forward. A user-origin filing
auto-approves, mirroring deviate and lapse.
devague oblige / evidence / delta are
record-only moves — they take the agent's already-obtained result and
file it. Running the suite is the agent's job, agent-side, exactly like
/summarize-delivery's read-only verification step (issue #20).llm-origin filings stay proposed until the user confirms. Same
anti-fabrication contract as every other origin vocabulary in this
method — an agent's own proposal never self-confirms./deviate, /validate-delivery does not
add a fourth standing human gate — it produces the record /summarize- delivery and the final PR review consume; the three gates (spec,
implementation split plan, final PR) are unchanged.Wave 2 of a plan merged the export --format widget-md verb. The plan's
confirmed success_signal claim c9 said "round-tripping a widget through
export and back loses no fields." A behavioral test exists for it,
marked @pytest.mark.behavioral, plus two more behavioral tests for
adjacent claims — one of which fails.
# 1. Identify the obligation (echoes its id, e.g. o1)
devague oblige c9 --seam "widget export round-trip" \
--behavior "round-tripping a widget loses no fields"
# 2. Locate and run the behavioral tests agent-side (read-only)
pytest -m behavioral -q
# -> tests/behavioral/test_widget_export.py::test_round_trip PASSED
# -> tests/behavioral/test_widget_export.py::test_empty_field_rendering FAILED
# 3. File evidence for each outcome — the failure included, not smoothed over
devague evidence --obligation o1 \
--test tests/behavioral/test_widget_export.py::test_round_trip \
--behavior "asserts an exported-then-reimported widget compares equal field by field" \
--contract "round-tripping a widget loses no fields" \
--type automated --strength execution \
--basis "behavioral test ran green at the named commit" \
--outcome pass --run-commit abc1234 --run-timestamp 2026-08-31T12:00:00
devague evidence --obligation o2 \
--test tests/behavioral/test_widget_export.py::test_empty_field_rendering \
--behavior "asserts an absent widget field renders as an empty line" \
--contract "an absent field renders honestly, never as filler" \
--type automated --strength execution \
--basis "behavioral test ran red at the named commit" \
--outcome fail --run-commit abc1234 --run-timestamp 2026-08-31T12:00:00
# 4. File a delta if the failure reveals a real behavioral divergence
devague delta --kind amended \
--behavior "empty widget fields render as garbled text, not an empty line" \
--caused-by c11 --evidence e2
# 5. Report faithfully: c9 is validated; c11's claimed behavior is unmet —
# say so plainly, hand it to /summarize-delivery as Remaining Work, not
# as a passing claim./summarize-delivery then reads these back — c9's Delivery Claims row
cites evidence e1 at high confidence (a passing behavioral test); c11's
row is unverified or explicitly failing, never rounded up.
Once every obligation in scope has an evidence record (or is reported as not
yet checkable) and any behavioral deltas are filed, this leg is done — there
is nothing separate to export, the filed records already live in devague
state. Continue with the sibling /summarize-delivery skill: its
Delivery Claims table reads the evidence and deltas filed here directly
(devague summary), so the confidence a claim carries in the final delivery
artifact traces back to a test that actually ran, not to memory. Don't stop
at "tests ran" — the standing flow is file the evidence, then
/summarize-delivery.
The Reasoning Degradation Ledger (devague lapse, issue
agentculture/devague#97)
exists because of this, cited verbatim: "Four graders failed in that
cycle... Every one was found by reading data afterwards; none by a test
failing." That gap — a corrections record reconstructed only at the end,
from memory, because nothing forced a behavioral check to run and be filed
along the way — is exactly what /validate-delivery closes for the
execution side, the same way /challenge closes it for the spec side.
The design itself traces to issue
agentculture/devague#107,
"Suggestion: behavioral validation and a derived current spec," which
proposed behavior as the primary contract, four evidence types, a strength
ladder, and the current spec as a projection of a behavior ledger rather
than a hand-maintained document. This skill is the method-only front door to
that idea: it does not implement the full ledger or the derived-spec
projection — it establishes where in the flow behavioral checking happens,
what gets filed, and how the failure mode #97 documented gets closed instead
of rediscovered.
Previous leg: deviate
Next leg: summarize-deliveryAfter every successful, non-exempt move, the CLI prints one next: <recommended move> line to stderr — follow it, or run devague status when unsure what
comes next. The evidence and deltas filed here also feed devague today's
read-only projection of current behavior into the committed
docs/current-spec.md.
This is a first-party skill — its origin is agentculture/devague, the
eighth in the outbound family after /scope, /think, /challenge,
/spec-to-plan, /assign-to-workforce, /deviate, and
/summarize-delivery, covering the execution-to-evidence leg that runs
after a plan's waves merge and before the delivery summary is written.
guildmaster pulls it from here and broadcasts it to the AgentCulture mesh;
because devague is upstream, it is never re-vendored back from
guildmaster's re-broadcast copy. The cite, don't import policy still
holds: downstream repos copy it, they don't symlink or depend on it. See
docs/skill-sources.md.
© agentculture, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .claude/skills/validate-delivery of agentculture/culture.
Open the folder on GitHubat commit 5d12851
Validate Delivery next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Validate Delivery this skillagentculture/culture | 114 | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | |
| Adk Verify Snippetsgoogle/adk-python | 22k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | |
| Hermetic Python Unit TestsdimensionalOS/dimos | 4.6k | — | ~1.4k | Automated safety check: Pass | Custom licence | |
| Test GuardamElnagdy/guard-skills | 1.3k | 2 repos | ~2.1k | Automated safety check: Pass | MIT | |
| Pytest Runnersaleor/saleor | 23k | — | ~251 | Automated safety check: Pass | BSD-3-Clause | |
| Port Node Red Nodeoldrev/edgelinkd | 125 | — | ~3k | Automated safety check: Pass | Apache-2.0 |
google/adk-python
Checks that every Python code block in a Markdown file actually compiles and runs, by extracting each block to a temporary file, executing it in an isolated subprocess, and writing a pass/fail…
dimensionalOS/dimos
Rules for writing, fixing and reviewing pytest unit tests that are hermetic: behavior-focused, deterministic, isolated and cheap to run.
amElnagdy/guard-skills
Reviews newly written or edited tests against nine rules that cut test bloat, such as mock-heavy checks and near-duplicate cases, before they are committed.
saleor/saleor
Run pytest tests with automatic virtual environment activation. Use this skill whenever running tests, executing pytest, or when asked to "run tests", "test…
oldrev/edgelinkd
Port a Node-RED node into EdgeLinkd the way this repo does it: implement the node in Rust under crates/core/src/runtime/nodes, mirror Node-RED's mocha spec as pytest tests under tests/, register the…
microsoft/onnxruntime
Runs and debugs ONNX Runtime tests: Google Test executables for C++ and unittest or pytest for Python, with filters and build-directory guidance.
agentculture/culture
Show a Culture agent's full configuration in one read-only view: its system-prompt file (CLAUDE.md / AGENTS.md / GEMINI.md), the parallel culture.yaml, and the agent's local .claude/skills index.
agentculture/culture
CI/CD lane for culture: branch, commit, push, create PR, wait for automated reviewers, fetch comments, fix or pushback, reply, resolve threads.
agentculture/culture
All agent communication from culture: in-mesh chat (channels, DMs, mentions, knowledge sharing) via culture channel CLI, AND cross-repo hand-off briefs to sibling-repo agents (agentirc, steward…
agentculture/culture
Cross-repo + mesh communication: file tracked GitHub issues on sibling repos, comment on existing issues, fetch issues with body + comments to inline current state into briefs, and send live…
agentculture/culture
Switch a PyPI package install between the production index, TestPyPI pre-release builds, and a local editable checkout.
agentculture/culture
Run pytest with parallel execution and coverage. An agent skill from agentculture/culture.
Works with
Categories
Run the confirmed plan's behavioral tests agent-side after assign-to-workforce merges its waves and before summarize-delivery closes the loop, then file what was found — evidence for what passed…. Validate Delivery is an agent skill from agentculture/culture. Run the confirmed plan's behavioral tests agent-side after assign-to-workforce merges its waves and before summarize-delivery closes the loop, then file what was found — evidence for what passed, behavioral deltas for what the run added, amended, or removed — as first-class, record-only entries via the devague CLI.
Validate Delivery fits situations like: the user says validate delivery; run behavioral tests; check what actually behaves; record a behavioral delta.
Run `npx skills add agentculture/culture --skill validate-delivery -a claude-code`. Or copy the skill folder (.claude/skills/validate-delivery in agentculture/culture) into .claude/skills/validate-delivery in your project. Claude Code loads it when a task matches its description.
Run `npx skills add agentculture/culture --skill validate-delivery -a codex`. Or copy the skill folder (.claude/skills/validate-delivery in agentculture/culture) into .agents/skills/validate-delivery in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add agentculture/culture --skill validate-delivery -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/validate-delivery, .gemini/skills/validate-delivery, .github/skills/validate-delivery and .opencode/skills/validate-delivery in your project.
Going by SKILL.md and its folder, Validate Delivery needs the command-line tools its instructions call (pytest).
SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Validate Delivery is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.3k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Validate Delivery: Adk Verify Snippets (google/adk-python, 22k stars), Hermetic Python Unit Tests (dimensionalOS/dimos, 4.6k stars), Test Guard (amElnagdy/guard-skills, 1.3k stars) and Pytest Runner (saleor/saleor, 23k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
agentculture (a GitHub organization) maintains it in agentculture/culture, which has 114 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 10, 2026.
Source: agentculture/culture on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.