GreptimeDB Fuzz CI Failure Investigation
GreptimeTeam/greptimedb
Diagnoses a failed GreptimeDB fuzz CI job by pulling its GitHub Actions logs and fuzz artifacts, then matching the evidence to the local source code.
Detects flaky Go tests by analyzing GitHub Actions workflow runs across the last 7 days and all PRs — covering both the run-tests job (unit/integration) and the e2e-test job (gVisor and microVM…
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install agent-substrate/substrate detect-flaky-tests --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .claude/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .claude/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-testsType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install agent-substrate/substrate detect-flaky-tests --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .agents/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .agents/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install agent-substrate/substrate detect-flaky-tests --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .cursor/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .cursor/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/agent-substrate/substrate.git --path .agents/skills/detect-flaky-tests--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install agent-substrate/substrate detect-flaky-tests --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .gemini/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .gemini/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install agent-substrate/substrate detect-flaky-testsInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .github/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .github/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add agent-substrate/substrate --skill detect-flaky-tests -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install agent-substrate/substrate detect-flaky-tests --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/agent-substrate/substrate.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/detect-flaky-tests .opencode/skills/detect-flaky-tests && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "detect-flaky-tests" agent skill from https://github.com/agent-substrate/substrate/tree/main/.agents/skills/detect-flaky-tests into .opencode/skills/detect-flaky-tests/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "detect-flaky-tests", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
detect-flaky-testsDetects flaky Go tests by analyzing GitHub Actions workflow runs across the last 7 days and all PRs — covering both the run-tests job (unit/integration) and the e2e-test job (gVisor and microVM…
Detect Flaky Tests is an agent skill from agent-substrate/substrate. Detects flaky Go tests by analyzing GitHub Actions workflow runs across the last 7 days and all PRs — covering both the run-tests job (unit/integration) and the e2e-test job (gVisor and microVM lanes). For each newly-detected flaky test or infra issue, opens a GitHub issue with full evidence and a draft fix PR. Does not touch BigQuery, dashboards, or any external storage — those are cron-job concerns layered on top.
Its SKILL.md is about 3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Testing & QA, covering Failing and flaky tests and End-to-end testing. It works with Google BigQuery, GitHub and GitHub Actions. The repository describes itself as: Agent Substrate: the core system. The licence is Apache-2.0.
6 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 0b91488. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
ghgokindFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use gh, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Detect Flaky Tests loads about 3k tokens when it runs. Until then it costs about 110 tokens; SKILL.md has 1,181 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from agent-substrate/substrate at commit 0b91488, republished under its Apache-2.0 licence (© agent-substrate). 1,181 words, ~3,003 tokens.
.claude/skills/detect-flaky-tests/SKILL.md (or your agent's skills folder).A test is flaky when it produces both PASS and FAIL outcomes across multiple independent CI runs in the last 7 days, with no code change to that test's package explaining the inconsistency. Cross-PR analysis provides the strongest signal: if the same test fails on PR-A but passes on PR-B, that inconsistency is almost certainly non-determinism, not a legitimate regression.
A test is flagged only when all three conditions hold in the 7-day window:
| Condition | Rationale |
|---|---|
fail_count >= 2 | One failure could be infra noise |
pass_count >= 2 | One pass could be a pre-fix lucky run |
0.05 < fail_rate < 0.95 | Outside this band it is either reliably broken or reliably passing |
SINCE=$(date -u -v-7d +%Y-%m-%dT%H:%M:%SZ 2>/dev/null || date -u -d '7 days ago' +%Y-%m-%dT%H:%M:%SZ)
gh api --paginate \
"repos/agent-substrate/substrate/actions/workflows/pr-workflow.yaml/runs?status=completed&per_page=100&created=>=$SINCE" \
--jq '.workflow_runs[] | {id: .id, conclusion: .conclusion, head_sha: .head_sha, created_at: .created_at}'--paginate is required: a typical week has several hundred completed runs (600+ as of
August 2026), far more than one page of 100.
Collect all run IDs. Process both successful and failed runs — both contain test output.
For each run, you need logs from two jobs:
| Job name | Coverage |
|---|---|
run-tests | Unit + integration tests (go test -race -v ./...) |
e2e-test | E2E suite, both sandbox classes — two sequential steps in the one job |
# List all jobs for a run
gh api "repos/agent-substrate/substrate/actions/runs/<RUN_ID>/jobs" \
--jq '.jobs[] | {id: .id, name: .name, conclusion: .conclusion}'
# Download log for a specific job
gh api "repos/agent-substrate/substrate/actions/jobs/<JOB_ID>/logs" > /tmp/job_<JOB_ID>.logThe two e2e lanes are NOT a matrix — e2e-test is a single job that runs the step
"Run E2E tests (gVisor)" followed by "Run E2E tests (micro-VM)" (same
hack/run-e2e-kind.sh command, the second with E2E_SANDBOX_CLASS: microvm). Split the
one e2e-test job log at the "Run E2E tests (micro-VM)" step boundary: PASS/FAIL lines
before it belong to the gVisor lane, lines after it to the microVM lane. Because the
steps are sequential, a gVisor-lane failure means the micro-VM step never ran — record
no microVM results for that run rather than counting them as failures.
Parse go test -v output from each log:
grep -E '^--- (PASS|FAIL): ' /tmp/job_<JOB_ID>.log \
| awk '{print $2, $3}' | sed 's/://'Track results per (test_name, job_type) where job_type is one of:
unit, e2e-gvisor, e2e-microvm.
An infra failure is when the job itself breaks before or during setup — not when a test produces FAIL output. Infra failures must be identified and reported separately; they do NOT count toward a test's fail_count.
A job log is an infra failure (not a test failure) when it contains ANY of:
| Signal | Example log pattern |
|---|---|
| Go module proxy error | INTERNAL_ERROR, proxy.golang.org: dial, go mod download: ...500 |
| kind cluster creation failure | ERROR: failed to create cluster, node(s) not ready, timed out waiting for the condition |
| Image pull failure | failed to pull image, ErrImagePull, ImagePullBackOff |
| Docker/containerd failure | failed to start containerd, Error response from daemon |
| OOM / out of disk | OOMKilled, No space left on device |
| Network/DNS failure in setup | dial tcp: lookup, connection refused during setup steps (not inside a test) |
| No test output at all | Log ends before any --- PASS or --- FAIL line appears |
| Setup step non-zero exit | A step before the go test or hack/run-e2e-kind.sh command fails |
Critical rule: If a job log contains --- FAIL: TestFoo AND infra error patterns, you
must determine which came first chronologically. If the infra error appears before the first
test ran, treat the whole job as an infra failure (zero test results). If tests started
running and then an infra error interrupted them mid-run, count only the completed test
results and note the truncation.
For any run where you suspect an infra issue, verify it three ways:
=== RUN Test lines, or between test completions?Collect infra failures separately:
| infra_pattern | job_type | fail_count | example_run_ids |
|---|---|---|---|
proxy.golang.org INTERNAL_ERROR | unit | 4 | [run_1, run_2, ...] |
kind cluster: node(s) not ready | e2e-gvisor | 2 | [run_5, ...] |
An infra pattern that appears in ≥2 runs warrants a GitHub issue (Step 5b).
Using only the runs NOT classified as infra failures, build a per-test table:
| test_name | job_type | fail_count | pass_count | total_runs |
|---|
Treat each lane independently. A test that is flaky only in e2e-gvisor is still
flagged — it does not need to be flaky in e2e-microvm too.
Apply the threshold: fail_count >= 2 AND pass_count >= 2 AND 0.05 < fail_rate < 0.95.
Before creating, check for an existing open issue:
gh issue list \
--repo agent-substrate/substrate \
--state open \
--label "kind/bug,area/tests" \
--search "flaky: <TEST_NAME>" \
--json number,titleIf none exists, create:
gh issue create \
--repo agent-substrate/substrate \
--title "flaky: <TEST_NAME>" \
--label "kind/bug,area/tests" \
--body "$(cat <<'BODY'
## Flaky test detected
**Test:** `<TEST_NAME>`
**Job:** `<e2e-gvisor | e2e-microvm | unit>`
### Evidence (last 7 days)
| Metric | Value |
|---|---|
| Runs analysed | <TOTAL_RUNS> |
| Failures | <FAIL_COUNT> |
| Passes | <PASS_COUNT> |
| Flake rate | <FLAKE_RATE>% |
| Infra-failure runs excluded | <INFRA_EXCLUDED> |
### Failing run examples
<links to 2-3 failing runs>
### Passing run examples
<links to 1-2 passing runs>
### Infra triage
Infra failures were excluded before computing this flake rate. The remaining
failures cannot be explained by cluster setup, image pull, or proxy errors.
A draft fix PR will be opened by the detect-flaky-tests agent.
BODY
)"For each infra pattern appearing in ≥2 runs:
gh issue list \
--repo agent-substrate/substrate \
--state open \
--search "infra: <PATTERN_SUMMARY>" \
--json number,titleIf none exists, create:
gh issue create \
--repo agent-substrate/substrate \
--title "infra: <PATTERN_SUMMARY>" \
--label "kind/bug,area/dev-infra" \
--body "$(cat <<'BODY'
## Recurring infrastructure failure in CI
**Pattern:** `<infra error pattern>`
**Job type:** `<unit | e2e-gvisor | e2e-microvm>`
### Evidence
| Metric | Value |
|---|---|
| Occurrences in last 7 days | <COUNT> |
| Example runs | <links> |
### Impact
This failure causes entire CI jobs to abort before tests run. It inflates
apparent failure rates and masks real test flakiness. Fixing it will improve
flakiness signal quality.
### Log excerpt
\`\`\`
<paste 3-5 lines of the actual error from the log>
\`\`\`
BODY
)"Read the test source file. Diagnose the likely cause using the patterns below (unit tests and e2e tests share most root causes, but e2e has additional patterns):
| Pattern | Symptoms | Fix |
|---|---|---|
| Timing / sleep | time.Sleep before an assertion | Replace with require.Eventually or testutil.WaitFor |
| Shared global state | Package-level var mutated without cleanup | Move to test-local; t.Cleanup to restore |
| Port conflicts | Hardcoded port or race on ephemeral port | Use ln.Addr() from the actual listener |
| Goroutine leak | Goroutines from one test race the next | t.Cleanup(cancel) + wait for goroutines to exit |
| File system races | Shared temp path across parallel tests | Use t.TempDir() |
| Context not cancelled | Long operation outlives test | t.Context() (Go 1.21+) or t.Cleanup(cancel) |
| Order dependency | Test relies on prior test's side effects | Make each test self-contained |
| E2E: actor/pod not ready | Test proceeds before actor reaches Running state | Poll with require.Eventually on status, increase timeout with justification |
| E2E: resource cleanup race | Prior test's namespace/actor not fully deleted before next test | Add explicit WaitForDeletion in t.Cleanup |
| E2E: network policy timing | Policy applied but not yet enforced at assertion time | Retry the connectivity check, not just the policy application |
Steps for the fix PR:
fix/flaky-<test-name-kebab> from mainfix(tests): resolve flakiness in <TestName>\n\nFixes #<issue_number>gh pr create \
--repo agent-substrate/substrate \
--title "fix(tests): resolve flakiness in <TestName>" \
--draft \
--body "$(cat <<'BODY'
## Summary
Fixes the flaky test `<TestName>` in `<package>` (`<job_type>` lane).
**Root cause:** <one sentence>
**Fix:** <one sentence>
Closes #<issue_number>
## Evidence
Flake rate over last 7 days: <FLAKE_RATE>% (<FAIL_COUNT> fail / <PASS_COUNT> pass)
Infra-failure runs excluded from count: <INFRA_EXCLUDED>
Failing runs: <links>
Passing runs: <links>
## Test plan
- [ ] Run `go test -race -count=10 ./path/to/package/...` locally for unit tests
- [ ] For e2e: re-run the affected suite 3+ times against a kind cluster
BODY
)"Do not open a fix PR for infra issues — those require infra investigation, not test code changes.
Output two tables:
| Test | Job | Flake rate | Fail/Pass | Infra excluded | Issue | Fix PR | Action |
|---|---|---|---|---|---|---|---|
| TestFoo | e2e-gvisor | 40% | 4/6 | 2 runs | #NNN | #MMM | created |
| Pattern | Job | Occurrences | Issue | Action |
|---|---|---|---|---|
proxy.golang.org INTERNAL_ERROR | unit | 4 | #OOO | created |
If nothing found in either category: No new flaky tests or infra issues detected.
--add-label flag may fail if a label does not exist on the repo; fall back to
omitting labels and add a comment instead.© agent-substrate, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/detect-flaky-tests of agent-substrate/substrate.
Open the folder on GitHubat commit 0b91488
Detect Flaky Tests next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Detect Flaky Tests this skillagent-substrate/substrate | 4.5k | — | ~3k | Automated safety check: Pass | Apache-2.0 | |
| GreptimeDB Fuzz CI Failure InvestigationGreptimeTeam/greptimedb | 6.7k | — | ~4.4k | Automated safety check: Pass | Apache-2.0 | |
| Debugging Opik E2E Testscomet-ml/opik | 22k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | |
| Debug Playwrightquay/quay | 2.8k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | |
| Babysit PRZenUml/web-sequence | 150 | — | ~871 | Automated safety check: Pass | MIT | |
| Analysing CI Failuresgolemcloud/golem | 1.5k | — | ~862 | Automated safety check: Pass | Custom licence |
GreptimeTeam/greptimedb
Diagnoses a failed GreptimeDB fuzz CI job by pulling its GitHub Actions logs and fuzz artifacts, then matching the evidence to the local source code.
comet-ml/opik
Investigates a failed Opik end-to-end test from CI, TestOps or a local run, decides regression versus flake, and proposes a fix without editing tests.
quay/quay
Debug Playwright E2E test failures from GitHub Actions CI runs.
ZenUml/web-sequence
Monitor and diagnose GitHub Actions checks on ZenUML web-sequence PRs, fixing code-caused CI failures when appropriate.
golemcloud/golem
Analysing GitHub Actions CI failures from a run URL. An agent skill from golemcloud/golem.
dotnet/maui
Writes UI tests that reproduce a GitHub issue in .NET MAUI and keeps iterating until the tests actually fail, proving they catch the bug.
agent-substrate/substrate
Posts pull request review findings as GitHub draft (pending) inline comments for a human to edit and submit, instead of publishing them straight to the PR author.
agent-substrate/substrate
Triages open GitHub issues by applying the correct labels. An agent skill from agent-substrate/substrate.
agent-substrate/substrate
Reviews CRDs for compliance with Kubernetes API conventions.
agent-substrate/substrate
Generates a security status report based on docs/threats.json by spinning up sub-agents for each threat to compute a quality score.
agent-substrate/substrate
Generates or updates an AGENTS.md file
Works with
Categories
Detects flaky Go tests by analyzing GitHub Actions workflow runs across the last 7 days and all PRs — covering both the run-tests job (unit/integration) and the e2e-test job (gVisor and microVM…. Detect Flaky Tests is an agent skill from agent-substrate/substrate. Detects flaky Go tests by analyzing GitHub Actions workflow runs across the last 7 days and all PRs — covering both the run-tests job (unit/integration) and the e2e-test job (gVisor and microVM lanes).
Detect Flaky Tests fits situations like: tasks that involve Failing and flaky tests; tasks that involve End-to-end testing.
Run `npx skills add agent-substrate/substrate --skill detect-flaky-tests -a claude-code`. Or copy the skill folder (.agents/skills/detect-flaky-tests in agent-substrate/substrate) into .claude/skills/detect-flaky-tests in your project. Claude Code loads it when a task matches its description.
Run `npx skills add agent-substrate/substrate --skill detect-flaky-tests -a codex`. Or copy the skill folder (.agents/skills/detect-flaky-tests in agent-substrate/substrate) into .agents/skills/detect-flaky-tests in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add agent-substrate/substrate --skill detect-flaky-tests -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/detect-flaky-tests, .gemini/skills/detect-flaky-tests, .github/skills/detect-flaky-tests and .opencode/skills/detect-flaky-tests in your project.
Going by SKILL.md and its folder, Detect Flaky Tests needs the command-line tools its instructions call (gh, go and kind). Our summary lists: Docker.
SKILL.md contains no URLs. Its commands use gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Detect Flaky Tests is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Detect Flaky Tests: GreptimeDB Fuzz CI Failure Investigation (GreptimeTeam/greptimedb, 6.7k stars), Debugging Opik E2E Tests (comet-ml/opik, 22k stars), Debug Playwright (quay/quay, 2.8k stars) and Babysit PR (ZenUml/web-sequence, 150 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
agent-substrate (a GitHub organization) maintains it in agent-substrate/substrate, which has 4,459 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on October 7, 2026.
Source: agent-substrate/substrate on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.