Agent skill

Flaky Hunter

by Archive228 in Archive228/loopkit

Diagnose and fix tests that pass sometimes and fail other times.

MITAuto-check passedTesting & QA

Install Flaky Hunter

skills CLI
$ npx skills add Archive228/loopkit --skill flaky-hunter -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Archive228/loopkit flaky-hunter --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Archive228/loopkit.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/flaky-hunter .claude/skills/flaky-hunter && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
flaky-hunter
GitHub stars
756
Token cost
~222 tokens
SKILL.md length
106 words
Files
1
Skills in repo
43
Repo updated
First seen
Licence
MIT

At a glance

Diagnose and fix tests that pass sometimes and fail other times.

  • Works in 4 steps: Time/order — depends on test execution… → Async race — asserting before a… → Real network/clock/random — mock them.… → …
  • CI is red intermittently
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Flaky Hunter is an agent skill from Archive228/loopkit. Diagnose and fix tests that pass sometimes and fail other times. Use when CI is red intermittently.

Its SKILL.md is about 220 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA. The repository describes itself as: 33 battle-tested skills + minimal .claude harness for any coding agent (Claude Code, Cursor, Codex, Gemini CLI). The licence is MIT.

When your agent uses it

  • CI is red intermittently

Example prompts

  • “/flaky-hunter”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Time/order — depends on test execution order or shared mutable state. Isolate it; run alone.
  2. Async race — asserting before a promise/refetch resolves. Await the actual condition, not a sleep.
  3. Real network/clock/random — mock them. Freeze time, seed RNG, stub the call.
  4. Resource leak — a prior test left a connection/file/port open.

What it can do on your machine

Read from SKILL.md and the folder at commit 5ae033e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Flaky Hunter loads about 222 tokens when it runs. Until then it costs about 28 tokens; SKILL.md has 106 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~28
When it runs · the whole SKILL.md, loaded when a task matches
~222

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Archive228/loopkit at commit 5ae033e, republished under its MIT licence (© Archive228). 106 words, ~222 tokens.

Download SKILL.mdSave it as .claude/skills/flaky-hunter/SKILL.md (or your agent's skills folder).
name
flaky-hunter
description
Diagnose and fix tests that pass sometimes and fail other times. Use when CI is red intermittently.
when_to_use
intermittent CI failure, "passes locally fails in CI", a flaky test

Flaky Test Hunter

Run the suspect test 20x in a loop first — confirm it's actually flaky, not just broken. Common causes, in order of likelihood:

  1. Time/order — depends on test execution order or shared mutable state. Isolate it; run alone.
  2. Async race — asserting before a promise/refetch resolves. Await the actual condition, not a sleep.
  3. Real network/clock/random — mock them. Freeze time, seed RNG, stub the call.
  4. Resource leak — a prior test left a connection/file/port open. Fix the cause, not the symptom. retry(3) on a flaky test hides a real race that will bite in production. Quarantine only as a last resort, with a ticket.

© Archive228, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/flaky-hunter of Archive228/loopkit.

Open the folder on GitHubat commit 5ae033e

Compare with similar skills

Flaky Hunter next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Flaky Hunter compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Flaky Hunter this skillArchive228/loopkit756—~222Automated safety check: PassMIT
Web Application Testinganthropics/skills180k51 repos~966Automated safety check: PassApache-2.0
Diagnosing Bugsfossasia/eventyay-interpretation1.6k32 repos~2.1kAutomated safety check: PassApache-2.0
TDDpietheinstrengholt/rssmonster56430 repos~906Automated safety check: PassMIT
TDD WorkflowhellangleZ/burn-in-cceverywhere-ralph11211 repos~2.4kAutomated safety check: PassNone
TDDsanity-io/sanity6.4k20 repos~1kAutomated safety check: PassMIT

Similar skills

  • Web Application Testing

    anthropics/skills

    Official

    Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

    180k GitHub starsUsed in 51 repos~966 tokens
    Testing & QAAuto-check passed
  • Diagnosing Bugs

    fossasia/eventyay-interpretation

    Diagnosis loop for hard bugs and performance regressions. An agent skill from fossasia/eventyay-interpretation.

    1.6k GitHub starsUsed in 32 repos~2.1k tokens
    Testing & QAAuto-check passed
  • TDD

    pietheinstrengholt/rssmonster

    Test-driven development. An agent skill from pietheinstrengholt/rssmonster.

    564 GitHub starsUsed in 30 repos~906 tokens
    Testing & QAAuto-check passed
  • TDD Workflow

    hellangleZ/burn-in-cceverywhere-ralph

    A skill your agent uses when writing new features, fixing bugs, or refactoring code.

    112 GitHub starsUsed in 11 repos~2.4k tokens
    Testing & QAAuto-check passed
  • TDD

    sanity-io/sanity

    Official

    Test-driven development with red-green-refactor loop. An agent skill from sanity-io/sanity.

    6.4k GitHub starsUsed in 20 repos~1k tokens
    Testing & QAAuto-check passed
  • Context Driven Development

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when working with Conductor's context-driven development methodology, managing project context artifacts, or understanding the relationship between product.md, tech-stack.md…

    380 GitHub starsUsed in 9 repos~2.9k tokens
    Testing & QAAuto-check passed

More from Archive228/loopkit

All 43 skills in this repo
  • Hitl Escalate

    Archive228/loopkit

    Escalate blocked runs to a human via configured channel or fallback to BLOCKED.md and exit the loop.

    756 GitHub stars~1.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Structured Output

    Archive228/loopkit

    Get JSON out of the model reliably. An agent skill from Archive228/loopkit.

    756 GitHub stars~830 tokensUpdated 2 mo ago
    Auto-check passed
  • Using Loopkit

    Archive228/loopkit

    A skill your agent uses when starting any conversation in a loopkit-enabled project - establishes how to find and use loopkit's 49 skills, requiring skill invocation before ANY response including…

    756 GitHub stars~1.4k tokensUpdated 2 mo ago
    Auto-check passed
  • Active Memory Reminder

    Archive228/loopkit

    Before compaction Loopkit extracts decisions into claude-decisions.json (machine-readable).

    756 GitHub stars~1.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Eval Harness

    Archive228/loopkit

    Build a repeatable eval loop that grades agent output with an LLM judge, so prompt/skill changes get scored against a baseline instead of eyeballed.

    756 GitHub stars~876 tokensUpdated 2 mo ago
    Auto-check passed
  • Feature List JSON

    Archive228/loopkit

    Enumerate every end-to-end feature as strict JSON entries with passes:false, editable-passes-only discipline, and priority order.

    756 GitHub stars~1.2k tokensUpdated 2 mo ago
    Auto-check passed

Categories

Questions about Flaky Hunter

What does Flaky Hunter do?

Diagnose and fix tests that pass sometimes and fail other times. Flaky Hunter is an agent skill from Archive228/loopkit. Diagnose and fix tests that pass sometimes and fail other times.

When should I use Flaky Hunter?

Flaky Hunter fits situations like: CI is red intermittently.

How do I install Flaky Hunter in Claude Code?

Run `npx skills add Archive228/loopkit --skill flaky-hunter -a claude-code`. Or copy the skill folder (skills/flaky-hunter in Archive228/loopkit) into .claude/skills/flaky-hunter in your project. Claude Code loads it when a task matches its description.

How do I install Flaky Hunter in Codex?

Run `npx skills add Archive228/loopkit --skill flaky-hunter -a codex`. Or copy the skill folder (skills/flaky-hunter in Archive228/loopkit) into .agents/skills/flaky-hunter in your project. Codex loads it when a task matches its description.

Can I use Flaky Hunter in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Archive228/loopkit --skill flaky-hunter -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/flaky-hunter, .gemini/skills/flaky-hunter, .github/skills/flaky-hunter and .opencode/skills/flaky-hunter in your project.

What does Flaky Hunter need to run?

SKILL.md names no scripts, command-line tools or credentials: Flaky Hunter is instructions for the agent only.

Does Flaky Hunter access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Flaky Hunter safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Flaky Hunter use?

Flaky Hunter is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Flaky Hunter use?

About 222 tokens (SKILL.md is roughly 888 characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Flaky Hunter?

Skills that share tags, products or a category with Flaky Hunter: Web Application Testing (anthropics/skills, 180k stars), Diagnosing Bugs (fossasia/eventyay-interpretation, 1.6k stars), TDD (pietheinstrengholt/rssmonster, 564 stars) and TDD Workflow (hellangleZ/burn-in-cceverywhere-ralph, 112 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Flaky Hunter?

Archive228 (a GitHub user) maintains it in Archive228/loopkit, which has 756 GitHub stars. The repository holds 43 skills in this directory. The repository was last updated on July 14, 2026.

Source: Archive228/loopkit on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.