Reproduce Chat States
different-ai/openwork
Fires known chat states in the running OpenWork desktop app, such as provider errors, retries and tool steps, so you can check how each renders.
Track down a behavior that used to work and now fails, changed, or regressed.
$ npx skills add stella/stella --skill regression-hunt -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install stella/stella regression-hunt --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/regression-hunt .claude/skills/regression-hunt && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .claude/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/stella/stella/tree/main/.agents/skills/regression-huntType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add stella/stella --skill regression-hunt -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install stella/stella regression-hunt --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/regression-hunt .agents/skills/regression-hunt && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .agents/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add stella/stella --skill regression-hunt -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install stella/stella regression-hunt --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/regression-hunt .cursor/skills/regression-hunt && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .cursor/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/stella/stella.git --path .agents/skills/regression-hunt--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add stella/stella --skill regression-hunt -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install stella/stella regression-hunt --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/regression-hunt .gemini/skills/regression-hunt && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .gemini/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install stella/stella regression-huntInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add stella/stella --skill regression-hunt -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/regression-hunt .github/skills/regression-hunt && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .github/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add stella/stella --skill regression-hunt -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install stella/stella regression-hunt --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/stella/stella.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/regression-hunt .opencode/skills/regression-hunt && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "regression-hunt" agent skill from https://github.com/stella/stella/tree/main/.agents/skills/regression-hunt into .opencode/skills/regression-hunt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "regression-hunt", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
regression-huntTrack down a behavior that used to work and now fails, changed, or regressed.
Regression Hunt is an agent skill from stella/stella. Track down a behavior that used to work and now fails, changed, or regressed. Use this when a bug report points to a recent breakage, especially when the cause is not obvious yet.
Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Testing & QA, covering QA and bug reports. The repository describes itself as: Open-source legal workspace. The licence is Apache-2.0.
10 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit b225fd8. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
bungitFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Regression Hunt loads about 2.5k tokens when it runs. Until then it costs about 49 tokens; SKILL.md has 1,458 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from stella/stella at commit b225fd8, republished under its Apache-2.0 licence (© stella). 1,458 words, ~2,513 tokens.
.claude/skills/regression-hunt/SKILL.md (or your agent's skills folder).Track down a behavior that used to work and now fails, changed, or regressed. Use this when a bug report points to a recent breakage, especially when the cause is not obvious yet.
$ARGUMENTS: a short description of what regressed.
Helpful extras when available:
A plain-English bug report is enough to start.
Everything else is mechanical. With a fast, deterministic, agent-runnable pass/fail signal, bisection and hypothesis testing all just consume it. Without one, no amount of staring at code will save you. Spend disproportionate effort here. Be aggressive. Be creative. Refuse to give up.
Resolve the repository's test runner and the flags its scripts wire from its instructions first; the commands below show the Bun shape. Try these in roughly this order:
bun run test -- --bail -t "<pattern>", so flags wired into the
script (--preload, custom setup) are preserved; calling bun test
directly bypasses them. Avoid bun --bun test from a worktree root.
Note that bun test positional arguments are file path patterns,
not test names; use -t "<pattern>" to filter by test name and
--bail to fast-fail.AbortSignal.timeout(10_000) so a hung request does not rot the loop.git bisect run bun run test -- --bail -t "<test-name>".
Verify the harness actually fails on a known-bad commit before
starting: bun test exits 0 when no test names match, which would
silently mark every commit as good and produce the wrong culprit.Then treat the loop as a product. Iterate on it:
A 30-second flaky loop is barely better than no loop. A 2-second deterministic loop is a debugging superpower.
For non-deterministic bugs the goal is a higher reproduction rate, not a clean repro. Loop the trigger 100×, parallelise, narrow timing windows, inject sleeps. A 50%-flake bug is debuggable; 1% is not. Keep raising the rate until it is.
If you genuinely cannot build a loop, stop and say so explicitly. List what you tried. Ask the user for: environment access, a captured artifact (HAR, log dump, screen recording with timestamps), or permission to add temporary instrumentation. Do not proceed to hypothesise without a loop.
Only if a correct seam exists, one where the test exercises the real bug pattern as it occurs at the call site. A test at the wrong seam (too shallow, single-caller test when the bug needs multiple callers, unit test that cannot replicate the trigger chain) gives false confidence.
If no correct seam exists, that itself is the finding. Note it and flag the architectural gap for step 10. The codebase is preventing the bug from being locked down.
Prefer focused integration tests over deep unit tests when the regression crosses layers. If the area has no automated harness, build the smallest reproducible check and say why a proper regression test was not added yet.
Single-hypothesis generation anchors on the first plausible idea. Each hypothesis must be falsifiable; state its prediction:
If
<X>is the cause, then<changing Y>will make it disappear /<changing Z>will make it worse.
If you cannot state a prediction, the hypothesis is a vibe; sharpen or discard. Show the ranked list to the user before testing; domain knowledge often re-ranks instantly ("we just deployed a change to #3"). Don't block on it if the user is AFK; proceed with your own ranking.
Use git history when helpful, but do not stop at blame; verify the actual
cause. For regressions specifically, git bisect run against your step 2
loop is often the fastest path to the offending commit.
Each probe must map to a specific prediction from step 4. Change one variable at a time.
Preference:
bun --inspect-brk is a runtime flag that must
attach to the process actually running the code, which means wrapping
it in bun run <script> will not propagate to the spawned child.
Either prepend --inspect-brk to the test command inside your
package script temporarily, or invoke directly while replicating the
flags the script wires, e.g. bun --inspect-brk test --preload ./setup.ts <file-path>. Open the printed devtools:// URL in
Chrome, set one breakpoint at the suspected fault. One breakpoint
beats ten logs.Tag every debug log with a unique prefix, e.g. [DEBUG-a4f2]. Cleanup at
the end becomes a single grep: untagged logs survive; tagged logs die.
For performance regressions, logs are usually wrong. Establish a baseline
measurement first (prefer Bun.nanoseconds() over performance.now() on
the backend (sharper resolution, no Node-portability tax), or a query plan
for DB regressions), then bisect. Measure first, fix second.
[DEBUG-...] instrumentation removed (grep the prefix)Ask: what would have prevented this regression? Make this recommendation after the fix is in; you have more information now than when you started.
Classify the failure before choosing a guard:
Add the strongest feasible guard for a substantial bug. If the strongest guard needs a broader architectural change, land the immediate fix and create a concrete follow-up rather than merely mentioning the idea.
Additional lenses:
© stella, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/regression-hunt of stella/stella.
Open the folder on GitHubat commit b225fd8
Regression Hunt next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Regression Hunt this skillstella/stella | 258 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | |
| Reproduce Chat Statesdifferent-ai/openwork | 24k | — | ~673 | Automated safety check: Pass | Custom licence | |
| Dynamo Jira TicketDynamoDS/Dynamo | 2k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | |
| Moav E2EMotherofallVPNs/MoaV | 448 | — | ~1.9k | Automated safety check: Notes | MIT | |
| Creating A Coral TaskHuman-Agent-Society/CORAL | 1.1k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | |
| Launch Rlmarin-community/marin | 3.9k | — | ~894 | Automated safety check: Pass | Apache-2.0 |
different-ai/openwork
Fires known chat states in the running OpenWork desktop app, such as provider errors, retries and tool steps, so you can check how each renders.
DynamoDS/Dynamo
Create structured Jira tickets for Dynamo from bug reports, failing tests, or feature requests.
MotherofallVPNs/MoaV
Run and debug MoaV's end-to-end tests — real protocol connectivity (client-test.sh) and the moav CLI smoke test — against a LIVE server, via the self-hosted e2e workflow or a local test VPS.
Human-Agent-Society/CORAL
Author a new CORAL task — the three pieces that must line up (task.yaml, seed/, a packaged grader/), the coral init → coral validate → smoke-test loop, and how to pick a grader pattern (stdout…
marin-community/marin
Define, validate, submit, or restart a Marin SkyRL experiment through its artifact main.
actionbook/actionbook
Run browser-based web tests against websites using Actionbook CLI.
stella/stella
Create a concise, evidence-backed implementation plan in the repository planning area when the user explicitly asks for a plan.
stella/stella
Answers data-protection (GDPR) questions grounded in the regulation and supervisory guidance, with a citation for every claim.
stella/stella
Reviews a non-disclosure agreement against the firm's NDA checklist and reports findings with citations.
stella/stella
Collects the facts of an unpaid invoice, then drafts a payment demand letter.
stella/stella
Apply when a performance-guard check (network baseline, bundle baseline, DB query count, loader-prefetch lint, RC bailouts) fails or when touching a hot route/endpoint.
stella/stella
Apply when writing or reviewing React effects in apps/web. An agent skill from stella/stella.
Categories
Track down a behavior that used to work and now fails, changed, or regressed. Regression Hunt is an agent skill from stella/stella. Track down a behavior that used to work and now fails, changed, or regressed.
Regression Hunt fits situations like: tasks that involve QA and bug reports.
Run `npx skills add stella/stella --skill regression-hunt -a claude-code`. Or copy the skill folder (.agents/skills/regression-hunt in stella/stella) into .claude/skills/regression-hunt in your project. Claude Code loads it when a task matches its description.
Run `npx skills add stella/stella --skill regression-hunt -a codex`. Or copy the skill folder (.agents/skills/regression-hunt in stella/stella) into .agents/skills/regression-hunt in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add stella/stella --skill regression-hunt -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/regression-hunt, .gemini/skills/regression-hunt, .github/skills/regression-hunt and .opencode/skills/regression-hunt in your project.
Going by SKILL.md and its folder, Regression Hunt needs the command-line tools its instructions call (bun and git).
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Regression Hunt is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.5k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Regression Hunt: Reproduce Chat States (different-ai/openwork, 24k stars), Dynamo Jira Ticket (DynamoDS/Dynamo, 2k stars), Moav E2E (MotherofallVPNs/MoaV, 448 stars) and Creating A Coral Task (Human-Agent-Society/CORAL, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
stella (a GitHub organization) maintains it in stella/stella, which has 258 GitHub stars. The repository holds 24 skills in this directory. The repository was last updated on October 9, 2026.
Source: stella/stella on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.