Spinning Up Deep Rl
alirezarezvani/claude-skills
Knowledge base from "Spinning Up in Deep RL" by Joshua Achiam (OpenAI, MIT-licensed).
Post-implementation review. An agent skill from ayoubben18/ab-method.
$ npx skills add ayoubben18/ab-method --skill review-implementation -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install ayoubben18/ab-method review-implementation --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/review-implementation .claude/skills/review-implementation && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .claude/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementationType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add ayoubben18/ab-method --skill review-implementation -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install ayoubben18/ab-method review-implementation --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/review-implementation .agents/skills/review-implementation && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .agents/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ayoubben18/ab-method --skill review-implementation -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install ayoubben18/ab-method review-implementation --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/review-implementation .cursor/skills/review-implementation && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .cursor/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/ayoubben18/ab-method.git --path .agents/skills/review-implementation--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add ayoubben18/ab-method --skill review-implementation -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install ayoubben18/ab-method review-implementation --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/review-implementation .gemini/skills/review-implementation && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .gemini/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install ayoubben18/ab-method review-implementationInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add ayoubben18/ab-method --skill review-implementation -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/review-implementation .github/skills/review-implementation && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .github/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add ayoubben18/ab-method --skill review-implementation -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install ayoubben18/ab-method review-implementation --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/ayoubben18/ab-method.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/review-implementation .opencode/skills/review-implementation && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "review-implementation" agent skill from https://github.com/ayoubben18/ab-method/tree/main/.agents/skills/review-implementation into .opencode/skills/review-implementation/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "review-implementation", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
review-implementationPost-implementation review. An agent skill from ayoubben18/ab-method.
Review Implementation is an agent skill from ayoubben18/ab-method. Post-implementation review. Spins up three read-only critics on a completed task's diff — cleaner-architecture, slop-defender, reusability-inspector — that push back ONLY on real issues. Autonomous runs apply safe fixes (tests-green-gated) and write everything to review.md next to progress-tracker.md; interactive runs present findings to pick. Use after a task's missions are done (from start-task / start-roadmap / create-task) or standalone on a diff.
Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `ARCHITECTURE.md`, `REUSE.md` and `SLOP.md`).
The repository describes itself as: A workflow system for Claude Code and Codex. It grills a problem into a domain-grounded plan, then either drives it through test-driven missions you review one at a time, or… The licence is MIT.
5 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 85946e3. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
gitnpxFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git and npx, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Review Implementation loads about 1.5k tokens when it runs. Until then it costs about 119 tokens; SKILL.md has 673 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from ayoubben18/ab-method at commit 85946e3, republished under its MIT licence (© ayoubben18). 673 words, ~1,516 tokens.
.claude/skills/review-implementation/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.Review a completed task's diff through three lenses, each a read-only critic subagent. The counterpart to ../critique-plan/SKILL.md: that guards the plan before coding; this guards the result after.
Silence is the default output. A clean diff produces no findings — that is the normal case. A finding survives only if a senior engineer, looking at this diff, would actually make the change. State each as a concrete cost, not a preference. Never invent refactors to look thorough.
ALWAYS check .ab-method/structure/index.yaml FIRST for paths (task location, domain model,
review.md). Review only what this task changed — the cohesive diff across its missions (git diff
for the commit range). This is not a whole-codebase audit; that's /improve-codebase-architecture.
Look for it under the project root (the current working directory) first — a project's own copy is how it
customises its paths, so it always wins. Only if the project has none (AB Method installed as a plugin rather than
with npx ab-method), read the bundled default: ../../../.ab-method/structure/index.yaml relative to this
SKILL.md. Either way, every path the index names is relative to the project root, never to the folder the
bundled file lives in.
The changed files and their git diff, plus (read only what exists): UBIQUITOUS_LANGUAGE.md /
CONTEXT.md (canonical terms, spotting reinvented concepts), docs/architecture/* (the documented way
things are built here), docs/adr/ (don't propose what an ADR settled), and the task's
unresolved-questions.md if it has one.
Black boxes are deliberate — seed every critic with the file. A TODO(UQ-n) seam whose entry is
recorded there is a decision the user signed off on: ship a placeholder rather than guess. No lens may
flag it as slop, a shallow module, a speculative abstraction, or dead code, and none may propose
"just implement it properly" — the answer isn't theirs to pick. Two things are fair game and should be
reported: a TODO(UQ-n) marker with no matching entry (an orphan black box nobody recorded), and a
placeholder that leaked past its named seam into several call sites — the seam was supposed to contain
it, and containing it again is a safe fix.
Spawn three subagents in one batch, named exactly:
| Agent | Lens | Brief |
|---|---|---|
cleaner-architecture | depth / deepening | ARCHITECTURE.md |
slop-defender | AI code-slop | SLOP.md |
reusability-inspector | duplication / reuse | REUSE.md |
Each is read-only — analyses the diff, returns a findings list (often empty), edits nothing. Seed
each with the diff + context + its lens file — including the task's unresolved-questions.md when it
exists, since every lens would otherwise read its placeholders as defects (each lens file says so too). On Codex (spawn_agent is one level deep) spawn them
at the orchestrator's own level — flat; on Claude they may nest. Flat works on both.
Merge the lists. Drop anything that's a preference (not a concrete cost), an ADR already settled, or a design change bigger than this task (note it open, don't act). Classify each survivor:
safe-fix — confirmed, mechanical, covered by existing tests, no interface/behavior change
(delete dead code, inline a pass-through, drop a redundant comment, call an existing util of identical
semantics).needs-judgment — real but design-level, risky, behavior/interface-affecting, or not test-covered.Interactive (manual create-task tail, or standalone): present findings grouped by lens, each marked
safe-fix/needs-judgment; apply what the user approves; run tests; report.
Autonomous (start-task / start-roadmap — afk): the orchestrator (never the parallel critics)
applies fixes, so there are no concurrent writes:
safe-fix: apply it, re-run the test suite (command from tech-stack.md). Green → keep.
Red → revert it, downgrade to needs-judgment with a note (attempted, reverted — broke <test>).refactor(<task>): post-review cleanup (repo convention). Nothing kept → no commit.docs/tasks/<task>/review.md (below) — every finding, applied or open. Never prompt;
anything uncertain stays open rather than being changed.review.md formatWritten next to progress-tracker.md. Always write it in autonomous mode — even when clean — so the afk
user knows the review ran.
# Post-Implementation Review: <Task Name>
**Reviewed**: YYYY-MM-DD **Scope**: missions <a–z> / commit <range>
## Applied — safe fixes ✅ (commit <hash>)
- [slop-defender] removed pass-through wrapper `fooProxy` — src/foo.ts
- [cleaner-architecture] inlined shallow `formatName` into its one caller — src/user.ts
## Open — need your judgment ⬜
### [reusability-inspector] duplicates `paymentService.calculateTax`
- **Files**: src/checkout/tax.ts
- **Why it matters**: two tax formulas will drift; a rate change must be made twice.
- **Suggested change**: call the existing `paymentService.calculateTax` instead.
## Clean lenses
- slop-defender: no further findingsIf all three lenses are clean and nothing was applied, the whole body is one line:
All three lenses clean — no findings.
© ayoubben18, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 3 other files in .agents/skills/review-implementation of ayoubben18/ab-method.
Open the folder on GitHubat commit 85946e3
Review Implementation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Review Implementation this skillayoubben18/ab-method | 192 | — | ~1.5k | Automated safety check: Pass | MIT | |
| Spinning Up Deep Rlalirezarezvani/claude-skills | 28k | — | ~2.8k | Automated safety check: Pass | MIT | |
| Adaptive Spin Waitmicrosoft/garnet | 12k | — | ~1k | Automated safety check: Pass | MIT | |
| Turnstile Spincloudflare/skills | 3k | 4 repos | ~7.2k | Automated safety check: Notes | Apache-2.0 | |
| Spin Sellinggetagentseal/founder-playbook | 724 | — | ~4.8k | Automated safety check: Pass | MIT | |
| Spin Upaspi6246/Claude-Code-Skills-for-Academics | 158 | — | ~1.9k | Automated safety check: Pass | None |
alirezarezvani/claude-skills
Knowledge base from "Spinning Up in Deep RL" by Joshua Achiam (OpenAI, MIT-licensed).
microsoft/garnet
Selects Garnet spin-wait policies based on the slow-path cost.
cloudflare/skills
Set up, repair, or migrate to Cloudflare Turnstile bot verification in an existing frontend and backend, including server-side Siteverify.
getagentseal/founder-playbook
Applies Neil Rackham's SPIN methodology (Situation/Problem/Implication/Need-payoff questions) to major B2B sales.
aspi6246/Claude-Code-Skills-for-Academics
Start-of-session orientation routine that briefs Claude on the current state of a project before work begins.
swyxio/skills
Add or repair Cloudflare Turnstile on an existing web form by creating or reusing a widget, embedding it, wiring mandatory server-side Siteverify in the existing backend, and validating the result.
ayoubben18/ab-method
Shared vocabulary and principles for designing deep modules — small interfaces, clean seams, testable through the interface.
ayoubben18/ab-method
Grilling session that challenges your plan against the existing domain model, sharpens terminology, and updates documentation (CONTEXT.md, ADRs) inline as decisions crystallise.
ayoubben18/ab-method
Draw a task's blast radius twice. An agent skill from ayoubben18/ab-method.
ayoubben18/ab-method
Compact the current conversation (or a side-topic that surfaced mid-grill) into a handoff document another agent can pick up.
ayoubben18/ab-method
Scan a codebase for deepening opportunities, present them as a visual HTML report, then grill through whichever one you pick.
ayoubben18/ab-method
Cross-plan coherence critic for a whole roadmap. An agent skill from ayoubben18/ab-method.
Post-implementation review. An agent skill from ayoubben18/ab-method. Review Implementation is an agent skill from ayoubben18/ab-method. Post-implementation review.
Run `npx skills add ayoubben18/ab-method --skill review-implementation -a claude-code`. Or copy the skill folder (.agents/skills/review-implementation in ayoubben18/ab-method) into .claude/skills/review-implementation in your project. Claude Code loads it when a task matches its description.
Run `npx skills add ayoubben18/ab-method --skill review-implementation -a codex`. Or copy the skill folder (.agents/skills/review-implementation in ayoubben18/ab-method) into .agents/skills/review-implementation in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ayoubben18/ab-method --skill review-implementation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/review-implementation, .gemini/skills/review-implementation, .github/skills/review-implementation and .opencode/skills/review-implementation in your project.
Going by SKILL.md and its folder, Review Implementation needs the command-line tools its instructions call (git and npx). Our summary lists: Node.js.
SKILL.md contains no URLs. Its commands use git and npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Review Implementation is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.5k tokens (SKILL.md is roughly 6.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Review Implementation: Spinning Up Deep Rl (alirezarezvani/claude-skills, 28k stars), Adaptive Spin Wait (microsoft/garnet, 12k stars), Turnstile Spin (cloudflare/skills, 3k stars) and Spin Selling (getagentseal/founder-playbook, 724 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
ayoubben18 (a GitHub user) maintains it in ayoubben18/ab-method, which has 192 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on October 1, 2026.
Source: ayoubben18/ab-method on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.