Auto Improve
crimeacs/auto-improve
GAN-style iterative improvement loop for any text artifact. An agent skill from crimeacs/auto-improve.
LLM-as-a-judge rubric for code comments (forbidden, false, stale, narration, noise, keep).
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install fmflurry/settings-opencode comment-judge --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/comment-judge .claude/skills/comment-judge && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .claude/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judgeType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install fmflurry/settings-opencode comment-judge --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/comment-judge .agents/skills/comment-judge && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .agents/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install fmflurry/settings-opencode comment-judge --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/comment-judge .cursor/skills/comment-judge && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .cursor/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/fmflurry/settings-opencode.git --path skills/comment-judge--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install fmflurry/settings-opencode comment-judge --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/comment-judge .gemini/skills/comment-judge && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .gemini/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install fmflurry/settings-opencode comment-judgeInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/comment-judge .github/skills/comment-judge && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .github/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add fmflurry/settings-opencode --skill comment-judge -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install fmflurry/settings-opencode comment-judge --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/fmflurry/settings-opencode.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/comment-judge .opencode/skills/comment-judge && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "comment-judge" agent skill from https://github.com/fmflurry/settings-opencode/tree/master/skills/comment-judge into .opencode/skills/comment-judge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "comment-judge", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
comment-judgeLLM-as-a-judge rubric for code comments (forbidden, false, stale, narration, noise, keep).
Comment Judge is an agent skill from fmflurry/settings-opencode. LLM-as-a-judge rubric for code comments (forbidden, false, stale, narration, noise, keep). Loaded in-process by code-reviewer, dotnet-cop and angular-cop to judge every added or changed comment in a review; used by the comment-judge agent for per-module comment purges. Use when judging whether comments are true, useful, or should be deleted/rewritten.
Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Education, covering Quizzes and assessments, Code review and Technical documentation. It works with Angular and .NET. The repository describes itself as: Custom OpenCode settings. The licence is MIT.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 0e6c33c. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
bashgitFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Comment Judge loads about 2.5k tokens when it runs. Until then it costs about 92 tokens; SKILL.md has 1,209 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from fmflurry/settings-opencode at commit 0e6c33c, republished under its MIT licence (© fmflurry). 1,209 words, ~2,543 tokens.
.claude/skills/comment-judge/SKILL.md (or your agent's skills folder).LLM-as-a-judge for code comments. Applies rubric-based verdict, harvests blocks, emits structured findings. Loaded in-process by code-reviewer/dotnet-cop/angular-cop during review; used by the comment-judge agent for module-scoped purges.
REVIEW: Candidates are added/changed comments in the diff. Source: bash scripts/check-added-comments.sh <base> if the repo provides that script; otherwise git diff -U0 <base>...HEAD | grep -nE '^\+.*(//|/\*|#|<!--)', merged with pre-existing comments within ±10 lines of changed hunks (git diff -U10 <base>...HEAD | grep -nE '^\+.*(//|/\*|#|<!--)'). Goal: verify each comment truthfully describes code behavior.
PURGE: Scans a directory for all comments (added or pre-existing). Source: bash scripts/check-added-comments.sh --path <dir> if available; otherwise grep -rnE '(//|/\*|#|<!--)' <dir> (directory-scoped only; repo-wide is prohibitively slow). Goal: delete stale, false, forbidden, or noisy comments in bulk.
EXPLICIT: Caller supplies a list of path:start-end line ranges (used for manual calibration and edge-case handling).
Merge consecutive path:line hits in the same file into blocks: start_line..end_line. One block = one contiguous region. Emit verdict and reasoning once per block, not per line.
REVIEW mode: Missing a false comment is costly (reviewers trust the code). A wrong flag is cheap (human ignores). Strategy: flag suspected false/stale/forbidden comments even with partial evidence. Tier as ❓ (low) when evidence is incomplete.
PURGE mode: Verdicts are bulk-applied to many files. Delete only at high confidence (🔴 or 🟡). Any doubt → keep + ❓ marker for follow-up.
Never auto-apply false: Comment may be correct and code may be wrong = possible bug. Always require human review.
Read the code the comment describes: enclosing symbol, or ±15 lines, or ≤60 lines below a doc block. Never read files >400 lines whole. For stale checks, use codememory_definitions or one targeted Grep.
Ask in order; first failure wins:
a) FORBIDDEN class (per the project's code-comments rule — ~/.claude/rules/common/code-comments.md / ~/.config/opencode/instructions/code-comments.md, or the repo's own copy if it has one)
(ADR-1042) or (RFC 7636 §4.2) for external specs (section allowed for RFCs only). Tier: 🟡.US\d+, EX\d+, SEC\d+, R\d{1,2}, F\d[a-z]?, #\d{5,}, WI \d+, TC #\d+, Phase \d, sub-phase, P\d(\.\d)?, Pin #\d+. Fix: strip the ID; keep the sentence if a why remains. Tier: 🟡.#region markers (tier: 🟡).b) FALSE or STALE (tier: 🔴)
path:line + exact line.codememory_definitions result.c) NARRATION (restates next lines or method signature)
// Creates a new user above public User Create(…) {}.d) NOISE (task/phase/agent history, verbose doc blocks, overstated impl detail on private members)
// As discussed in ticket 1234, // Agent wrote this.e) KEEP (non-obvious why, invariant with bare pointer, public/cross-module contract)
// Sort by created_at DESC; matches UI mockup (ADR-1042).// Retry budget is capped at 3; matches upstream rate-limit window (ADR-2031).f) UNSURE (evidence partial, contradiction unclear, or context ambiguous)
DELETE vs REWRITE (when verdict is forbidden, narration, or noise): After removing the offending part (ticket ids, history, narration, stale claim), does a non-obvious why, invariant, or constraint remain that a reader of the attached code would otherwise miss?
External evidence for FALSE: When a comment cites external evidence (RFC, spec, ADR section, standard name) as proof of its claim:
Rubric order (tie-breaker for equally valid verdicts): When two verdicts seem equally valid:
One finding per judged block, JSON:
{
"path": "src/file.ts",
"start_line": 42,
"end_line": 45,
"category": "comment",
"severity": "high",
"tier": "🔴",
"rule_ref": "code-comments.md#forbidden-class",
"source": "llm",
"existing_code": "// This is thread-safe.\nfunction update(cache, key, value) {",
"verdict": "false",
"suggestion_code": "",
"reason": "Code mutates cache[key] directly without synchronization; comment claims thread-safety but no locking present."
}Optional verdict field (if omitted, finding is recorded but not auto-applied):
keep — no change needed.delete — remove the comment (suggestion_code = "").rewrite — replace with suggestion_code (≤1 line, English, no narration).false — assertion proven false; must quote evidence.Tier (severity mapping):
critical, tier 🔴.medium, tier 🟡.low, tier ❓.Summary line format:
Comments: judged N · keep K · delete D · rewrite R · false F · ❓ QAdd caveat: "Calibrated to REVIEW/PURGE mode; refactor-cleaner may auto-apply delete+rewrite; human must approve false verdicts."
Per-file limit: ≤40 blocks per batch. Total dispatch limit: ≤200 blocks. Over cap: return ## Continuation: next=<path> with the next file to process.
Blind calibration on real cases showed: 65–70% inter-run agreement, catches roughly a quarter of false comments, produces about one wrong delete per run, and can fabricate external evidence (e.g., misattributing which section of an RFC defines a term).
Reliable: forbidden, noise, narration verdicts (consistent across independent runs).
Unreliable: false, stale verdicts (low recall; fabrication risk with external evidence).
Application:
REVIEW mode (code-reviewer / cops include judge findings): Judge false/stale verdicts are advisory flags only (tier ❓). They never block by themselves. Block only when the reviewer confirms by quoting a contradicting line from the repo (not memory or external recall). Deterministic forbidden hits still block as before.
PURGE mode (refactor-cleaner applies verdicts): May auto-apply only delete/rewrite findings whose verdict is forbidden, noise, or narration AND two independent judge runs agree on (same verdict, same block). Require explicit human review ("human skimmed: yes" in brief) before applying; false, stale, and ❓ always go to human. Never apply verdicts whose evidence cites external sources without human confirmation of the source.
Re-calibrate: widen these limits only after a substantial batch of human-approved cases accumulates.
© fmflurry, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/comment-judge of fmflurry/settings-opencode.
Open the folder on GitHubat commit 0e6c33c
Comment Judge next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Comment Judge this skillfmflurry/settings-opencode | 171 | — | ~2.5k | Automated safety check: Pass | MIT | |
| Auto Improvecrimeacs/auto-improve | 135 | — | ~651 | Automated safety check: Pass | MIT | |
| Create Skill Testdotnet/skills | 5.6k | 1 repos | ~6.1k | Automated safety check: Pass | MIT | |
| Malloy Reviewmalloydata/publisher | 116 | — | ~2.5k | Automated safety check: Pass | MIT | |
| PR ReviewNVIDIA/Megatron-LM | 18k | — | ~839 | Automated safety check: Pass | Apache-2.0 | |
| Copilot PR Autopilotgithub/awesome-copilot | 40k | — | ~3.4k | Automated safety check: Pass | MIT |
crimeacs/auto-improve
GAN-style iterative improvement loop for any text artifact. An agent skill from crimeacs/auto-improve.
dotnet/skills
Scaffolds eval.yaml evaluation specs for skills, custom agents, and redistributable gh-aw workflow packages in the dotnet/skills repository.
malloydata/publisher
Malloy semantic-model code review. An agent skill from malloydata/publisher.
NVIDIA/Megatron-LM
Review rubric for the /review pull-request command. An agent skill from NVIDIA/Megatron-LM.
github/awesome-copilot
Copilot left 14 review comments on your PR — half are nits. An agent skill from github/awesome-copilot.
aehrc/pathling
Review a FHIRPath implementation change in Pathling against a correctness rubric covering collection semantics, empty propagation, column cardinality, type coercion, error-vs-empty behaviour, spec…
fmflurry/settings-opencode
Keep a reviewable decision trail for long-running or unattended work: a TSV log with one row per decision (what, why, evidence, result).
fmflurry/settings-opencode
Scaffold and extend Playwright E2E tests for the gc.platform suite (tests/playwright), wiring every artifact to the real frontend (localhost:4200) + real .NET backend — never mocks.
fmflurry/settings-opencode
A skill your agent uses for 'why does X work this way', 'why we picked Y', design rationale, regressions, postmortems, or data-backed thresholds.
fmflurry/settings-opencode
Audit and fix common accessibility issues in Angular templates and Angular Material components.
fmflurry/settings-opencode
Scaffolds and extends Angular standalone feature MODULES under src/app/modules/{name} using Clean Architecture layering (presentation/application/core/infrastructure), a self-registering module…
fmflurry/settings-opencode
Pre-merge code review for Angular + TypeScript pull requests.
Categories
LLM-as-a-judge rubric for code comments (forbidden, false, stale, narration, noise, keep). Comment Judge is an agent skill from fmflurry/settings-opencode. LLM-as-a-judge rubric for code comments (forbidden, false, stale, narration, noise, keep).
Comment Judge fits situations like: judging whether comments are true; should be deleted/rewritten.
Run `npx skills add fmflurry/settings-opencode --skill comment-judge -a claude-code`. Or copy the skill folder (skills/comment-judge in fmflurry/settings-opencode) into .claude/skills/comment-judge in your project. Claude Code loads it when a task matches its description.
Run `npx skills add fmflurry/settings-opencode --skill comment-judge -a codex`. Or copy the skill folder (skills/comment-judge in fmflurry/settings-opencode) into .agents/skills/comment-judge in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add fmflurry/settings-opencode --skill comment-judge -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/comment-judge, .gemini/skills/comment-judge, .github/skills/comment-judge and .opencode/skills/comment-judge in your project.
Going by SKILL.md and its folder, Comment Judge needs the command-line tools its instructions call (bash and git).
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Comment Judge is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.5k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Comment Judge: Auto Improve (crimeacs/auto-improve, 135 stars), Create Skill Test (dotnet/skills, 5.6k stars), Malloy Review (malloydata/publisher, 116 stars) and PR Review (NVIDIA/Megatron-LM, 18k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
fmflurry (a GitHub user) maintains it in fmflurry/settings-opencode, which has 171 GitHub stars. The repository holds 20 skills in this directory. The repository was last updated on October 7, 2026.
Source: fmflurry/settings-opencode on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.