Scientific Workflow Tools
DrugClaw/DrugClaw
Research-method workflow guide for hypothesis framing, peer-review style critique, reproducibility planning, study-design checks, and scientific-writing structure.
Evaluates a draft research idea like a top-venue reviewer and advisor, scoring it on five dimensions, checking fit with your capacity and returning a clear verdict.
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluator --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/idea-evaluator .claude/skills/idea-evaluator && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .claude/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluatorType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluator --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/idea-evaluator .agents/skills/idea-evaluator && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .agents/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluator --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/idea-evaluator .cursor/skills/idea-evaluator && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .cursor/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/HKUSTDial/Supervisor-Skills.git --path skills/idea-evaluator--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluator --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/idea-evaluator .gemini/skills/idea-evaluator && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .gemini/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluatorInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/idea-evaluator .github/skills/idea-evaluator && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .github/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install HKUSTDial/Supervisor-Skills idea-evaluator --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/HKUSTDial/Supervisor-Skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/idea-evaluator .opencode/skills/idea-evaluator && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "idea-evaluator" agent skill from https://github.com/HKUSTDial/Supervisor-Skills/tree/main/skills/idea-evaluator into .opencode/skills/idea-evaluator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "idea-evaluator", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
idea-evaluatorEvaluates a draft research idea like a top-venue reviewer and advisor, scoring it on five dimensions, checking fit with your capacity and returning a clear verdict.
The skill scores a preliminary idea on five improvement dimensions (Higher, Faster, Stronger, Cheaper, Broader), matches its lifecycle against your actual capability and weekly hours, probes for paradigm-shift potential and audits it for fatal flaws. It ends with one of three verdicts: Strong Accept, Accept with Revisions, or Reject and Pivot. The stated aim is to drop weak ideas before months are spent on them and to strengthen promising ones before writing starts.
The first step positions the idea as a novel problem, novel method or new setting and asks you to restate it if the story cannot be told in one sentence. The research paradigm is classified by method rather than department: experiment-based work continues with the five dimensions, while text-analysis or conceptual work routes to substitute frameworks. Reference files cover the dimensions, fatal flaws, lifecycle matching, paradigm probes and worked examples. It is not meant for finished work awaiting a paper, manuscript review or brainstorming from scratch.
12 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 207bc6f. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Research Idea Evaluator loads about 3.4k tokens when it runs, and up to ~23k if it reads all its reference files. Until then it costs about 141 tokens; SKILL.md has 1,759 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from HKUSTDial/Supervisor-Skills at commit 207bc6f, republished under its CC-BY-4.0 licence (© HKUSTDial). 1,759 words, ~3,400 tokens.
.claude/skills/idea-evaluator/SKILL.md (or your agent's skills folder). This skill also uses 11 other files; get the full folder from GitHub.This skill evaluates a preliminary research idea from the combined perspective of a top-venue reviewer and an experienced advisor. It scores the idea against five improvement dimensions from the idea-generation guide (Higher, Faster, Stronger, Cheaper, Broader), matches the idea's lifecycle against the user's actual capability and available hours per week, probes whether the idea has paradigm-shift potential, flags fatal flaws, and returns one of three verdicts: Strong Accept, Accept with Revisions, or Reject and Pivot.
The goal is to kill weak ideas before the student invests months, and to shape promising-but-underdeveloped ideas into stronger forms before writing begins.
intro-drafter, tech-paper-template, or
benchmark-paper-template (separate plugin) instead.pre-submission-reviewer.benchmark-paper-template (separate plugin) in targeted mode.Read the user's idea description. In one paragraph, state whether the idea reads as Novel Problem, Novel Method, or New Setting. Is the story compelling in one sentence? If you cannot write that sentence, the idea itself is probably not yet clear enough for evaluation; ask the user to restate.
While restating, classify the research paradigm by method, not by department name: experiments, benchmarks, ablations, or model architectures mean STEM (continue with the five dimensions and fatal flaws below); text analysis, archives, or conceptual argument mean humanities; surveys, interviews, statistics, or fieldwork mean empirical social science; regressions, IV, DID, RDD, or panel data mean finance or economics; statutes, cases, or doctrine mean law.
See: references/domain-evaluation-frameworks.md for the substitute five-dimension frameworks and fatal-flaw substitutions used by the non-STEM paradigms. When the paradigm is unclear, ask one question: "is the core method experiments, surveys, text analysis, or theoretical derivation?"
See: references/fatal-flaws.md for the ten canonical fatal flaws, each with a detection rule and a defense strategy.
Run the fatal-flaws audit before the scoring steps rather than after them. Identify at most two fatal flaws. For each, state the flaw, cite the detection rule, and recommend a concrete defense.
Ground the novelty flaw (F1) in real retrieval whenever the environment has a literature-search capability (a scholarly search tool, web search over scholarly indexes, or shell access to public APIs). Extract two or three keyword groups from the idea (core method plus domain; mechanism plus task; technique plus benchmark), search, and name the three to five closest published works with title, authors, and year. For each, state which axis actually differs: the object acted on, the mechanism, the input granularity, or the problem setting. A similar title alone never establishes duplication; duplication requires failing to find even one differing axis. Retrieval results support metadata-level judgments only (who did what, where); never quote numbers or method details from search snippets. And "not found" does not prove novelty: report it as "no directly overlapping work retrieved under these keywords". With no retrieval capability, label the novelty judgment "unverified; literature check required".
A data-refuted core mechanism is an automatic CRITICAL. If the user's own reported data or attachments already show the core mechanism matched or beaten by a baseline or simple control, lock the verdict to Reject and Pivot; write no defense and invent no optimistic threshold. See: references/fatal-flaws.md, the data-refuted section, for the exact boundary between "refuted by data" and "merely untested".
Short-circuit rule. If any fatal flaw is tagged CRITICAL in the severity taxonomy (single-handedly causes rejection, unfixable within the lifecycle), stop here and emit the verdict directly:
If no CRITICAL flaw is found, continue to Step 3.
See: references/lifecycle-capability-matching.md for the six-category lifecycle matrix, capability self-assessment rubric, and mismatch recovery strategies.
Map the idea onto one of six categories (Application, Foundational Theory, Cross-Disciplinary, Frontier Exploration, Data-Intensive, Innovative Technique). Match against the user's declared capability (effective hours per week, skill depth, theoretical versus applied strength). Output a mismatch flag if lifecycle is shorter than the user's realistic execution window.
See: references/five-dimensions.md for each dimension's entry strategies, scoring rubric, and worked examples.
Score the idea on each of:
Score each 1-10 with explicit evidence from the user's stated contribution. Identify the two or three dimensions where the idea has the highest ceiling and recommend emphasising those in the paper.
Scoring discipline: start every dimension at 5 and justify movement. Two kinds of grounds move a score up, and both count: measured results the user reported (quote them), or a mechanism argument that holds up (label the score "mechanism-based, not yet confirmed by data"). A solid, untested mechanism can reach 8 or 9 with that label plus a named validation experiment; do not systematically cap untested ideas. A dimension with neither data nor mechanism stays at 5 with "no grounds given". Watch attribution: when an impressive gain plausibly comes from a peripheral factor (routing, post-processing, a stronger base model, favorable samples), cap that dimension until an ablation isolates the core mechanism. The scoring reference's final two sections cover both rules in detail.
For non-STEM paradigms, score the substitute dimensions from references/domain-evaluation-frameworks.md instead, under the same discipline.
See: references/paradigm-shift-probe.md for the four probing principles (First Principles, Elephant in the Room, Technology Cycle, Hamming's Rule) and the cross-reference to handbook section 2.3 when deeper disruptive-innovation exploration is needed.
Test the idea against four questions:
Two or more yes answers means the idea has disruptive potential. Note that, and recommend reading handbook 2.3 to deepen the thinking on disruptive-innovation dimensions.
Against the user's stated resources (hardware, data access, team size, engineering skills, timeline), assess:
If any risk is high, flag it explicitly with a suggested mitigation.
Before emitting the verdict, run the checks in the Integrity gate section below.
Issue one of three verdicts:
When the high scores are mechanism-based rather than data-based, qualify the verdict as "worth pursuing, pending the validation experiment", and name that experiment in the top-three actions.
Emit the evaluation in the Output format below.
Each bullet is tagged with an enforceability class. [inspection] means the LLM can verify the bullet from the produced output alone. [attestation] means the LLM states it has done the check, but the user remains responsible for verification. [user-attest] means the bullet is a user-side rule the skill cannot confirm.
Before returning the verdict:
If any [inspection] check fails, downgrade the verdict and mark the corresponding output section as "needs user attention". For [attestation] bullets, the skill states the check was run and the user confirms the result.
Run the gate silently. Do not print a per-gate pass or fail report; a failure surfaces as a concrete finding inside the affected output section, and the delivered evaluation stays free of internal checking rituals.
<Novel Problem or Novel Method or New Setting>| # | Flaw | Severity | Defense |
|---|---|---|---|
| 1 | ... | CRITICAL or MAJOR | ... |
If any CRITICAL flaw is present, skip sections 3-6 and go to section 7 with verdict Reject and Pivot.
| Aspect | User's input | Assessment |
|---|---|---|
| Idea category | ... | ... |
| Lifecycle | ... months | ... |
| Weekly effective hours | ... | ... |
| Fit | ... | Green or Yellow or Red |
| Dimension | Score 1-10 | Evidence | Lift suggestion |
|---|---|---|---|
| Higher | ... | ... | ... |
| Faster | ... | ... | ... |
| Stronger | ... | ... | ... |
| Cheaper | ... | ... | ... |
| Broader | ... | ... | ... |
| Probe | Yes or No | Rationale |
|---|---|---|
| First Principles | ... | ... |
| Elephant in the Room | ... | ... |
| Technology Cycle | ... | ... |
| Hamming's Rule | ... | ... |
Disruptive potential: <none, possible, strong>.
| Risk | Level | Mitigation |
|---|---|---|
| Compute | ... | ... |
| Data | ... | ... |
| Engineering | ... | ... |
| Timeline | ... | ... |
<Strong Accept or Accept with Revisions or Reject and Pivot>
(mechanism-based high scores: append "worth pursuing, pending the
validation experiment")
Top three actions to take first:
© HKUSTDial, CC-BY-4.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 11 other files (references) in skills/idea-evaluator of HKUSTDial/Supervisor-Skills.
Open the folder on GitHubat commit 207bc6f
Research Idea Evaluator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Research Idea Evaluator this skillHKUSTDial/Supervisor-Skills | 8.5k | — | ~3.4k | Automated safety check: Pass | CC-BY-4.0 | |
| Scientific Workflow ToolsDrugClaw/DrugClaw | 125 | — | ~712 | Automated safety check: Pass | Apache-2.0 | |
| Academic Paper Writing PipelineImbad0202/academic-research-skills | 51k | — | ~16k | Automated safety check: Pass | Custom licence | |
| Academic Research PipelineImbad0202/academic-research-skills | 51k | — | ~15k | Automated safety check: Pass | Custom licence | |
| Social Science Paper Writingfakerqwq/social-science-paper-writing-skill | 375 | — | ~7k | Automated safety check: Pass | None | |
| Paper GlanceDiaugeia/paper-glance-skill | 106 | — | ~433 | Automated safety check: Pass | None |
DrugClaw/DrugClaw
Research-method workflow guide for hypothesis framing, peer-review style critique, reproducibility planning, study-design checks, and scientific-writing structure.
Imbad0202/academic-research-skills
Runs a 12-agent pipeline that plans, drafts, cites, reviews and formats academic papers, with modes for revision, rebuttals, abstracts and citation checks.
Imbad0202/academic-research-skills
Orchestrates a ten-stage academic workflow from research to finished manuscript, including integrity checks, two rounds of peer review and revision.
fakerqwq/social-science-paper-writing-skill
Helps draft, diagnose, review and revise social science papers, from topic and research question to literature review, citation risks and pre-submission checks.
Diaugeia/paper-glance-skill
Reads an academic paper from a PDF, pasted text or abstract and offers five outputs: an analysis report, a mind map, a peer review, promotion scripts and a podcast.
vishalsachdev/canvas-mcp
Student weekly assignment planner for Canvas LMS. An agent skill from vishalsachdev/canvas-mcp.
HKUSTDial/Supervisor-Skills
Structures benchmark and evaluation papers around five pillars, with a completeness audit, an Introduction logic chain, a section skeleton and a pre-submission checklist.
HKUSTDial/Supervisor-Skills
Rebuilds a reference diagram image as an editable, high-fidelity Draw.io file, mixing native elements, SVG icons and cropped PNGs, with a batch workflow for a folder of images.
HKUSTDial/Supervisor-Skills
Drafts the Introduction of a technical paper as six paragraphs of flowing prose, positioning the work and matching contributions to challenges, with an outline on request.
HKUSTDial/Supervisor-Skills
Turns pasted peer-review comments into a per-concern rebuttal plan, with reviewer mindset matching and strategy priorities, but not the final rebuttal text.
HKUSTDial/Supervisor-Skills
Runs a survey-grade literature investigation: fixes the research questions, searches from adversarial angles, verifies citations and writes an evidence-first report.
HKUSTDial/Supervisor-Skills
Advises on designing the three core figures of a technical paper, then audits them against rules for format, fonts, color and captions.
Categories
Evaluates a draft research idea like a top-venue reviewer and advisor, scoring it on five dimensions, checking fit with your capacity and returning a clear verdict. The skill scores a preliminary idea on five improvement dimensions (Higher, Faster, Stronger, Cheaper, Broader), matches its lifecycle against your actual capability and weekly hours, probes for paradigm-shift potential and audits it for fatal flaws. It ends with one of three verdicts: Strong Accept, Accept with Revisions, or Reject and Pivot.
Research Idea Evaluator fits situations like: deciding whether a draft research idea is worth pursuing; running a novelty or feasibility check before committing to a paper scope; comparing two or three candidate ideas with a structured trade-off; checking an idea for scope creep.
Run `npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a claude-code`. Or copy the skill folder (skills/idea-evaluator in HKUSTDial/Supervisor-Skills) into .claude/skills/idea-evaluator in your project. Claude Code loads it when a task matches its description.
Run `npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a codex`. Or copy the skill folder (skills/idea-evaluator in HKUSTDial/Supervisor-Skills) into .agents/skills/idea-evaluator in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add HKUSTDial/Supervisor-Skills --skill idea-evaluator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/idea-evaluator, .gemini/skills/idea-evaluator, .github/skills/idea-evaluator and .opencode/skills/idea-evaluator in your project.
SKILL.md names no scripts, command-line tools or credentials: Research Idea Evaluator is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Research Idea Evaluator is published under the CC-BY-4.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.4k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 19k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Research Idea Evaluator: Scientific Workflow Tools (DrugClaw/DrugClaw, 125 stars), Academic Paper Writing Pipeline (Imbad0202/academic-research-skills, 51k stars), Academic Research Pipeline (Imbad0202/academic-research-skills, 51k stars) and Social Science Paper Writing (fakerqwq/social-science-paper-writing-skill, 375 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
HKUSTDial (a GitHub organization) maintains it in HKUSTDial/Supervisor-Skills, which has 8,523 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on September 5, 2026.
Source: HKUSTDial/Supervisor-Skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.