Agent skill

Scholar Brainstorm

by joshzyj in joshzyj/open-scholar-skill

Generate research questions from existing materials — codebooks, survey questionnaires, datasets, or published papers/abstracts.

Custom licenceAuto-check passedAgent Workflows

Install Scholar Brainstorm

skills CLI
$ npx skills add joshzyj/open-scholar-skill --skill scholar-brainstorm -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install joshzyj/open-scholar-skill scholar-brainstorm --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/joshzyj/open-scholar-skill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/scholar-brainstorm .claude/skills/scholar-brainstorm && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scholar-brainstorm
GitHub stars
168
Token cost
~4.2k tokens
SKILL.md length
1,613 words
Files
5 (incl. references)
Skills in repo
30
Repo updated
First seen
Licence
Custom licence

At a glance

Generate research questions from existing materials — codebooks, survey questionnaires, datasets, or published papers/abstracts.

  • Works in 3 steps: DATA — user provides data files (.csv,… → MATERIALS — user provides… → PAPER — user provides a published paper…
  • Tasks that involve Brainstorming
  • SKILL.md covers Arguments, Setup, Mode Detection and Primary Goal, plus 5 more sections
  • Calls bash

What it does

Scholar Brainstorm is an agent skill from joshzyj/open-scholar-skill. Generate research questions from existing materials — codebooks, survey questionnaires, datasets, or published papers/abstracts. Three modes: DATA (data files with safety scan + empirical signal tests), MATERIALS (codebook/questionnaire only with theory-driven ranking), PAPER (published paper/abstract → follow-up research ideas via SciThinker-30B + multi-agent evaluation). Auto-detects mode from file extensions. Explores the data landscape and proposes a ranked Top 10 list of publishable research questions using…

Its SKILL.md is about 4.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files (for example `references/brainstorm-patterns.md`, `references/mode-data.md` and `references/mode-paper.md`).

It sits in Agent Workflows, covering Brainstorming, Hypothesis generation and Model hubs and datasets. It works with Hugging Face. The repository describes itself as: Open scholar skill, a claude code plugin, for academic research.

When your agent uses it

  • Tasks that involve Brainstorming
  • Tasks that involve Hypothesis generation
  • Tasks that involve Model hubs and datasets

Example prompts

  • “/scholar-brainstorm”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. DATA — user provides data files (.csv, .dta, .rds, etc.) → safety scan + empirical signal tests
  2. MATERIALS — user provides codebooks/questionnaires without data → theory-driven ranking
  3. PAPER — user provides a published paper PDF, DOI, or pasted abstract → follow-up research ideation via SciThinker + Claude brainstorming

What it can do on your machine

Read from SKILL.md and the folder at commit 6e5ac8e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • bash

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scholar Brainstorm loads about 4.2k tokens when it runs, and up to ~24k if it reads all its reference files. Until then it costs about 171 tokens; SKILL.md has 1,613 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~171
When it runs · the whole SKILL.md, loaded when a task matches
~4.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~24k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 1,613 words (~4,202 tokens).

“You are a senior social scientist who discovers publishable research questions by deeply exploring codebooks, questionnaires, and datasets. Your approach is bottom-up: start from what the data contains, then build theoretically grounded questions.”

— opening of SKILL.md by joshzyj, Custom licence
name
scholar-brainstorm
tools
Read, Bash, WebSearch, WebFetch, Write, Agent, Glob, Grep
argument-hint
[path to codebook/questionnaire/data file(s)/paper PDF] [optional: field, population, target journal] — e.g., 'paper.pdf for NHB' or 'brainstorm from…
user-invocable
true

Read the full SKILL.md on GitHub

Files

SKILL.md and 4 other files (references) in .claude/skills/scholar-brainstorm of joshzyj/open-scholar-skill.

  • SKILL.md
  • references/brainstorm-patterns.md
  • references/mode-data.md
  • references/mode-paper.md
  • references/shared-evaluation.md

Open the folder on GitHubat commit 6e5ac8e

Compare with similar skills

Scholar Brainstorm next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scholar Brainstorm compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scholar Brainstorm this skilljoshzyj/open-scholar-skill168—~4.2kAutomated safety check: PassCustom licence
Ideer Daily PaperAI45Lab/iDeer416—~2.3kAutomated safety check: NotesAGPL-3.0
Idea Generationvoidful/academic-skills134—~1.6kAutomated safety check: PassMIT
Research Ideationmaxwell2732/paper-replicate-agent-demo1371 repos~914Automated safety check: PassNone
Interview Mepedrohcgs/claude-code-my-workflow1.7k—~1.8kAutomated safety check: PassMIT
Research Ideationpedrohcgs/claude-code-my-workflow1.7k—~1.7kAutomated safety check: PassMIT

Similar skills

  • Ideer Daily Paper

    AI45Lab/iDeer

    Daily paper/repo digest where YOU are the reader. An agent skill from AI45Lab/iDeer.

    416 GitHub stars~2.3k tokensUpdated 2 mo ago
    Research & ScienceAuto-check: notes
  • Idea Generation

    voidful/academic-skills

    學術研究的 Idea 產生技能——從發散到收斂,系統化地產出高品質研究構想。當使用者想腦力激盪研究方向、找新 research idea、或問「我接下來可以做什麼研究」時,一定要使用此技能。觸發詞包括:brainstorm、想 idea、研究方向、下一步做什麼、有什麼可以研究的、找 gap、research proposal。適用於任何階段的學術研究構想生成。

    134 GitHub stars~1.6k tokensUpdated 6 mo ago
    Agent WorkflowsAuto-check passed
  • Research Ideation

    maxwell2732/paper-replicate-agent-demo

    Generate structured research questions, testable hypotheses, and empirical strategies from a topic or dataset

    137 GitHub starsUsed in 1 repo~914 tokens
    Agent WorkflowsAuto-check passed
  • Interview Me

    pedrohcgs/claude-code-my-workflow

    Interactive interview that formalizes a fuzzy research idea into a structured spec (RQ, hypotheses, identification, data needs, empirical strategy).

    1.7k GitHub stars~1.8k tokensUpdated 12 days ago
    Agent WorkflowsAuto-check passed
  • Research Ideation

    pedrohcgs/claude-code-my-workflow

    Generate structured research questions, testable hypotheses, and candidate empirical strategies from a topic, phenomenon, or dataset description.

    1.7k GitHub stars~1.7k tokensUpdated 12 days ago
    Agent WorkflowsAuto-check passed
  • Light Idea Generation

    Light0305/Light-skills

    Light 科研主线第 3 步·提 idea:从模糊方向/数据/文献结构化发散(激发算子系统生成,不是泛泛头脑风暴) → 产值得做且做得成的分层候选 idea(moonshot 冲刺/solid 稳妥/safe 保底),每个必答为什么值得做·创新点· 比现有强在哪·解决什么具体问题·能投什么层次,且提出时就自带撞车前置自查(最像的前作+delta,吃上游 literature-search…

    640 GitHub stars~4.6k tokensUpdated 3 mo ago
    Agent WorkflowsAuto-check passed

More from joshzyj/open-scholar-skill

All 30 skills in this repo
  • Scholar Annotate

    joshzyj/open-scholar-skill

    Turn unstructured text into validated, structured variables at corpus scale with LLMs: codebook design, dev/gold-set construction, DSPy prompt optimization, a hard reliability gate (Cohen κ ≥ 0.70)…

    168 GitHub stars~4.6k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar Auto Research

    joshzyj/open-scholar-skill

    Stable, deterministic social-science research-paper pipeline from idea or data to verified manuscript, citations, replication package, and final md/docx/tex/pdf outputs.

    168 GitHub stars~21k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar RAG

    joshzyj/open-scholar-skill

    Build and query a local vector database + GraphRAG over your entire reference library (Zotero or a PDF folder) for literature review.

    168 GitHub stars~7.4k tokensUpdated 21 days ago
    Auto-check: notes
  • Scholar Causal

    joshzyj/open-scholar-skill

    Comprehensive causal inference toolkit for social science research.

    168 GitHub stars~10k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar Data

    joshzyj/open-scholar-skill

    Comprehensive open data directory (100+ datasets across 14 categories) with auto-fetch capability, plus data collection instrument design, variable dictionaries, data management, IRB materials, and…

    168 GitHub stars~23k tokensUpdated 21 days ago
    Auto-check: notes
  • Scholar Eda

    joshzyj/open-scholar-skill

    Conduct exploratory data analysis (EDA) before hypothesis testing.

    168 GitHub stars~12k tokensUpdated 21 days ago
    Auto-check passed

Works with

Questions about Scholar Brainstorm

What does Scholar Brainstorm do?

Generate research questions from existing materials — codebooks, survey questionnaires, datasets, or published papers/abstracts. Scholar Brainstorm is an agent skill from joshzyj/open-scholar-skill. Generate research questions from existing materials — codebooks, survey questionnaires, datasets, or published papers/abstracts.

When should I use Scholar Brainstorm?

Scholar Brainstorm fits situations like: tasks that involve Brainstorming; tasks that involve Hypothesis generation; tasks that involve Model hubs and datasets.

How do I install Scholar Brainstorm in Claude Code?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-brainstorm -a claude-code`. Or copy the skill folder (.claude/skills/scholar-brainstorm in joshzyj/open-scholar-skill) into .claude/skills/scholar-brainstorm in your project. Claude Code loads it when a task matches its description.

How do I install Scholar Brainstorm in Codex?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-brainstorm -a codex`. Or copy the skill folder (.claude/skills/scholar-brainstorm in joshzyj/open-scholar-skill) into .agents/skills/scholar-brainstorm in your project. Codex loads it when a task matches its description.

Can I use Scholar Brainstorm in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add joshzyj/open-scholar-skill --skill scholar-brainstorm -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scholar-brainstorm, .gemini/skills/scholar-brainstorm, .github/skills/scholar-brainstorm and .opencode/skills/scholar-brainstorm in your project.

What does Scholar Brainstorm need to run?

Going by SKILL.md and its folder, Scholar Brainstorm needs the command-line tools its instructions call (bash).

Does Scholar Brainstorm access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Scholar Brainstorm safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scholar Brainstorm use?

Scholar Brainstorm has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Scholar Brainstorm use?

About 4.2k tokens (SKILL.md is roughly 17k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 20k tokens, read only when the agent opens those files.

What are the alternatives to Scholar Brainstorm?

Skills that share tags, products or a category with Scholar Brainstorm: Ideer Daily Paper (AI45Lab/iDeer, 416 stars), Idea Generation (voidful/academic-skills, 134 stars), Research Ideation (maxwell2732/paper-replicate-agent-demo, 137 stars) and Interview Me (pedrohcgs/claude-code-my-workflow, 1.7k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scholar Brainstorm?

joshzyj (a GitHub user) maintains it in joshzyj/open-scholar-skill, which has 168 GitHub stars. The repository holds 30 skills in this directory. The repository was last updated on September 18, 2026.

Source: joshzyj/open-scholar-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.