Agent skill

Scholar Compute

by joshzyj in joshzyj/open-scholar-skill

Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction…

Custom licenceAuto-check passedAI & LLM Engineering

Install Scholar Compute

skills CLI
$ npx skills add joshzyj/open-scholar-skill --skill scholar-compute -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install joshzyj/open-scholar-skill scholar-compute --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/joshzyj/open-scholar-skill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/scholar-compute .claude/skills/scholar-compute && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scholar-compute
GitHub stars
168
Token cost
~15k tokens
SKILL.md length
4,632 words
Files
50 (incl. references)
Skills in repo
30
Repo updated
First seen
Licence
Custom licence

At a glance

Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction…

  • Works in 4 steps: Consolidate early. Write the module's… → Review: load and execute… → Fix loop:… → …
  • Tasks that involve Test data and fixtures
  • SKILL.md covers Arguments and Module Routing, MODULE 0: Setup (All Modules), MODULE 0.5: Data Safety Gate… and MODULE 0.7: Pre-Scale Review…, plus 5 more sections
  • Calls bash, node and python3

What it does

Scholar Compute is an agent skill from joshzyj/open-scholar-skill. Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction, XLM-R/mBERT); ML (supervised, Double ML, Causal Forests, Bayesian brms/Stan, conformal prediction); network analysis (ERGMs, SAOMs, relational event models, GNN via PyTorch Geometric); agent-based modeling (Mesa 3.x, NetLogo, LLM agents, ODD, SALib); computer vision (DINOv2, CLIP, ViT, multimodal LLMs, VideoMAE); LLM…

Its SKILL.md is about 15k tokens, which your agent loads only when the skill is triggered. The skill folder holds 52 other files, including reference files (for example `references/computer-vision.md`, `references/ml-methods.md` and `references/module-01-nlp.md`).

It sits in AI & LLM Engineering, covering Test data and fixtures, Geospatial analysis and Embeddings. It works with PyTorch. The repository describes itself as: Open scholar skill, a claude code plugin, for academic research.

When your agent uses it

  • Tasks that involve Test data and fixtures
  • Tasks that involve Geospatial analysis
  • Tasks that involve Embeddings

Example prompts

  • “/scholar-compute”

Requirements

  • Docker

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Consolidate early. Write the module's planned pipeline into its self-contained script(s) NOW (the Script Archive Protocol's "After Each…
  2. Review: load and execute scholar-code-review in pre-execution mode, fast 3 subset (statistics data-handling correctness) — 3 SEPARATE…
  3. Fix loop: _shared/code-review-fix-loop.md on CRITICAL findings (max 2 iterations; re-review per scholar-code-review Step 6).
  4. Gate, then execute the reviewed scripts directly

What it can do on your machine

Read from SKILL.md and the folder at commit 6e5ac8e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • bash
    • node
    • python3
    • claude

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scholar Compute loads about 15k tokens when it runs, and up to ~134k if it reads all its reference files. Until then it costs about 258 tokens; SKILL.md has 4,632 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~258
When it runs · the whole SKILL.md, loaded when a task matches
~15k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~134k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 4,632 words (~14,578 tokens).

“You are an expert computational social scientist. You run executable analyses, validate results, and produce publication-quality outputs with prose ready for the Methods and Results sections. You meet the reproducibility standards of Nature Computational Science, Science Advances, and top sociology…”

— opening of SKILL.md by joshzyj, Custom licence
name
scholar-compute
tools
Read, Bash, Write, WebSearch, Agent
argument-hint
[text|network|ml|abm|reproduce|spatial|bayesian|dsl|audio|life2vec] [description of data and research question]
user-invocable
true

Read the full SKILL.md on GitHub

Files

SKILL.md and 49 other files (references) in .claude/skills/scholar-compute of joshzyj/open-scholar-skill.

  • SKILL.md
  • references/computer-vision.md
  • references/ml-methods.md
  • references/module-01-nlp.md
  • references/module-02-ml.md
  • references/module-03-network.md
  • references/module-04-abm.md
  • references/module-05-reproducibility.md
  • references/module-06-cv.md
  • references/module-07-llm-analysis.md
  • references/module-08-synthetic-data.md
  • references/module-09-geospatial.md
  • references/module-10-audio.md
  • references/module-11-life2vec.md
  • references/network-analysis.md
  • references/nlp-pipeline.md
  • references/prompt-optimization-demo/README.md
  • references/prompt-optimization-demo/dspy-demo/CODEBOOK.md
  • … and 32 more

Open the folder on GitHubat commit 6e5ac8e

Compare with similar skills

Scholar Compute next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scholar Compute compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scholar Compute this skilljoshzyj/open-scholar-skill168—~15kAutomated safety check: PassCustom licence
CLIP Image-Text MatchingOrchestra-Research/AI-Research-SKILLs13k7 repos~1.7kAutomated safety check: PassMIT
Sentence Transformers EmbeddingsOrchestra-Research/AI-Research-SKILLs13k2 repos~1.6kAutomated safety check: PassMIT
Video Camera DemosVectorSpaceLab/AREX-Skill330—~626Automated safety check: PassApache-2.0
Deep Learning NLPDrchronx/ai-agent-research-starter-kit137—~516Automated safety check: PassCustom licence
Text Analysis BasicDrchronx/ai-agent-research-starter-kit137—~737Automated safety check: PassCustom licence

Similar skills

  • CLIP Image-Text Matching

    Orchestra-Research/AI-Research-SKILLs

    Explains OpenAI's CLIP model for zero-shot image classification, image-text similarity, semantic image search and content moderation, with install steps and code patterns.

    13k GitHub starsUsed in 7 repos~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Sentence Transformers Embeddings

    Orchestra-Research/AI-Research-SKILLs

    Generates text embeddings locally with the sentence-transformers library for RAG, semantic search, clustering and similarity, with model picks for general, multilingual and legal text.

    13k GitHub starsUsed in 2 repos~1.6k tokens
    AI & LLM EngineeringAuto-check passed
  • Video Camera Demos

    VectorSpaceLab/AREX-Skill

    Guide safe video-file, webcam, and optional half-precision demo use for pytorch-yolo-v3.

    330 GitHub stars~626 tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Deep Learning NLP

    Drchronx/ai-agent-research-starter-kit

    Paddle-based deep learning workflows from the course materials, including DNN/RNN text-style baselines and the CNN/LeNet image classification case using folder-labeled digit images.

    137 GitHub stars~516 tokensUpdated 4 mo ago
    AI & LLM EngineeringAuto-check passed
  • Text Analysis Basic

    Drchronx/ai-agent-research-starter-kit

    Basic Chinese NLP and text analysis workflows for PDF text/table extraction, jieba tokenization, word and sentence frequency, word clouds, TF-IDF, Word2Vec, sentence embeddings, and text similarity.

    137 GitHub stars~737 tokensUpdated 4 mo ago
    AI & LLM EngineeringAuto-check passed
  • AI ML Skills

    wentorai/research-plugins

    27 ai & machine learning skills. An agent skill from wentorai/research-plugins.

    298 GitHub starsUsed in 1 repo~993 tokens
    AI & LLM EngineeringAuto-check passed

More from joshzyj/open-scholar-skill

All 30 skills in this repo
  • Scholar Annotate

    joshzyj/open-scholar-skill

    Turn unstructured text into validated, structured variables at corpus scale with LLMs: codebook design, dev/gold-set construction, DSPy prompt optimization, a hard reliability gate (Cohen κ ≥ 0.70)…

    168 GitHub stars~4.6k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar Auto Research

    joshzyj/open-scholar-skill

    Stable, deterministic social-science research-paper pipeline from idea or data to verified manuscript, citations, replication package, and final md/docx/tex/pdf outputs.

    168 GitHub stars~21k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar RAG

    joshzyj/open-scholar-skill

    Build and query a local vector database + GraphRAG over your entire reference library (Zotero or a PDF folder) for literature review.

    168 GitHub stars~7.4k tokensUpdated 21 days ago
    Auto-check: notes
  • Scholar Causal

    joshzyj/open-scholar-skill

    Comprehensive causal inference toolkit for social science research.

    168 GitHub stars~10k tokensUpdated 21 days ago
    Auto-check passed
  • Scholar Data

    joshzyj/open-scholar-skill

    Comprehensive open data directory (100+ datasets across 14 categories) with auto-fetch capability, plus data collection instrument design, variable dictionaries, data management, IRB materials, and…

    168 GitHub stars~23k tokensUpdated 21 days ago
    Auto-check: notes
  • Scholar Eda

    joshzyj/open-scholar-skill

    Conduct exploratory data analysis (EDA) before hypothesis testing.

    168 GitHub stars~12k tokensUpdated 21 days ago
    Auto-check passed

Works with

Questions about Scholar Compute

What does Scholar Compute do?

Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction…. Scholar Compute is an agent skill from joshzyj/open-scholar-skill.

When should I use Scholar Compute?

Scholar Compute fits situations like: tasks that involve Test data and fixtures; tasks that involve Geospatial analysis; tasks that involve Embeddings.

How do I install Scholar Compute in Claude Code?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-compute -a claude-code`. Or copy the skill folder (.claude/skills/scholar-compute in joshzyj/open-scholar-skill) into .claude/skills/scholar-compute in your project. Claude Code loads it when a task matches its description.

How do I install Scholar Compute in Codex?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-compute -a codex`. Or copy the skill folder (.claude/skills/scholar-compute in joshzyj/open-scholar-skill) into .agents/skills/scholar-compute in your project. Codex loads it when a task matches its description.

Can I use Scholar Compute in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add joshzyj/open-scholar-skill --skill scholar-compute -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scholar-compute, .gemini/skills/scholar-compute, .github/skills/scholar-compute and .opencode/skills/scholar-compute in your project.

What does Scholar Compute need to run?

Going by SKILL.md and its folder, Scholar Compute needs the command-line tools its instructions call (bash, node, python3 and claude). Our summary lists: Docker.

Does Scholar Compute access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Scholar Compute safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scholar Compute use?

Scholar Compute has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Scholar Compute use?

About 15k tokens (SKILL.md is roughly 58k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 119k tokens, read only when the agent opens those files.

What are the alternatives to Scholar Compute?

Skills that share tags, products or a category with Scholar Compute: CLIP Image-Text Matching (Orchestra-Research/AI-Research-SKILLs, 13k stars), Sentence Transformers Embeddings (Orchestra-Research/AI-Research-SKILLs, 13k stars), Video Camera Demos (VectorSpaceLab/AREX-Skill, 330 stars) and Deep Learning NLP (Drchronx/ai-agent-research-starter-kit, 137 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scholar Compute?

joshzyj (a GitHub user) maintains it in joshzyj/open-scholar-skill, which has 168 GitHub stars. The repository holds 30 skills in this directory. The repository was last updated on September 18, 2026.

Source: joshzyj/open-scholar-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.