Agent skill

Scholar Safety

by joshzyj in joshzyj/open-scholar-skill

Real-time data privacy and leakage protection layer for AI-assisted research.

Custom licenceAuto-check passedLegal & Compliance

Install Scholar Safety

skills CLI
$ npx skills add joshzyj/open-scholar-skill --skill scholar-safety -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install joshzyj/open-scholar-skill scholar-safety --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/joshzyj/open-scholar-skill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/scholar-safety .claude/skills/scholar-safety && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scholar-safety
GitHub stars
168
Token cost
~11k tokens
SKILL.md length
2,785 words
Files
4 (incl. references)
Skills in repo
30
Repo updated
First seen
Licence
Custom licence

At a glance

Real-time data privacy and leakage protection layer for AI-assisted research.

  • Works in 12 steps: 0 — Quick Gate Check (REQUIRED) → 1 — File Inventory → 2 — Detect Data Type (Quantitative vs.… → …
  • Tasks that involve Privacy and GDPR
  • SKILL.md covers Arguments, Dispatch Table, Setup and CRITICAL OPERATING RULE, plus 5 more sections
  • Calls bash and python3

What it does

Scholar Safety is an agent skill from joshzyj/open-scholar-skill. Real-time data privacy and leakage protection layer for AI-assisted research. Before any file is read by Claude Code and transmitted to Anthropic's API, proactively scans it locally for sensitive content (PII, HIPAA, IRB-protected, restricted licensed data) using local Bash pattern matching — so the sensitive data itself never enters the AI context during the scan. Issues tiered warnings (green/yellow/red), requests explicit user permission before transmitting any sensitive data, and offers safe alternatives…

Its SKILL.md is about 11k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `references/permission-templates.md`, `references/safe-mode-protocols.md` and `references/sensitivity-patterns.md`).

It sits in Legal & Compliance, covering Privacy and GDPR and Healthcare and finance regulation. It works with Bash. The repository describes itself as: Open scholar skill, a claude code plugin, for academic research.

When your agent uses it

  • Tasks that involve Privacy and GDPR
  • Tasks that involve Healthcare and finance regulation

Example prompts

  • “/scholar-safety”

Workflow steps

12 steps, taken from the step headings in SKILL.md.

  1. 0 — Quick Gate Check (REQUIRED)
  2. 1 — File Inventory
  3. 2 — Detect Data Type (Quantitative vs. Qualitative)
  4. 3 — Local Sensitivity Pattern Scan
  5. 4 — Risk Classification
  6. 5 — Safety Alert Output
  7. 1 — Parse the proposed operation
  8. 2 — Run MODE 1 scan on each file
  9. 3 — Gate decision
  10. 4 — Log the gate outcome
  11. 1 — Project Data Profile
  12. 2 — Generate the Safety Protocol Document

What it can do on your machine

Read from SKILL.md and the folder at commit 6e5ac8e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • bash
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scholar Safety loads about 11k tokens when it runs, and up to ~21k if it reads all its reference files. Until then it costs about 252 tokens; SKILL.md has 2,785 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~252
When it runs · the whole SKILL.md, loaded when a task matches
~11k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~21k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 2,785 words (~11,170 tokens).

“You are the data safety guardian for AI-assisted research. Your job is to intercept potentially risky operations before any sensitive data reaches Anthropic's servers, warn the researcher with precise details, and give them real options — not just vague caution.”

— opening of SKILL.md by joshzyj, Custom licence
name
scholar-safety
tools
Bash, Read, Write
argument-hint
[scan|gate|protocol|status|level] [file path / operation description / level name] [optional: data type, project name, journal target]
user-invocable
true

Read the full SKILL.md on GitHub

Files

SKILL.md and 3 other files (references) in .claude/skills/scholar-safety of joshzyj/open-scholar-skill.

  • SKILL.md
  • references/permission-templates.md
  • references/safe-mode-protocols.md
  • references/sensitivity-patterns.md

Open the folder on GitHubat commit 6e5ac8e

Compare with similar skills

Scholar Safety next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scholar Safety compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scholar Safety this skilljoshzyj/open-scholar-skill168—~11kAutomated safety check: PassCustom licence
HIPAA Safe Harbor Coverage Auditmaziyarpanahi/openmed5.5k—~1.7kAutomated safety check: PassApache-2.0
Hipaa ComplianceSushegaad/Claude-Skills-Governance-Risk-and-Compliance9461 repos~2.3kAutomated safety check: PassMIT
Audit Reportharness/harness-skills115—~1.3kAutomated safety check: PassApache-2.0
Anne WojcickiK-Dense-AI/mimeographs129—~1.5kAutomated safety check: PassMIT
Dpa Checklist ReviewLegalQuants/lq-ai150—~3.7kAutomated safety check: PassApache-2.0

Similar skills

  • Checks OpenMed de-identified clinical text against the 18 HIPAA Safe Harbor identifier categories and reports gaps and residual re-identification risk.

    5.5k GitHub stars~1.7k tokensUpdated today
    Legal & ComplianceAuto-check passed
  • Hipaa Compliance

    Sushegaad/Claude-Skills-Governance-Risk-and-Compliance

    Expert HIPAA compliance assistant for healthcare and software contexts.

    946 GitHub starsUsed in 1 repo~2.3k tokens
    Legal & ComplianceAuto-check passed
  • Audit Report

    harness/harness-skills

    Generate audit reports and compliance trails using Harness audit trail data via MCP v2 tools.

    115 GitHub stars~1.3k tokensUpdated 4 days ago
    Legal & ComplianceAuto-check passed
  • Anne Wojcicki

    K-Dense-AI/mimeographs

    Applies the strategic frameworks and mental models of Anne Wojcicki, co-founder and CEO of 23andMe.

    129 GitHub stars~1.5k tokensUpdated 1 mo ago
    Legal & ComplianceAuto-check passed
  • Dpa Checklist Review

    LegalQuants/lq-ai

    A skill your agent uses when the user provides a Data Processing Agreement, Data Processing Addendum, or HIPAA Business Associate Agreement and asks whether it contains the terms required under the…

    150 GitHub stars~3.7k tokensUpdated today
    Legal & ComplianceAuto-check passed
  • Auditing Deidentification Runs

    maziyarpanahi/openmed

    Produce a signed, reproducible, no-PHI audit trail for an OpenMed de-identification run via deidentify(audit=True).

    5.5k GitHub stars~1.8k tokensUpdated today
    Legal & ComplianceAuto-check passed

More from joshzyj/open-scholar-skill

All 30 skills in this repo
  • Scholar Annotate

    joshzyj/open-scholar-skill

    Turn unstructured text into validated, structured variables at corpus scale with LLMs: codebook design, dev/gold-set construction, DSPy prompt optimization, a hard reliability gate (Cohen κ ≥ 0.70)…

    168 GitHub stars~4.6k tokensUpdated 22 days ago
    Auto-check passed
  • Scholar Auto Research

    joshzyj/open-scholar-skill

    Stable, deterministic social-science research-paper pipeline from idea or data to verified manuscript, citations, replication package, and final md/docx/tex/pdf outputs.

    168 GitHub stars~21k tokensUpdated 22 days ago
    Auto-check passed
  • Scholar RAG

    joshzyj/open-scholar-skill

    Build and query a local vector database + GraphRAG over your entire reference library (Zotero or a PDF folder) for literature review.

    168 GitHub stars~7.4k tokensUpdated 22 days ago
    Auto-check: notes
  • Scholar Causal

    joshzyj/open-scholar-skill

    Comprehensive causal inference toolkit for social science research.

    168 GitHub stars~10k tokensUpdated 22 days ago
    Auto-check passed
  • Scholar Data

    joshzyj/open-scholar-skill

    Comprehensive open data directory (100+ datasets across 14 categories) with auto-fetch capability, plus data collection instrument design, variable dictionaries, data management, IRB materials, and…

    168 GitHub stars~23k tokensUpdated 22 days ago
    Auto-check: notes
  • Scholar Eda

    joshzyj/open-scholar-skill

    Conduct exploratory data analysis (EDA) before hypothesis testing.

    168 GitHub stars~12k tokensUpdated 22 days ago
    Auto-check passed

Works with

Questions about Scholar Safety

What does Scholar Safety do?

Real-time data privacy and leakage protection layer for AI-assisted research. Scholar Safety is an agent skill from joshzyj/open-scholar-skill. Real-time data privacy and leakage protection layer for AI-assisted research.

When should I use Scholar Safety?

Scholar Safety fits situations like: tasks that involve Privacy and GDPR; tasks that involve Healthcare and finance regulation.

How do I install Scholar Safety in Claude Code?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-safety -a claude-code`. Or copy the skill folder (.claude/skills/scholar-safety in joshzyj/open-scholar-skill) into .claude/skills/scholar-safety in your project. Claude Code loads it when a task matches its description.

How do I install Scholar Safety in Codex?

Run `npx skills add joshzyj/open-scholar-skill --skill scholar-safety -a codex`. Or copy the skill folder (.claude/skills/scholar-safety in joshzyj/open-scholar-skill) into .agents/skills/scholar-safety in your project. Codex loads it when a task matches its description.

Can I use Scholar Safety in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add joshzyj/open-scholar-skill --skill scholar-safety -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scholar-safety, .gemini/skills/scholar-safety, .github/skills/scholar-safety and .opencode/skills/scholar-safety in your project.

What does Scholar Safety need to run?

Going by SKILL.md and its folder, Scholar Safety needs the command-line tools its instructions call (bash and python3).

Does Scholar Safety access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Scholar Safety safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scholar Safety use?

Scholar Safety has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Scholar Safety use?

About 11k tokens (SKILL.md is roughly 45k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 10k tokens, read only when the agent opens those files.

What are the alternatives to Scholar Safety?

Skills that share tags, products or a category with Scholar Safety: HIPAA Safe Harbor Coverage Audit (maziyarpanahi/openmed, 5.5k stars), Hipaa Compliance (Sushegaad/Claude-Skills-Governance-Risk-and-Compliance, 946 stars), Audit Report (harness/harness-skills, 115 stars) and Anne Wojcicki (K-Dense-AI/mimeographs, 129 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scholar Safety?

joshzyj (a GitHub user) maintains it in joshzyj/open-scholar-skill, which has 168 GitHub stars. The repository holds 30 skills in this directory. The repository was last updated on September 18, 2026.

Source: joshzyj/open-scholar-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.