Agent skill

Librarian

by jdforsythe in jdforsythe/forge

Reviews, curates, and maintains the Forge library of agents, skills, and templates.

MITAuto-check passedAgent Workflows

Install Librarian

skills CLI
$ npx skills add jdforsythe/forge --skill librarian -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jdforsythe/forge librarian --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jdforsythe/forge.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/librarian .claude/skills/librarian && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
librarian
GitHub stars
151
Token cost
~5.1k tokens
SKILL.md length
1,943 words
Files
4 (incl. references)
Skills in repo
4
Repo updated
First seen
Licence
MIT

At a glance

Reviews, curates, and maintains the Forge library of agents, skills, and templates.

  • Works in 4 steps: Load Inventory → Run Checks → Produce Report → …
  • The user wants to review the library
  • SKILL.md covers Expert Vocabulary Payload, Anti-Pattern Watchlist, Behavioral Instructions and Output Format, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Librarian is an agent skill from jdforsythe/forge. Reviews, curates, and maintains the Forge library of agents, skills, and templates. Performs deduplication analysis, staleness detection, quality promotion, and orphan reference checking. Produces structured review reports with actionable recommendations for merging, archiving, or promoting library items. Use this skill when the user wants to review the library, clean up agents or skills, check what's available, find duplicates, trim unused items, see library statistics, or says "what's in my library?" Also…

Its SKILL.md is about 5.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `references/review-criteria.md`).

It sits in Agent Workflows, covering Data cleaning, Building AI agents and Skill authoring. The repository describes itself as: Skills for creating high quality skills and agents. The licence is MIT.

When your agent uses it

  • The user wants to review the library
  • Clean up agents
  • Check whats available
  • Find duplicates

Example prompts

  • “s in my library?”
  • “/librarian”

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Load Inventory
  2. Run Checks
  3. Produce Report
  4. Execute Changes

What it can do on your machine

Read from SKILL.md and the folder at commit b192c5c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Librarian loads about 5.1k tokens when it runs, and up to ~8.6k if it reads all its reference files. Until then it costs about 185 tokens; SKILL.md has 1,943 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~185
When it runs · the whole SKILL.md, loaded when a task matches
~5.1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~8.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from jdforsythe/forge at commit b192c5c, republished under its MIT licence (© jdforsythe). 1,943 words, ~5,133 tokens.

Download SKILL.mdSave it as .claude/skills/librarian/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
librarian
description
Reviews, curates, and maintains the Forge library of agents, skills, and templates. Performs deduplication analysis, staleness detection, quality promotion, and orphan reference checking. Produces structured review reports with actionable recommendations for merging, archiving, or promoting library items. Use this skill when the user wants to review the library, clean up agents or skills, check what's available, find duplicates, trim unused items, see library statistics, or says "what's in my library?" Also triggers on scheduled review intervals or when the library grows beyond 20 items. Do NOT use for creating new agents (use Agent Creator), creating skills (use Skill Creator), or planning teams (use Mission Planner).

Librarian

The curation and maintenance engine for the Forge library. Audits inventory, detects quality issues, and produces actionable review reports to keep the library lean, accurate, and useful.


Expert Vocabulary Payload

Deduplication & Similarity: deduplication, semantic similarity, tag overlap, merge candidate, near-duplicate detection, Jaccard coefficient, synonym clustering Lifecycle & Quality: quality promotion, lifecycle stage, quality tier, curation, retention policy, staleness threshold, maturation criteria Inventory & Structure: orphan detection, dependency graph, index integrity, catalog maintenance, inventory audit, library hygiene, reference traversal Usage & Telemetry: usage frequency, usage telemetry, usage decay, access recency, impact-weighted frequency, archive candidate


Anti-Pattern Watchlist

Hoarder Library
  • Detection: Library exceeds 50 items with fewer than 20% showing any usage in the past 90 days. The index grows monotonically — items are added but never removed.
  • Why it fails: Search and matching degrade as irrelevant items dilute results. Users lose trust in library quality when most items are stale or broken. Cognitive overhead increases with every unused entry.
  • Resolution: Flag all items with zero usage in the past 90 days for review. Present the list sorted by staleness (oldest unused first). Recommend archive for items with no usage data at all and no recent modification.
Premature Deletion
  • Detection: Item has low usage frequency (fewer than 3 uses total) but the uses that exist are high-impact — referenced in complex blueprints, used for critical domains, or explicitly requested by name.
  • Why it fails: Specialized items (security auditors, compliance reviewers, incident responders) are infrequently needed but irreplaceable when they are. Deleting them forces recreation from scratch at the worst possible time.
  • Resolution: Before recommending removal of any low-frequency item, check usage context. If any usage occurred within a complex team blueprint or critical domain, flag as "low-frequency but high-value" and recommend keeping. Apply a higher staleness threshold (180 days instead of 90) for items in critical domains like security, compliance, and incident response.
Duplicate Blindness
  • Detection: Two or more items share more than 70% tag overlap, or their descriptions are semantically similar (same domain, same deliverables, similar SOPs). Common pattern: items created at different times for the same purpose with slightly different names (e.g., "api-developer" and "backend-engineer").
  • Why it fails: Users get inconsistent results depending on which duplicate is matched. Quality improvements to one copy do not propagate to the other. Library size inflates without capability gain.
  • Resolution: Recommend merge. Keep the version with the higher quality tier. If tied, keep the more recently modified version. If still tied, keep the version with more usage. Present both items side-by-side so the user can make the final call.
Orphan Accumulation
  • Detection: A template's roles array references agent names that do not exist in index.json agents. This happens when agents are deleted or renamed without updating the templates that reference them.
  • Why it fails: Templates that reference nonexistent agents will fail at runtime. Users who load the template get errors or incomplete teams. Trust in the library erodes.
  • Resolution: For each orphaned reference, present two options: (1) create the missing agent, or (2) update the template to remove or replace the reference. Flag the severity — a template with one orphan out of four roles is recoverable; a template where all roles are orphaned should be archived.
Quality Stagnation
  • Detection: Items remain at "untested" quality tier for more than 30 days despite having one or more recorded uses. Items at "tested" for more than 60 days with continued active usage and no modifications.
  • Why it fails: Quality tiers lose meaning if they never advance. Users cannot distinguish battle-tested agents from first drafts. The promotion system exists to build confidence — stagnation defeats that purpose.
  • Resolution: Automatically recommend promotion based on usage thresholds. Present the evidence (usage count, time at current tier, modification history) and let the user approve the promotion.

Behavioral Instructions

Phase 1: Load Inventory
  1. Read library/index.json for the current catalog. PARSE: total item count, items by type (agents, skills, templates), items by domain, items by quality tier. IF index is empty or missing: Report "Library is empty. Nothing to review." and STOP. OUTPUT: Inventory summary object.

  2. Read library/usage-log.jsonl for usage data. PARSE: per-item usage count, last-used timestamp, action history (loaded, created, modified, archived, promoted). IF usage log is empty or missing: Note "No usage data available. Staleness and promotion checks will be limited to metadata only." OUTPUT: Usage data map keyed by item path.

Phase 2: Run Checks
  1. Deduplication check. FOR each pair of items within the same type (agent-agent, skill-skill, template-template): a. Calculate tag overlap: count of shared tags divided by count of union of tags (Jaccard coefficient). b. IF tag overlap > 0.70: Flag as merge candidate. c. IF tag overlap > 0.50 AND same domain: Flag as potential merge candidate (review manually). d. Compare descriptions for semantic similarity: same domain + similar deliverables or purpose. OUTPUT: List of merge candidate pairs with overlap scores.

  2. Staleness check. FOR each item in the index: a. Find most recent usage log entry (any action type). b. Calculate days since last activity. c. IF no usage data exists AND item quality is "untested": Flag as stale (unknown — never used). d. IF last activity > 90 days (default threshold): Flag as stale. e. IF last activity > 90 days BUT item is in a critical domain (security, compliance, incident-response): Apply extended threshold of 180 days instead. f. Usage staleness and evidence staleness are independent axes — an item can be actively used and still cite outdated research. See check 8 (Evidence-currency check) below; do not skip it just because an item passes this check. OUTPUT: List of stale items with days-since-last-activity and recommended action.

  3. Quality promotion check. FOR each item with a quality tier: a. Count total "loaded" actions in usage log. b. Check for any "modified" actions in usage log. c. IF quality is "untested" AND loaded count >= 5: Recommend promotion to "tested." d. IF quality is "tested" AND loaded count >= 10 AND no "modified" actions after initial creation: Recommend promotion to "iterated." e. IF quality is "iterated" AND loaded count >= 20 AND user has explicitly reviewed: Eligible for "curated" (requires manual approval). f. Evidence-currency issues (check 8) do not block a usage-based promotion, but surface them alongside the recommendation — a promotion candidate with unverified claims still needs a correction pass. OUTPUT: List of promotion candidates with current tier, recommended tier, and evidence.

  4. Orphan detection. FOR each template in index: a. Read the template's roles array. b. FOR each role name: Check if an agent with that name exists in the index agents array. c. IF agent not found: Flag as orphaned reference. OUTPUT: List of orphaned references with template name, missing agent name, and severity.

  5. Core overlap check. FOR each library item: a. Compare against the core skills (mission-planner, agent-creator, skill-creator, librarian). b. IF a library item's described functionality substantially overlaps with a core skill: Flag as core overlap. OUTPUT: List of items that may duplicate core skill functionality.

  6. Evidence-currency check. Evidence baseline: July 2026, targeting Claude 4.5+/5. Apply this check to every item that cites research, states a quantitative claim, or gives model-behavior guidance — it is independent of the usage-based staleness check above. FOR each such item: a. IF the item cites a source: Trace the citation to docs/research/source-index.md. IF it cannot be traced: Flag as unverified citation. b. IF the item states a quantitative claim tied to a citation: Confirm the number appears in the cited source. IF it does not: Flag as unsupported number. c. IF the item presents a Forge convention (team size, vocabulary payload counts, iteration caps, identity length, etc.) as a research finding rather than a labeled convention: Flag as mislabeled convention — recommend an explicit "Forge design standard" label. d. IF the item contains model-behavior guidance older than ~18 months with neither re-confirmation against current-generation evidence nor a "Forge design standard" label: Flag as needing re-confirmation. e. Regardless of citation status, check the two current-model guidance rules (see references/review-criteria.md — Evidence Currency Criteria): - Reasoning-echo instruction: does the item tell an agent to narrate or print its internal reasoning? Flag if so. - Rule-stuffing: does the item substitute a long enumerated rule list for brief goal/boundary/verification framing? Flag if so. OUTPUT: List of evidence-currency issues with item name, issue type, and recommended fix. These are findings, not blockers — see step 9.

Show full SKILL.md (576 more words)Show less
Phase 3: Produce Report
  1. Compile review report. Assemble all findings into a structured markdown report (see Output Format below). Include:

    • Summary statistics
    • Merge recommendations (with which item to keep and why)
    • Archive/removal recommendations (with reason and staleness data)
    • Quality promotion recommendations (with usage evidence)
    • Orphaned references (with fix options)
    • Core overlap warnings
    • Evidence-currency issues (with issue type and recommended fix) IF no issues found: Produce a clean report with summary statistics only.
  2. Present report and WAIT for user approval. Do NOT modify any files until the user explicitly approves. IF user approves all recommendations: Proceed to Phase 4. IF user approves selectively: Execute only the approved changes. IF user rejects: STOP. No changes made.

Phase 4: Execute Changes
  1. Execute approved changes. FOR each approved merge: a. Remove the lower-quality item from index.json. b. Optionally: copy unique tags from the removed item to the kept item. c. Delete or archive the removed item's file. FOR each approved archive: a. Remove the item from index.json. b. Move the item's file to library/archive/ (create directory if needed). FOR each approved promotion: a. Update the item's quality field in index.json. FOR each approved orphan fix: a. Update the template's roles array in index.json as directed. Evidence-currency fixes are content edits (correcting a claim, adding a label, rewriting a rule-stuffed instruction) made directly in the item's file — they do not have an index.json field of their own. Update index.json updated timestamp.

  2. Log all changes. FOR each change executed: Append a usage-log.jsonl entry with:

    • ts: current ISO 8601 timestamp
    • item: path of the affected item
    • type: item type (agent, skill, template)
    • action: "archived" for removals, "promoted" for promotions, "modified" for merges, orphan fixes, and evidence-currency corrections
    • context: "librarian"

Output Format

The review report uses structured markdown:

markdown
# Library Review Report

**Date:** [ISO 8601 date]
**Total items:** [count] ([agents] agents, [skills] skills, [templates] templates)

## Summary Statistics

| Domain | Agents | Skills | Templates | Total |
|--------|--------|--------|-----------|-------|
| [domain] | [n] | [n] | [n] | [n] |
| ...    | ...    | ...    | ...       | ...   |

### Quality Distribution

| Tier | Count | Percentage |
|------|-------|------------|
| Curated | [n] | [%] |
| Iterated | [n] | [%] |
| Tested | [n] | [%] |
| Untested | [n] | [%] |

## Merge Candidates

### [Item A] + [Item B]
- **Tag overlap:** [%]
- **Recommendation:** Keep [Item A] (reason: higher quality tier / more recent / more used)
- **Action required:** Merge and remove [Item B]

## Archive Candidates

### [Item Name]
- **Last activity:** [date] ([N] days ago)
- **Total uses:** [count]
- **Reason:** [Unused for >90 days / Never used / Duplicates core skill]
- **Action required:** Archive to library/archive/

## Quality Promotions

### [Item Name]
- **Current tier:** [tier]
- **Recommended tier:** [tier]
- **Evidence:** [N] uses, [N] days at current tier, [modified/unmodified]

## Orphaned References

### Template: [template-name]
- **Missing agent:** [agent-name]
- **Options:** Create agent / Remove from template / Replace with [alternative]

## Evidence Currency Issues

### [Item Name]
- **Issue type:** [Unverified citation / Unsupported number / Mislabeled convention / Needs re-confirmation / Reasoning-echo instruction / Rule-stuffing]
- **Detail:** [what was found and where]
- **Recommended fix:** [trace or remove citation / correct the number / add "Forge design standard" label / re-confirm against current-generation evidence / remove the reasoning-echo instruction / replace with goal-boundary-verification framing]

## No Issues

[Only shown if all checks pass]
Library is clean. No merge candidates, stale items, promotions, orphans, or evidence-currency issues detected.

Examples

Example 1: Library With Issues

Scenario: Library contains 12 agents, 3 skills, and 2 templates across software and marketing domains.

Review report output:

markdown
# Library Review Report

**Date:** 2026-03-28T12:00:00Z
**Total items:** 17 (12 agents, 3 skills, 2 templates)

## Summary Statistics

| Domain | Agents | Skills | Templates | Total |
|--------|--------|--------|-----------|-------|
| software | 8 | 2 | 1 | 11 |
| marketing | 4 | 1 | 1 | 6 |

### Quality Distribution

| Tier | Count | Percentage |
|------|-------|------------|
| Curated | 0 | 0% |
| Iterated | 1 | 8% |
| Tested | 3 | 25% |
| Untested | 8 | 67% |

## Merge Candidates

### api-developer + backend-engineer
- **Tag overlap:** 78% (shared: api, rest, backend, nodejs, database, testing; unique to api-developer: openapi; unique to backend-engineer: microservices)
- **Recommendation:** Keep backend-engineer (reason: higher quality tier — tested vs untested)
- **Action required:** Merge tags and remove api-developer

### content-writer + copywriter
- **Tag overlap:** 72% (shared: writing, content, marketing, seo, editing; unique to content-writer: blog, longform; unique to copywriter: ads, conversion)
- **Recommendation:** Keep content-writer (reason: more total uses — 8 vs 3)
- **Action required:** Merge tags and remove copywriter

## Archive Candidates

### seo-analyst
- **Last activity:** 2025-11-28 (120 days ago)
- **Total uses:** 1
- **Reason:** Unused for >90 days, low total usage
- **Action required:** Archive to library/archive/

## Quality Promotions

### product-manager
- **Current tier:** untested
- **Recommended tier:** tested
- **Evidence:** 7 uses over 45 days, no modifications

### frontend-developer
- **Current tier:** untested
- **Recommended tier:** tested
- **Evidence:** 5 uses over 30 days, no modifications

### qa-engineer
- **Current tier:** tested
- **Recommended tier:** iterated
- **Evidence:** 12 uses over 60 days, unmodified since creation

## Orphaned References

### Template: marketing-campaign
- **Missing agent:** brand-strategist
- **Options:** Create brand-strategist agent / Remove from template / Replace with content-writer

## Evidence Currency Issues

### backend-engineer
- **Issue type:** Unsupported number
- **Detail:** Agent definition claims "pair review catches three times as many defects" with no citation. No source in docs/research/source-index.md states this figure.
- **Recommended fix:** Remove the claim or replace with a sourced figure; if it describes a Forge convention, label it as such instead.

### qa-engineer
- **Issue type:** Reasoning-echo instruction
- **Detail:** SOP step 4 instructs the agent to "explain your step-by-step reasoning before giving the verdict."
- **Recommended fix:** Remove the reasoning-echo instruction; rely on adaptive thinking and require only the verdict plus supporting evidence.
Example 2: Clean Library

Scenario: Library contains 6 agents, 1 skill, and 1 template. All items are actively used, no duplicates, no orphans.

Review report output:

markdown
# Library Review Report

**Date:** 2026-03-28T12:00:00Z
**Total items:** 8 (6 agents, 1 skill, 1 template)

## Summary Statistics

| Domain | Agents | Skills | Templates | Total |
|--------|--------|--------|-----------|-------|
| software | 4 | 1 | 1 | 6 |
| marketing | 2 | 0 | 0 | 2 |

### Quality Distribution

| Tier | Count | Percentage |
|------|-------|------------|
| Curated | 1 | 17% |
| Iterated | 2 | 33% |
| Tested | 2 | 33% |
| Untested | 1 | 17% |

## No Issues

Library is clean. No merge candidates, stale items, promotions, orphans, or evidence-currency issues detected.

All items have been used within the last 90 days. Quality tiers are up to date. All template references resolve to existing agents. All citations trace to docs/research/source-index.md and no item shows reasoning-echo instructions or rule-stuffing.

Environment Branching

Claude Code Environment
  • Operates on the filesystem directly. Reads and writes library/index.json, library/usage-log.jsonl, and item files under library/.
  • Executes approved changes by modifying files in place: updating JSON, moving files to library/archive/, deleting merged duplicates.
  • Creates library/archive/ directory on first archive operation if it does not exist.
Cowork / Claude.ai Environment
  • Reads from the working folder. Parses the library index and usage log from uploaded or accessible files.
  • Cannot directly modify files. Instead, produces a report with specific recommendations.
  • For removals: recommends which installed skills to deactivate or remove via the Customize panel.
  • For promotions: provides the updated index.json content for the user to apply manually.
  • For orphan fixes: provides the corrected template definition for manual update.

Questions This Skill Answers

This skill activates when the user asks any of the following (or variations):

  • "Review the library"
  • "What's in my library?"
  • "Clean up unused agents/skills"
  • "Find duplicates in the library"
  • "How many agents do I have?"
  • "Which agents haven't been used?"
  • "Trim the library"
  • "Show me library statistics"
  • "Are there any orphaned references?"
  • "Promote tested agents"
  • "What quality tier are my agents?"
  • "Is there anything stale in the library?"
  • "Is anything in the library citing outdated research?"

References

  • ./references/review-criteria.md — Scoring rubrics, threshold definitions, and merge strategy details
  • ./schemas/index-schema.json — Library index format specification
  • ./schemas/usage-log-schema.json — Usage log entry format specification
  • docs/research/source-index.md — Bibliography used by the evidence-currency check to verify citations and quantitative claims

© jdforsythe, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (references) in skills/librarian of jdforsythe/forge.

  • SKILL.md
  • library
  • references/review-criteria.md
  • schemas

Open the folder on GitHubat commit b192c5c

Compare with similar skills

Librarian next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Librarian compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Librarian this skilljdforsythe/forge151—~5.1kAutomated safety check: PassMIT
Prediction Report WriterNVIDIA-AI-Blueprints/deep-researcher-agent886—~2.5kAutomated safety check: PassApache-2.0
Microsoft Skill CreatorMicrosoftDocs/mcp1.9k3 repos~2.1kAutomated safety check: PassCC-BY-4.0
Code EngineeropenJiuwen-ai/sciencediscovery159—~2.8kAutomated safety check: PassApache-2.0
Guidancetestdouble/han281—~1.8kAutomated safety check: PassMIT
Create Agent Skillsglittercowboy/taches-cc-resources2k—~1.7kAutomated safety check: PassMIT

Similar skills

  • Prediction Report Writer

    NVIDIA-AI-Blueprints/deep-researcher-agent

    A skill your agent uses when the final answer strategy calls for a prediction, forecast, probability estimate, price target, expected value, threshold outcome, scenario outlook, or prediction-style…

    886 GitHub stars~2.5k tokensUpdated yesterday
    Agent WorkflowsAuto-check passed
  • Microsoft Skill Creator

    MicrosoftDocs/mcp

    Official

    Create agent skills for Microsoft technologies using official documentation.

    1.9k GitHub starsUsed in 3 repos~2.1k tokens
    Agent WorkflowsAuto-check passed
  • Code Engineer

    openJiuwen-ai/sciencediscovery

    A skill your agent uses when you need to write and execute Python/R code to process, transform, and analyze data, delivering reproducible computational results with complete code-level methodology…

    159 GitHub stars~2.8k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Guidance

    testdouble/han

    Authoritative guidance for building Claude Code skills, agents, and plugins, plus init and update steps that install and refresh the plugin-building skills in the current repository.

    281 GitHub stars~1.8k tokensUpdated 9 days ago
    Agent WorkflowsAuto-check passed
  • Create Agent Skills

    glittercowboy/taches-cc-resources

    Expert guidance for creating, writing, building, and refining Claude Code Skills.

    2k GitHub stars~1.7k tokensUpdated 6 mo ago
    AI & LLM EngineeringAuto-check passed
  • Data Analysis

    xiaoyuge886/aigc

    Perform data analysis tasks including data cleaning, statistical analysis, visualization, and insight generation.

    198 GitHub stars~794 tokensUpdated 2 mo ago
    Data & AnalyticsAuto-check passed

More from jdforsythe/forge

  • Agent Creator

    jdforsythe/forge

    Creates structured agent definitions using the 7-component format grounded in persona science (the alignment-accuracy tradeoff), vocabulary routing, and the MAST failure taxonomy + Forge watchlist.

    151 GitHub stars~4.5k tokensUpdated 3 mo ago
    Auto-check passed
  • Mission Planner

    jdforsythe/forge

    Decomposes goals into team blueprints using evidence-based scaling laws, topology selection, and role design.

    151 GitHub stars~3.5k tokensUpdated 3 mo ago
    Auto-check passed
  • Skill Creator

    jdforsythe/forge

    Creates high-quality Claude Code and Cowork skills using evidence-based principles: expert vocabulary payloads for knowledge routing, dual-register descriptions for reliable triggering, named…

    151 GitHub stars~4.1k tokensUpdated 3 mo ago
    Auto-check passed

Questions about Librarian

What does Librarian do?

Reviews, curates, and maintains the Forge library of agents, skills, and templates. Librarian is an agent skill from jdforsythe/forge. Reviews, curates, and maintains the Forge library of agents, skills, and templates.

When should I use Librarian?

Librarian fits situations like: the user wants to review the library; clean up agents; check whats available; find duplicates.

How do I install Librarian in Claude Code?

Run `npx skills add jdforsythe/forge --skill librarian -a claude-code`. Or copy the skill folder (skills/librarian in jdforsythe/forge) into .claude/skills/librarian in your project. Claude Code loads it when a task matches its description.

How do I install Librarian in Codex?

Run `npx skills add jdforsythe/forge --skill librarian -a codex`. Or copy the skill folder (skills/librarian in jdforsythe/forge) into .agents/skills/librarian in your project. Codex loads it when a task matches its description.

Can I use Librarian in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jdforsythe/forge --skill librarian -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/librarian, .gemini/skills/librarian, .github/skills/librarian and .opencode/skills/librarian in your project.

What does Librarian need to run?

SKILL.md names no scripts, command-line tools or credentials: Librarian is instructions for the agent only.

Does Librarian access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Librarian safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Librarian use?

Librarian is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Librarian use?

About 5.1k tokens (SKILL.md is roughly 21k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 3.4k tokens, read only when the agent opens those files.

What are the alternatives to Librarian?

Skills that share tags, products or a category with Librarian: Prediction Report Writer (NVIDIA-AI-Blueprints/deep-researcher-agent, 886 stars), Microsoft Skill Creator (MicrosoftDocs/mcp, 1.9k stars), Code Engineer (openJiuwen-ai/sciencediscovery, 159 stars) and Guidance (testdouble/han, 281 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Librarian?

jdforsythe (a GitHub user) maintains it in jdforsythe/forge, which has 151 GitHub stars. The repository holds 4 skills in this directory. The repository was last updated on July 3, 2026.

Source: jdforsythe/forge on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.