Agent skill

Extract Document Insights

by Abilityai in Abilityai/cornelius

Extract insights from external documents (research papers, books, articles).

MITAuto-check passedResearch & Science

Install Extract Document Insights

skills CLI
$ npx skills add Abilityai/cornelius --skill extract-document-insights -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Abilityai/cornelius extract-document-insights --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Abilityai/cornelius.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/extract-document-insights .claude/skills/extract-document-insights && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
extract-document-insights
GitHub stars
109
Token cost
~649 tokens
SKILL.md length
221 words
Files
1
Skills in repo
51
Repo updated
First seen
Licence
MIT

At a glance

Extract insights from external documents (research papers, books, articles).

  • Works in 4 steps: Parse arguments to extract session name… → Spawn document-insight-extractor… → Report results - number of insights… → …
  • Tasks that involve Subagents
  • SKILL.md covers Purpose, When to Use, Usage and Process, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Extract Document Insights is an agent skill from Abilityai/cornelius. Extract insights from external documents (research papers, books, articles). Spawns document-insight-extractor subagent. Requires session name.

Its SKILL.md is about 650 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Research & Science, covering Subagents. The repository describes itself as: AI-powered second brain template for Claude Code + Obsidian. The licence is MIT.

When your agent uses it

  • Tasks that involve Subagents

Example prompts

  • “/extract-document-insights”

Requirements

  • Pre-approved tools (allowed-tools): Task

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Parse arguments to extract session name and source
  2. Spawn document-insight-extractor subagent via Task tool with
  3. Report results - number of insights extracted, session folder path
  4. Suggest insight interview - Present to user

What it can do on your machine

Read from SKILL.md and the folder at commit fd5e9a4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Task

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Extract Document Insights loads about 649 tokens when it runs. Until then it costs about 42 tokens; SKILL.md has 221 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~42
When it runs · the whole SKILL.md, loaded when a task matches
~649

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Abilityai/cornelius at commit fd5e9a4, republished under its MIT licence (© Abilityai). 221 words, ~649 tokens.

Download SKILL.mdSave it as .claude/skills/extract-document-insights/SKILL.md (or your agent's skills folder).
name
extract-document-insights
description
Extract insights from external documents (research papers, books, articles). Spawns document-insight-extractor subagent. Requires session name.
allowed-tools
Task
user-invocable
true
arg-description
<session name> <file path or content description>

Extract Document Insights

Extract insights from EXTERNAL documents with proper epistemic classification.

Purpose

User-invocable wrapper that spawns the document-insight-extractor subagent to:

  • Extract research findings, theoretical frameworks, hypotheses
  • Classify epistemic status (confirmed, theoretical, speculative)
  • Store in session-based folders
  • Deduplicate against existing knowledge base
  • Save to Brain/Document Insights/[session]/

When to Use

  • Analyzing research papers, academic articles
  • Processing books or book chapters
  • Extracting insights from web articles, reports
  • Any EXTERNAL content (not your personal thoughts)

NOT for personal content - use /extract-insights for your conversations, transcripts.

Usage

/extract-document-insights "2025-02-21 Ancient Wisdom Research" /path/to/document.md
/extract-document-insights "AI Agent Papers" "the transcript above"

Session name is REQUIRED - describes the research session for organization.

Process

  1. Parse arguments to extract session name and source

  2. Spawn document-insight-extractor subagent via Task tool with:

    Task(
      subagent_type="document-insight-extractor",
      prompt="Extract insights from: [source] into session '[session-name]'. Follow your mandatory workflow - contextualize, classify epistemically, search duplicates, create notes, update changelog."
    )
  3. Report results - number of insights extracted, session folder path

  4. Suggest insight interview - Present to user:

    "[N] insights from [source] saved to Brain/Document Insights/[session]/.

    Before these get indexed, consider running /insight-interview [topic] to capture your own angles - where you agree, disagree, or see connections the research missed. Your responses save to Brain/AI Extracted Notes/ and both sets will be available for the next connection discovery run.

    Run /insight-interview [topic] to continue, or skip if you want to index these as-is."

Outputs

  • Insight notes in Brain/Document Insights/[session]/
  • Session changelog in same folder
  • Summary with epistemic breakdown (confirmed vs hypothetical)
  • Suggested next step: /insight-interview [topic] to layer in your personal perspective

© Abilityai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .claude/skills/extract-document-insights of Abilityai/cornelius.

Open the folder on GitHubat commit fd5e9a4

Compare with similar skills

Extract Document Insights next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Extract Document Insights compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Extract Document Insights this skillAbilityai/cornelius109—~649Automated safety check: PassMIT
Web ResearchJuncai22/spring-ai-agent-learning1233 repos~1.1kAutomated safety check: PassApache-2.0
Verify Openspec Docsgoern/forgejo-mcp1411 repos~1.1kAutomated safety check: PassGPL-3.0
Deep Research312362115/claude107—~6.6kAutomated safety check: PassMIT
ULW Deep Researchcode-yeongyu/oh-my-openagent70k—~14kAutomated safety check: PassCustom licence
Deep ResearchXiaomiMiMo/MiMo-Code14k—~1.2kAutomated safety check: PassMIT

Similar skills

  • Web Research

    Juncai22/spring-ai-agent-learning

    A skill your agent uses for requests related to web research; it provides a structured approach to conducting comprehensive web research

    123 GitHub starsUsed in 3 repos~1.1k tokens
    Research & ScienceAuto-check passed
  • Verify Openspec Docs

    goern/forgejo-mcp

    Fact-checks OpenSpec user documentation with a fresh-context subagent that re-runs commands and checks claims against source.

    141 GitHub starsUsed in 1 repo~1.1k tokens
    Research & ScienceAuto-check passed
  • Deep Research

    312362115/claude

    深度调研技能:对任意命题进行系统性调研并输出专业研究报告. An agent skill from 312362115/claude.

    107 GitHub stars~6.6k tokensUpdated 4 mo ago
    Research & ScienceAuto-check passed
  • ULW Deep Research

    code-yeongyu/oh-my-openagent

    Runs an exhaustive, team-based research session that stands up cooperating agents, debates findings and delivers a report where every claim has a citation or proof.

    70k GitHub stars~14k tokensUpdated today
    Research & ScienceAuto-check passed
  • Deep Research

    XiaomiMiMo/MiMo-Code

    Runs a multi-source investigation with parallel sub-agents and built-in web tools, then writes one cited report. Meant for open-ended topics, not quick lookups.

    14k GitHub stars~1.2k tokensUpdated 4 days ago
    Research & ScienceAuto-check passed
  • Seven Pass Review

    pedrohcgs/claude-code-my-workflow

    Mechanize Pattern 15 — the seven-pass adversarial review protocol for academic manuscripts.

    1.6k GitHub stars~3.3k tokensUpdated 10 days ago
    Research & ScienceAuto-check: notes

More from Abilityai/cornelius

All 51 skills in this repo
  • Nano Banana Image Generator

    Abilityai/cornelius

    Generate images using Google's Nano Banana (Gemini 2.5 Flash Image).

    109 GitHub stars~1.2k tokensUpdated 15 days ago
    Auto-check: notes
  • Changelog Protocol

    Abilityai/cornelius

    Protocol for creating dated changelog files after significant agent sessions.

    109 GitHub stars~555 tokensUpdated 15 days ago
    Auto-check passed
  • Create Article

    Abilityai/cornelius

    Create long-form articles from knowledge base insights. An agent skill from Abilityai/cornelius.

    109 GitHub stars~2.1k tokensUpdated 15 days ago
    Auto-check: notes
  • Epistemic Classification

    Abilityai/cornelius

    Framework for distinguishing research findings from hypotheses and speculative synthesis.

    109 GitHub stars~1.5k tokensUpdated 15 days ago
    Auto-check passed
  • Get Youtube Transcript

    Abilityai/cornelius

    Extract the transcript from a YouTube video by URL or video ID.

    109 GitHub stars~525 tokensUpdated 15 days ago
    Auto-check: notes
  • Insight Capture Format

    Abilityai/cornelius

    Standard format for capturing and documenting insights in the knowledge base.

    109 GitHub stars~616 tokensUpdated 15 days ago
    Auto-check passed

Questions about Extract Document Insights

What does Extract Document Insights do?

Extract insights from external documents (research papers, books, articles). Extract Document Insights is an agent skill from Abilityai/cornelius. Extract insights from external documents (research papers, books, articles).

When should I use Extract Document Insights?

Extract Document Insights fits situations like: tasks that involve Subagents.

How do I install Extract Document Insights in Claude Code?

Run `npx skills add Abilityai/cornelius --skill extract-document-insights -a claude-code`. Or copy the skill folder (.claude/skills/extract-document-insights in Abilityai/cornelius) into .claude/skills/extract-document-insights in your project. Claude Code loads it when a task matches its description.

How do I install Extract Document Insights in Codex?

Run `npx skills add Abilityai/cornelius --skill extract-document-insights -a codex`. Or copy the skill folder (.claude/skills/extract-document-insights in Abilityai/cornelius) into .agents/skills/extract-document-insights in your project. Codex loads it when a task matches its description.

Can I use Extract Document Insights in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Abilityai/cornelius --skill extract-document-insights -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/extract-document-insights, .gemini/skills/extract-document-insights, .github/skills/extract-document-insights and .opencode/skills/extract-document-insights in your project.

What does Extract Document Insights need to run?

SKILL.md names no scripts, command-line tools or credentials: Extract Document Insights is instructions for the agent only. Its frontmatter pre-approves these tools: Task.

Does Extract Document Insights access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Extract Document Insights safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Extract Document Insights use?

Extract Document Insights is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Extract Document Insights use?

About 649 tokens (SKILL.md is roughly 2.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Extract Document Insights?

Skills that share tags, products or a category with Extract Document Insights: Web Research (Juncai22/spring-ai-agent-learning, 123 stars), Verify Openspec Docs (goern/forgejo-mcp, 141 stars), Deep Research (312362115/claude, 107 stars) and ULW Deep Research (code-yeongyu/oh-my-openagent, 70k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Extract Document Insights?

Abilityai (a GitHub organization) maintains it in Abilityai/cornelius, which has 109 GitHub stars. The repository holds 51 skills in this directory. The repository was last updated on September 22, 2026.

Source: Abilityai/cornelius on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.