Agent skill

Read Book

by coreyhaines31 in coreyhaines31/makerskills

When you want to read and extract structured notes from a book — PDF, EPUB, MOBI, markdown, .txt, pasted text, or URL to a public-domain work.

MITAuto-check passedDocuments & Office

Install Read Book

skills CLI
$ npx skills add coreyhaines31/makerskills --skill read-book -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install coreyhaines31/makerskills read-book --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/coreyhaines31/makerskills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/read-book .claude/skills/read-book && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
read-book
GitHub stars
851
Token cost
~2.2k tokens
SKILL.md length
932 words
Files
3 (incl. references)
Skills in repo
21
Repo updated
First seen
Licence
MIT

At a glance

When you want to read and extract structured notes from a book — PDF, EPUB, MOBI, markdown, .txt, pasted text, or URL to a public-domain work.

  • Works in 7 steps: Parse input → Parse mode → Get the text + chunk → …
  • Extract notes from this PDF
  • SKILL.md covers Step 1 — Parse input, Step 2 — Parse mode, Step 3 — Get the text + chunk and Step 4 — Read each chunk, plus 7 more sections
  • Calls brew, pdftotext and pandoc

What it does

Read Book is an agent skill from coreyhaines31/makerskills. When you want to read and extract structured notes from a book — PDF, EPUB, MOBI, markdown, .txt, pasted text, or URL to a public-domain work. Reads in chunks (by chapter when a TOC exists, by 50-page blocks otherwise), extracts per-chapter TL;DR + key concepts + quotes + action items + frameworks, and offers to capture to second-brain raw/ as a highlights- file. Four modes — notes (default, chapter-by-chapter), summary (whole-book TL;DR + 3–5 takeaways), quotes (pull-quote highlights only), study (notes + Q&A…

Its SKILL.md is about 2.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/output-modes.md` and `references/sources.md`).

It sits in Documents & Office, covering Summarization, Second brain and Study guides and flashcards. It works with Pandoc. The repository describes itself as: AI agent skills for the personal operator's craft — decisions, research, second-brain, content rotation, scenario modeling, and meta-skills to author more. Works with Claude… The licence is MIT.

When your agent uses it

  • Extract notes from this PDF
  • Whats in this book
  • Summarize this ebook
  • Pull quotes from this. Sibling to watch-video (same content-consumption pattern

Example prompts

  • “/read-book,”
  • “read this book,”
  • “extract notes from this PDF,”
  • “/read-book”

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Parse input
  2. Parse mode
  3. Get the text + chunk
  4. Read each chunk
  5. Aggregate into final notes file
  6. Offer to capture to second-brain
  7. Report

What it can do on your machine

Read from SKILL.md and the folder at commit cc31579. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • brew
    • pdftotext
    • pandoc

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Read Book loads about 2.2k tokens when it runs, and up to ~4.6k if it reads all its reference files. Until then it costs about 192 tokens; SKILL.md has 932 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~192
When it runs · the whole SKILL.md, loaded when a task matches
~2.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from coreyhaines31/makerskills at commit cc31579, republished under its MIT licence (© coreyhaines31). 932 words, ~2,160 tokens.

Download SKILL.mdSave it as .claude/skills/read-book/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
read-book
description
When you want to read and extract structured notes from a book — PDF, EPUB, MOBI, markdown, .txt, pasted text, or URL to a public-domain work. Reads in chunks (by chapter when a TOC exists, by 50-page blocks otherwise), extracts per-chapter TL;DR + key concepts + quotes + action items + frameworks, and offers to capture to second-brain raw/ as a highlights- file. Four modes — notes (default, chapter-by-chapter), summary (whole-book TL;DR + 3–5 takeaways), quotes (pull-quote highlights only), study (notes + Q&A spaced-rep prep). Triggers on "/read-book," "read this book," "extract notes from this PDF," "what's in this book," "summarize this ebook," "pull quotes from this." Sibling to watch-video (same content-consumption pattern, different medium).
metadata.version
0.1.1

/read-book — Extract structured notes from books and long PDFs

Sibling to watch-video. Same content-consumption pattern: ingest → chunk → extract → optionally capture to second-brain.

Step 1 — Parse input

Accept:

  • PDF: file path (Claude reads PDFs natively in chunks via Read pages:"X-Y")
  • EPUB / MOBI: file path (needs pandoc or ebook-convert to extract — see references/sources.md)
  • Markdown / .txt: file path (read directly)
  • Pasted text: just use what was pasted
  • URL to public-domain text: fetch it with your agent's URL-fetch tool or curl (Project Gutenberg, archive.org, etc.)

Detect type from file extension. If ambiguous, ask.

Step 2 — Parse mode

InvocationModeWhat you get
/read-book <input>notes (default)Chapter-by-chapter: TL;DR + key concepts + quotes + action items + frameworks
/read-book <input> summarysummaryWhole-book TL;DR (1 paragraph) + 3–5 key takeaways + who-it's-for
/read-book <input> quotesquotesPull-quote highlights only, with chapter context and page refs
/read-book <input> studystudyNotes mode + 10–20 spaced-repetition Q&A cards

If the book is long (>200 pages) and mode is unspecified, default to notes but warn it'll take many tool calls.

Step 3 — Get the text + chunk

See references/sources.md for per-source ingestion. Output of this step: text content + a chunking plan.

Chunking strategy (hybrid, in priority order):

  1. By chapter if a TOC exists (PDF with bookmarks, EPUB/MOBI converted via pandoc preserves chapter headers)

    • Use pdfinfo <pdf> | grep "Pages" for PDFs
    • Use pdftotext -layout <pdf> | grep -i "^chapter\|^part" for chapter detection, or read TOC from page 1–5
    • EPUB: after pandoc <epub> -o tmp.md, chunks are between # Chapter X headers
  2. By page count for PDFs without TOC: 50 pages per chunk

  3. By character count for text/markdown: 30,000 chars per chunk (~7,500 words)

Save the chunking plan as ~/Documents/books/<author>-<title-slug>-<YYYY-MM-DD>/chunks.json:

json
{
  "source": "<path>",
  "title": "<book title>",
  "author": "<author>",
  "type": "pdf",
  "total_pages": 287,
  "chunking": "by-chapter",
  "chunks": [
    {"i": 0, "label": "Introduction", "pages": "1-12"},
    {"i": 1, "label": "Chapter 1: The Problem", "pages": "13-32"},
    ...
  ]
}

Step 4 — Read each chunk

Loop:

  1. Read chunk N (Read tool with pages: for PDF, full file for text/MD)
  2. Extract per the chosen mode (see references/output-modes.md for templates)
  3. Append the chunk's notes to ~/Documents/books/<workdir>/notes-<NNN>-<label-slug>.md

For PDFs, don't read the whole book in one call — Claude's PDF tool maxes around 10 pages. Process chunks individually.

If a chunk fails to extract anything useful (e.g., it's mostly diagrams or front-matter), log the skip and continue.

Step 5 — Aggregate into final notes file

Combine all chunk notes into a single ~/Documents/books/<workdir>/notes.md matching the mode's full-book template (see references/output-modes.md).

Top of the file always has the metadata block + the second-brain-compatible frontmatter:

markdown
source: <file path or URL>
captured: YYYY-MM-DD
type: book
book_title: <title>
author: <author>
mode: notes
chunks: <count>
chunking: <strategy>

# <title> by <author>

## TL;DR
<2–3 sentences>

## Key takeaways
1. ...

## Chapter notes
...

## Cross-references (suggested for wiki)
- Could connect to [[Longevity Biomarkers]] (per Chapter 3 discussion of biomarkers)
- Could connect to [[Productivity & Systems]] (per Chapter 7 framework)

The cross-reference suggestions are advisory — they're suggestions for /sb compile to act on, not auto-applied. Keep responsibilities separated.

Step 6 — Offer to capture to second-brain

Ask:

"Want to capture this to second-brain? I'll write it to ${SECOND_BRAIN_VAULT:-$HOME/Documents/SecondBrain}/raw/highlights-<slug>.md matching your vault's highlights- type prefix."

Default is ask, never auto-write. If yes:

  1. Copy the final notes.md (with the second-brain-compatible frontmatter at top) to ${SECOND_BRAIN_VAULT:-$HOME/Documents/SecondBrain}/raw/highlights-<slug>.md
  2. Tell the user the path
  3. Suggest: "Run /sb compile later to merge this into wiki pages — the cross-reference suggestions in the footer are starting points."

If the user skips capture, the workdir still has everything — they can grab the file later.

Step 7 — Report

In chat:

  • One-line headline: <title> · <author> · <total_pages or word count> · <mode> · <chunks processed>
  • Workdir path
  • The TL;DR section
  • For notes / study modes: brief list of top 3 takeaways
  • For quotes mode: top 3 quotes
  • If captured to second-brain: that path too
Show full SKILL.md (398 more words)Show less

Modes (quick invocations)

InvocationModeBehavior
/read-book <input>notesFull pipeline, default mode
/read-book <input> summarysummaryJust TL;DR + key takeaways (1 read pass for short books, sampled chapters for long)
/read-book <input> quotesquotesChapter-by-chapter, but only output quotes
/read-book <input> studystudyNotes + Q&A spaced-rep cards
/read-book <input> --capture(any)Skip the ask step, auto-write to second-brain raw/
/read-book <input> --render pdf(any)Also render the final notes.md to PDF via pandoc (uses ~/.local/share/makerskills/render.css). See references/output-modes.md.
/read-book <input> --render html(any)Same as above but HTML

Composes with

  • second-brain — primary integration: writes highlights-<slug>.md to raw/. Then /sb compile merges into wiki pages.
  • deep-research — when a research question turns up a book, /read-book is the next step. Notes feed back into the research brief.
  • business-brainstorm — when scoring an idea (e.g., business books on similar models), read-book provides the structured evidence.
  • decide — when a decision hinges on what an authority has written (e.g., "should I take VC money?" → read Naval / Jason Cohen), read-book extracts the relevant chapter.
  • slide-deck — book takeaways → talk material (book talk pattern).
  • watch-video — sibling skill, same content-consumption pattern. Audiobook? Use watch-video transcript mode.
  • nonfictionskills / fictionskills — when researching to write a book, this skill reads the comp titles.

Error handling

FailureResponse
EPUB/MOBI without pandoc / ebook-convertTell the user: brew install pandoc or brew install calibre (calibre includes ebook-convert)
PDF is scanned (no text layer)Suggest OCR first: brew install ocrmypdf && ocrmypdf <pdf> <pdf-ocr.pdf>
PDF has no detectable TOCFall back to 50-page chunks. Note in the metadata.
Book is unusually long (>500 pages)Warn cost / time, ask if the user wants summary mode instead of full notes
Chunk extraction emptySkip the chunk, log, continue. Don't fail the whole run.

Notes on quality

  • Don't summarize beyond recognition. A 30-page chapter should produce 8–15 lines of notes, not 3. Compression is good; flattening is bad.
  • Preserve specifics. Names, numbers, dates, quotes — keep them. The whole point is later-the user can grep "what did Andy Wilkinson say about X" and find it.
  • Quotes are sacred. When you flag a quote, copy it verbatim. Note the page if possible.
  • Action items are explicit. If the book makes you think "I should do X," flag it explicitly. These are the highest-leverage outputs.
  • Frameworks deserve their own bullets. When the author names a framework (e.g., "the 9-dimension filter," "Save the Cat beats"), call it out by name in the notes.

© coreyhaines31, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (references) in skills/read-book of coreyhaines31/makerskills.

  • SKILL.md
  • references/output-modes.md
  • references/sources.md

Open the folder on GitHubat commit cc31579

Compare with similar skills

Read Book next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Read Book compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Read Book this skillcoreyhaines31/makerskills851—~2.2kAutomated safety check: PassMIT
Lecture Slides SummarizerLi-Baichuan-James/summarize-slides-skill313—~6.7kAutomated safety check: PassMIT
Second Brain Import PDFPieroSierra/SecondBrain153—~2.9kAutomated safety check: PassNone
Harness Book Best Practicewquguru/harness-books3.2k—~4.1kAutomated safety check: PassNone
Read URLs and PDFstw93/Waza7.2k—~1.8kAutomated safety check: PassMIT
Huashu Markdown Publishing Pipelinealchaincyf/huashu-md-html910—~4.8kAutomated safety check: PassMIT

Similar skills

  • Lecture Slides Summarizer

    Li-Baichuan-James/summarize-slides-skill

    Condenses a lecture PDF into an exam-focused LaTeX cheat sheet and compiled PDF with page citations, bilingual terms, formulas and only the diagrams that help.

    313 GitHub stars~6.7k tokensUpdated 5 mo ago
    EducationAuto-check passed
  • Second Brain Import PDF

    PieroSierra/SecondBrain

    Extract text from a PDF, convert it to markdown, and write the result into raw/pdf/, ready for ingestion.

    153 GitHub stars~2.9k tokensUpdated 5 days ago
    Documents & OfficeAuto-check passed
  • Harness Book Best Practice

    wquguru/harness-books

    Best practices for working on the Harness books repo. An agent skill from wquguru/harness-books.

    3.2k GitHub stars~4.1k tokensUpdated 5 mo ago
    Documents & OfficeAuto-check passed
  • Fetches web pages and PDFs and returns a source-grounded summary, clean Markdown, quotes or citations, routing each kind of link to a suitable fetch method.

    7.2k GitHub stars~1.8k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Huashu Markdown Publishing Pipeline

    alchaincyf/huashu-md-html

    Converts files and web pages into clean Markdown, then turns Markdown into polished HTML, Word, PDF and EPUB using four templates.

    910 GitHub stars~4.8k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed
  • Quant Paper Extractor

    CamusGIT/EvoQuant

    Convert quantitative research report PDFs to markdown, then extract structured knowledge (paperId, title, year, source, keywords, tldr, abstract, strategy, method, experiment, result) into JSONL…

    151 GitHub stars~2.4k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed

More from coreyhaines31/makerskills

All 21 skills in this repo
  • Business Brainstorm

    coreyhaines31/makerskills

    When you want to pressure-test a potential new business, product, or side project against the serial-founder filter.

    851 GitHub stars~1.7k tokensUpdated 2 days ago
    Auto-check passed
  • Company Brain

    coreyhaines31/makerskills

    Your team's shared, AI-ready knowledge base — people, companies, meetings, SOPs, and decisions structured so an agent can answer on your team's behalf.

    851 GitHub stars~4.9k tokensUpdated 2 days ago
    Auto-check passed
  • Company Cfo

    coreyhaines31/makerskills

    Monthly CFO workflow for a company or agency — pull raw data from bank + payment processor + payroll + expense management, categorize and reconcile, compute end-of-month cash via transaction-sum…

    851 GitHub stars~4k tokensUpdated 2 days ago
    Auto-check: notes
  • Decide

    coreyhaines31/makerskills

    When you have a decision to make and want a structured workflow that picks the load-bearing questions, walks through them, reaches a call (or "wait"), and archives the rationale for future reference.

    851 GitHub stars~1.9k tokensUpdated 2 days ago
    Auto-check passed
  • Ingest

    coreyhaines31/makerskills

    When you paste raw human input — a call transcript (Grain, Zoom, Granola, Fathom), a text or email from a client/partner/friend, a voice-memo dump, or meeting notes — and want it converted into…

    851 GitHub stars~1.8k tokensUpdated 2 days ago
    Auto-check passed
  • Maker Council

    coreyhaines31/makerskills

    When you want multiple expert perspectives on a founder/operator question — a simulated board of advisors (Jason Fried, Elon Musk, Jeff Bezos, Jensen Huang, Bob Iger, Paul Graham, Naval Ravikant…

    851 GitHub stars~2.9k tokensUpdated 2 days ago
    Auto-check passed

Works with

Questions about Read Book

What does Read Book do?

When you want to read and extract structured notes from a book — PDF, EPUB, MOBI, markdown, .txt, pasted text, or URL to a public-domain work. Read Book is an agent skill from coreyhaines31/makerskills.txt, pasted text, or URL to a public-domain work.

When should I use Read Book?

Read Book fits situations like: extract notes from this PDF; whats in this book; summarize this ebook; pull quotes from this. Sibling to watch-video (same content-consumption pattern.

How do I install Read Book in Claude Code?

Run `npx skills add coreyhaines31/makerskills --skill read-book -a claude-code`. Or copy the skill folder (skills/read-book in coreyhaines31/makerskills) into .claude/skills/read-book in your project. Claude Code loads it when a task matches its description.

How do I install Read Book in Codex?

Run `npx skills add coreyhaines31/makerskills --skill read-book -a codex`. Or copy the skill folder (skills/read-book in coreyhaines31/makerskills) into .agents/skills/read-book in your project. Codex loads it when a task matches its description.

Can I use Read Book in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add coreyhaines31/makerskills --skill read-book -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/read-book, .gemini/skills/read-book, .github/skills/read-book and .opencode/skills/read-book in your project.

What does Read Book need to run?

Going by SKILL.md and its folder, Read Book needs the command-line tools its instructions call (brew, pdftotext and pandoc).

Does Read Book access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Read Book safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Read Book use?

Read Book is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Read Book use?

About 2.2k tokens (SKILL.md is roughly 8.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.5k tokens, read only when the agent opens those files.

What are the alternatives to Read Book?

Skills that share tags, products or a category with Read Book: Lecture Slides Summarizer (Li-Baichuan-James/summarize-slides-skill, 313 stars), Second Brain Import PDF (PieroSierra/SecondBrain, 153 stars), Harness Book Best Practice (wquguru/harness-books, 3.2k stars) and Read URLs and PDFs (tw93/Waza, 7.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Read Book?

coreyhaines31 (a GitHub user) maintains it in coreyhaines31/makerskills, which has 851 GitHub stars. The repository holds 21 skills in this directory. The repository was last updated on October 8, 2026.

Source: coreyhaines31/makerskills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.