Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .claude/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Type this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
skills CLI
$ npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a codex
Project install goes to .agents/skills/; add -g for ~/.codex/skills/.
Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .agents/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a cursor
Project install goes to .agents/skills/; add -g for ~/.cursor/skills/.
Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .cursor/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
skills CLI
$ npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a gemini-cli
Project install goes to .agents/skills/; add -g for ~/.gemini/skills/.
Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .gemini/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Installs for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
skills CLI
$ npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a github-copilot
Project install goes to .agents/skills/; add -g for ~/.copilot/skills/.
Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .github/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a opencode
OpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
Install the "pp-substack-reader" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-substack-reader into .opencode/skills/pp-substack-reader/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-substack-reader", then confirm the skill loads.
OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Facts
Skill name
pp-substack-reader
GitHub stars
2.1k
Token cost
~7.9k tokens
SKILL.md length
3,268 words
Files
1
Skills in repo
505
Repo updated
First seen
Licence
Apache-2.0
At a glance
Read any Substack publication as a local, full-text-searchable corpus — keyless for free posts, your own session for what you subscribe to.
Works in 6 steps: recall before any discovery → decision tree → always read warnings → …
Phrases: archive this Substack
SKILL.md covers Prerequisites: Install the CLI, When to Use This CLI, Anti-triggers and Unique Capabilities, plus 8 more sections
Calls go, claude and npx; reaches substack.com
What it does
Pp Substack Reader is an agent skill from mvanhorn/printing-press-library. Read any Substack publication as a local, full-text-searchable corpus — keyless for free posts, your own session for what you subscribe to. Trigger phrases: archive this Substack, read this Substack post, search my Substack corpus, what's new in my newsletters, use substack-reader, run substack.
Its SKILL.md is about 7.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Writing & Content, covering Newsletters. It works with Substack. The repository describes itself as: Official library of CLIs generated by the CLI Printing Press. Endorsed, tested, and community-contributed. The licence is Apache-2.0.
When your agent uses it
Phrases: archive this Substack
Read this Substack post
Search my Substack corpus
Whats new in my newsletters
Example prompts
“/pp-substack-reader”
Requirements
Node.js
Pre-approved tools (allowed-tools): Read, Bash
Workflow steps
6 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 7638ad4. It shows what the files ask for, not the result of running them.
Tool permissions
Pre-approves these tools, so the agent can use them without asking each time:
Read
Bash
From allowed-tools in the SKILL.md frontmatter.
Runs code
Shell commands in SKILL.md call:
go
claude
npx
From the folder's file list and the shell code blocks in SKILL.md.
Network
Hosts in commands or code, which the agent is likely to contact:
substack.com
From URLs in SKILL.md, links to its own repository left out.
Credentials
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Context cost
Pp Substack Reader loads about 7.9k tokens when it runs. Until then it costs about 82 tokens; SKILL.md has 3,268 words of instructions outside code blocks.
Always· name and description, kept in context so the agent knows when to use it
~82
When it runs· the whole SKILL.md, loaded when a task matches
~7.9k
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
Safety
Auto-check: notes
The automated check noted patterns worth knowing about, such as sudo or a known installer.
NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
allowed-tools: Read, Bash
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
Download SKILL.mdSave it as .claude/skills/pp-substack-reader/SKILL.md (or your agent's skills folder).
name
pp-substack-reader
description
Read any Substack publication as a local, full-text-searchable corpus — keyless for free posts, your own session for what you subscribe to. Trigger phrases: `archive this Substack`, `read this Substack post`, `search my Substack corpus`, `what's new in my newsletters`, `use substack-reader`, `run substack`.
allowed-tools
Read, Bash
author
Maxime Delavergne
license
Apache-2.0
argument-hint
<command> [args] | install cli|mcp
<!-- GENERATED FILE — DO NOT EDIT.
This file is a verbatim mirror of library/media-and-entertainment/substack-reader/SKILL.md,
regenerated post-merge by tools/generate-skills/. Hand-edits here are
silently overwritten on the next regen. Edit the library/ source instead.
See the repository agent guide, section "Generated artifacts: registry.json, cli-skills/". -->
Substack Reader — Printing Press CLI
Prerequisites: Install the CLI
This skill drives the substack-reader-pp-cli binary. You must verify the CLI is installed before invoking any command from this skill. If it is missing, install it first:
Install via the Printing Press installer. It defaults binaries to $HOME/.local/bin on macOS/Linux and %LOCALAPPDATA%\Programs\PrintingPress\bin on Windows:bash
Ensure the reported install directory is on $PATH for the agent/runtime that will invoke this skill.
If the npx install fails (no Node, offline, etc.), fall back to a direct Go install (requires Go 1.26.6 or newer). This installs into $GOPATH/bin (default $HOME/go/bin), so add that directory to $PATH instead:
bash
go install github.com/mvanhorn/printing-press-library/library/media-and-entertainment/substack-reader/cmd/substack-reader-pp-cli@latest
If --version reports "command not found" after install, the runtime cannot see the binary directory on $PATH. Do not proceed with skill commands until verification succeeds.
Substack Reader archives whole publications into a local SQLite mirror you can search, SQL-query, and read offline. Free posts need no login; paid posts you're entitled to unlock with your own session cookie — never redistributed, always opt-in. Unlike every other Substack tool it builds a corpus that compounds instead of fetching live per call.
When to Use This CLI
Use Substack Reader when you want a durable, searchable local copy of one or more Substack publications for reading, agent workflows, or analysis — especially reading a specific post's full text or searching across newsletters offline. It is the right tool when you value a corpus that compounds over live per-call fetching.
Anti-triggers
Do not use this CLI for:
Do not use it to publish, schedule, or manage a Substack you own (this is read-only) — use the Substack web app or a publishing tool.
Do not use it to bulk-scrape or redistribute paid content you are not entitled to — it reads only your own entitled content, on demand.
Do not use it to manage subscribers, payments, or analytics for your own publication.
Unique Capabilities
These capabilities aren't available in any other tool for this API.
Local corpus that compounds
archive — Archive a whole Substack publication into a local SQLite mirror you can read, search, and query offline — no other Substack tool builds a persistent corpus. --limit 0 walks the whole archive; when a run stops at --limit instead of exhausting the archive, the output says so (JSON carries "exhausted": false) — never treat a limit-stopped run as a complete archive.
Reach for this to turn a live newsletter into a durable, queryable knowledge base instead of re-fetching every time.
sql — Run read-only SQL over your local Substack corpus for arbitrary analytics — post cadence, audience mix, longest posts — from data you've already archived.
Reach for this for ad-hoc analytics over what you've archived, without re-fetching or writing code.
bash
substack-reader-pp-cli sql "SELECT json_extract(data,'$.audience') AS audience, COUNT(*) FROM resources WHERE resource_type='posts' GROUP BY audience"
Entitlement-aware reading
read — Read one or more posts' full text in a single call; free posts keyless, and paid posts you subscribe to via your own session cookie — with an honest 'preview only, you're not entitled' signal. With several posts, JSON mode emits an array of envelopes (a failed post becomes a {"post", "error"} entry and the rest still return); a single post keeps the single-object envelope.
Use to pull posts' full text into an agent workflow — several slugs in one call, no shell loop — respecting exactly what the user is entitled to.
substack-reader-pp-cli categories browse — List publications in a category
substack-reader-pp-cli categories list — List all Substack categories
publications — Discover Substack publications
substack-reader-pp-cli publications <query> — Search Substack publications by name (best-effort; may return few results anonymously)
Finding the right command
When you know what you want to do but not which command does it, ask the CLI directly:
bash
substack-reader-pp-cli which "<capability in your own words>"
which resolves a natural-language capability query to the best matching command from this CLI's curated feature index. Exit code 0 means at least one match; exit code 2 means no confident match — fall back to --help or use a narrower query.
Mirror a whole publication (--limit 0 = no cap) then search it offline with FTS ranking. Compact/agent search hits carry key fields plus a bounded snippet — never full post bodies; use read (or --select) when you want a hit's full text.
Batch-read a list of slugs without a shell loop; JSON mode returns an array of post envelopes, and one bad slug doesn't sink the rest.
Audience mix analytics
bash
substack-reader-pp-cli sql "SELECT json_extract(data,'$.audience') AS audience, COUNT(*) FROM resources WHERE resource_type='posts' GROUP BY audience"
Read-only SQL over the local corpus for arbitrary analytics.
Auth Setup
Free and public posts are keyless — zero setup. Every command works the moment the CLI is installed: no account, no API key, nothing to configure. Authentication only ever matters for paid posts.
There is one optional layer: your own Substack session cookie (substack.sid). Substack decides entitlement server-side from the session you present, so handing the reader your own cookie unlocks exactly the paid posts you already subscribe to — and nothing else. This is your own cookie, never an API key, it is never required for free content, and one cookie covers every subscription on your account (entitlement is account-level, not per-publication).
Provide it any of these three ways (first hit wins):
bash
# 1. Environment variable — the bare substack.sid value...
export SUBSTACK_SESSION="s%3A..."
# ...or the whole cookie fragment, if that is what you copied
export SUBSTACK_SESSION="substack.sid=s%3A..."
# 2. A JSON cookie file at the default config-dir path
# (~/.config/substack-reader-pp-cli/cookie.json on the platform default).
# The directory does not exist on a clean install — create it first.
mkdir -p ~/.config/substack-reader-pp-cli && chmod 700 ~/.config/substack-reader-pp-cli
echo '{"substack.sid":"s%3A..."}' > ~/.config/substack-reader-pp-cli/cookie.json
chmod 600 ~/.config/substack-reader-pp-cli/cookie.json
# 3. The same JSON cookie file, kept anywhere on disk
export SUBSTACK_COOKIE_FILE=/path/to/substack-cookie.json
Precedence is SUBSTACK_SESSION, then SUBSTACK_COOKIE_FILE, then the default config-dir file. SUBSTACK_COOKIE_FILE is authoritative once set: a missing, unreadable, unparseable or substack.sid-less file there is a hard error, never a silent fall-through to a stale default. The default config-dir file is best-effort: a file that exists but cannot be read or parsed degrades to anonymous with a warning, while an absent file — or one that parses but carries no substack.sid — degrades to anonymous silently. The config directory itself is relocatable with SUBSTACK_READER_CONFIG_DIR or SUBSTACK_READER_HOME.
Getting the cookie out of your browser. Open DevTools → Application → Cookies → https://substack.com, and copy the value of substack.sid. Only substack.sid is needed — connect.sid is not. Copy it rather than retyping it: a single capital-I/lowercase-l slip yields a cookie that silently reads as anonymous.
With the MCP server. The MCP server resolves the session exactly as the CLI does, so the cookie file in the config directory needs no env block at all — save cookie.json and the server picks it up. If you would rather configure it explicitly, set either variable under env in your Claude Desktop config (the MCPB bundle prompts for it as an optional field).
Your cookie stays yours. It never leaves your machine except to Substack itself, is masked in every output path so scripted runs cannot leak it, and is never redistributed. Any copy the CLI persists is written 0600 under a 0700 parent, and a group/other-readable cookie file earns a warning.
Failure is honest, never silent. With no session, a paid post returns Substack's public preview and says so: access: preview, authenticated: false, full: false, plus a preview only (N of ~M words) — no session configured line on stderr. A downgrade is always visible; it never passes itself off as a full read.
Run substack-reader-pp-cli doctor to verify setup.
Agent Mode
Add --agent to any command. Expands to: --json --compact --no-input --no-color --yes.
Pipeable — JSON on stdout, errors on stderr
Filterable — --select keeps a subset of fields. Dotted paths descend into nested structures; arrays traverse element-wise. Critical for keeping context small on verbose APIs:
bash
substack-reader-pp-cli categories list --agent --select id,name,status
Previewable — --dry-run shows the request without sending
Offline-friendly — sync/search commands can use the local SQLite store when available
Non-interactive — never prompts, every input is a flag
Read-only — do not use this CLI for create, update, delete, publish, comment, upvote, invite, order, send, or other mutating requests
Response envelope
Commands that read from the local store or the API wrap output in a provenance envelope:
Parse .results for data and .meta.source to know whether it's live or local. A human-readable N results (live) summary is printed to stderr only when stdout is a terminal AND no machine-format flag (--json, --csv, --compact, --quiet, --plain, --select) is set — piped/agent consumers and explicit-format runs get pure JSON on stdout.
Paths and state
Agents should treat the CLI's path resolver as part of the runtime contract:
Use --home <dir> for one invocation, or set SUBSTACK_READER_HOME=<dir> to relocate all four path kinds under one root.
Use per-kind env vars only when a specific kind must diverge: SUBSTACK_READER_CONFIG_DIR, SUBSTACK_READER_DATA_DIR, SUBSTACK_READER_STATE_DIR, SUBSTACK_READER_CACHE_DIR.
Resolution order is per-kind env var, --home, SUBSTACK_READER_HOME, XDG (XDG_CONFIG_HOME, XDG_DATA_HOME, XDG_STATE_HOME, XDG_CACHE_HOME), then platform defaults.
config contains settings like config.toml and profiles. data contains credentials.toml, data.db, cookies, and auth sidecars. state contains persisted queries, jobs, and teach.log. cache contains regenerable HTTP/cache files.
Stored secrets live in credentials.toml under the data dir. Existing legacy config.toml secrets are read for compatibility and leave config.toml on the first auth write.
Run substack-reader-pp-cli doctor --fail-on warn to surface path and credential-location warnings. agent-context exposes a schema v4 paths block for agents that need the resolved dirs.
For MCP, pass relocation through the MCP host config. The MCP binary does not inherit CLI flags:
Fleet precedence: an inherited per-kind env var overrides an explicit --home for that kind. Use SUBSTACK_READER_HOME or per-kind vars as durable fleet levers, and use --home only for a single invocation. Relocation is not reversible by unsetting env vars; move files manually before clearing SUBSTACK_READER_HOME, or doctor will not find credentials left under the former root.
Automatic learning
This CLI ships a self-capturing learning loop. The CLI does its own bookkeeping: every invocation is journaled locally, a failed flag followed by a corrected retry auto-derives a flag_alias candidate, and a teach on a query family without a playbook auto-synthesizes a playbook_candidate from the session's journal. Your job is judgment only: recall first, act on surfaced candidates, teach the final answer, playbook amend when you observe a correction. You never record failures by hand.
Step 1: recall before any discovery
Before list/search/drill commands on a new user question, run:
Empty-store short-circuit: if the store has no learnings, playbooks, or candidates yet (recall finds nothing and learnings list and learnings candidates are both empty), skip recall for the rest of this session instead of taxing every query; resume recall-first once something has been taught.
Step 2: decision tree
Read candidates, playbook, notes, results[0], and warnings in that order:
if Candidates present (warnings include "candidates_present"):
-> candidates are try-then-confirm, never facts. Follow each candidate's
two-step next_action verbatim: run the trial command first, then run
`learnings confirm <id>` only after the trial verified the behavior.
Reject a wrong candidate with `learnings reject <id>`.
-> NEVER re-teach something recall surfaced as a candidate; confirm or
reject that candidate instead of teaching a duplicate.
-> candidates ride alongside playbooks and resource hits, not instead of
them; continue with the branches below after acting on them.
if Playbook present:
-> READ Playbook.notes verbatim FIRST (workarounds + gotchas the CLI surface doesn't expose)
-> replay Playbook.steps in order, substituting Playbook.slots_resolved entries
for the entity slot tokens. If a step's slot is unresolved, fall back to
discovery for that step only.
-> the Playbook's expected_tool_calls is a budget; if you find yourself running
materially more, record the divergence via `substack-reader-pp-cli playbook amend`
at end-of-session.
elif Notes present (no Playbook):
-> read Notes verbatim before any discovery step; they carry known gotchas
for this query family even when no structured choreography exists yet.
elif Found AND Results[0].EntityMatch == "exact" AND Results[0].Confidence >= 2:
-> skip discovery; fetch live data for Results[*].ResourceID in parallel
elif Found AND Results[0].EntityMatch == "partial":
-> candidate hint, NOT a hit; read the resource title to validate before trusting
elif (any row in Mismatches[] when --debug-mismatches was passed):
-> treat as cold start; the stored learning is for a different entity
(different canonical resolved from query_entities)
else: // Found == false, no playbook, no notes
-> cold start; run discovery normally; teach the answer afterward (Step 4).
If the family has no playbook yet, that teach auto-synthesizes a
playbook candidate from this session's journal - you do not need to
record one by hand.
Playbook and Notes are orthogonal to the per-resource path. A recall response can carry both a Playbook AND a Results[] hit - use both: the Playbook tells you which choreography to run; the resource hits short-circuit specific steps. Default to skipping mismatches; pass --debug-mismatches only when investigating cold-start surprises.
Candidate judgment details: learnings confirm <id> prints the candidate's full payload before materializing it - check that the printed payload matches the behavior you verified. learnings reject <id> tombstones the derivation signature so the same candidate does not resurface. The envelope carries only the few candidates worth acting on now; substack-reader-pp-cli learnings candidates lists the full open set.
Graceful degradation: if learnings confirm is an unknown command, you are driving an older binary - ignore the candidates guidance and follow the rest of the protocol.
Show full SKILL.md (1,399 more words)Show less
Step 3: always read warnings
low_confidence: row exists at confidence<2. Treat as a hint, not a skip-discovery hit.
resource_not_in_store: the local store doesn't have the resource the learning points at. The match validator couldn't classify entities — direct-fetch and re-evaluate.
cross_alias_match (per-result): the row was taught under a different alias and matched the live query's canonical via entity_lookups (e.g., a "USA" teach satisfying a "United States" recall). Trust the resource_id.
similar_shape_different_entity:<canonical> (top-level): a structurally matching row exists but its canonical entity differs from the live query's. Treated as cold start; the warning carries the conflicting canonical as a hint, but the row is NOT promoted into Results.
ambiguous_alias (top-level): a single query entity resolved to multiple canonicals (e.g., "Cards" → Arizona Cardinals + St. Louis Cardinals). Surface the ambiguity from context before committing to a resource.
candidates_present (top-level): the envelope carries a candidates section. Handle it via the candidates branch in Step 2 before anything else.
lookup_refresh_available (top-level): an entity in the query has no lookup row yet, but synced data could provide one. Run substack-reader-pp-cli sync to refresh entity lookups.
Top-level no_learnings_for_query_family: the table had no rows above the Jaccard floor. Pure cold start.
Step 4: teach & after finalizing your response - always
Teaching is unconditional. After resolving a query the store could not answer, background-teach the final resource mapping - no call-count threshold, no judging whether it was "worth" learning. The teach is the anchor of the loop: it triggers playbook synthesis for a family without a playbook, and same-referent phrasings fold into one family so near-duplicate teaches do not fragment the store. Fire it after assembling your user-facing response but BEFORE emitting it, with a shell & so the call returns immediately:
Silent on success. Errors only land in teach.log under the resolved state dir. Teach the most specific resource - if the user asked a broad question and you walked through parent records to find the specific answer, teach the leaf id, not the parent. The CLI uses seeded entity_lookups for cross-alias resolution at recall time, so a teach under one alias (e.g., "Niners") satisfies future queries under another alias (e.g., "49ers", "San Francisco") automatically.
PII rule: teach the structural question with identifiers stripped - never include names, emails, phone numbers, account ids, or other personal identifiers in taught queries or notes. The CLI scans teach queries for obvious email/phone shapes and warns, but does not block; strip before teaching rather than relying on the warning.
You do not need to decide whether a session "deserves" a playbook: a teach on a family without one auto-synthesizes a playbook_candidate from the session's journal, and the next session judges it via confirm/reject. Attach explicit playbook flags only when you already hold choreography worth recording verbatim - workarounds the CLI didn't surface (silently-dropped flags, undocumented params, pagination tricks, payload gotchas). Prefer the integrated one-call form - record the resource learning and the playbook in the same teach invocation:
bash
# Common case: record both the resource learning AND the playbook in one call.
substack-reader-pp-cli teach \
--query "<user's question>" \
--resource <id> \
--playbook-file ~/playbooks/<shape>.json \
--playbook-notes-file ~/playbooks/<shape>-notes.md
# (append shell `&` to background it)
# Alternate: playbook-only (no resource to record alongside).
substack-reader-pp-cli teach-playbook \
--query "<user's question>" \
--playbook-file ~/playbooks/<shape>.json \
--notes-file ~/playbooks/<shape>-notes.md
Playbook files are JSON with steps, entity_slots, expected_tool_calls. Notes files are markdown carrying the gotchas verbatim. File-free callers (MCP-only agents) pass the same content inline: --playbook-json and --playbook-notes on the integrated teach form, --playbook-json and --notes on teach-playbook. On the integrated teach form, the playbook flags are optional - omit them entirely for a resource-only teach. On the standalone teach-playbook form, at least one of the playbook and notes flags must be set; both empty is rejected. Playbooks are keyed on the structural query family (entities stripped) so a recipe taught from one entity-shaped query applies to every other query of the same shape, with slots_resolved binding the live query's canonical at recall time.
When you DO find a playbook on a future recall, treat it as ground truth: replay the steps with slots_resolved substitutions, skip the discovery that the choreography already documents, and read notes before any step.
Step 6: playbook amend & when your debug response identifies a correction
If your debug-protocol response identifies a concrete correction the notes or playbook should know — a workaround, an undocumented endpoint shape, a stale field name, observed schema drift, an empty-payload fallback — fire playbook amend BEFORE emitting your user-facing response. Same fire-and-forget posture as teach.
What counts as worth amending: a behavior you OBSERVED this session that future-you would benefit from knowing. Examples worth amending:
A workaround for a CLI surface that silently drops or misorders a flag.
An undocumented endpoint shape (response wrapped in {meta, results}, payload nested two levels deeper than the docs claim).
Observed schema drift (a field renamed, an index that shifted between seasons, a category label that the API now returns lower-cased).
What does NOT belong in notes:
The year-specific or entity-specific answer to the user's question. That's the response, not a learning.
Per-team / per-athlete / per-row data the playbook already retrieves at runtime.
Statements that paraphrase what the existing notes already say.
The amend command appends to the family's existing notes with a timestamped marker ([amend YYYY-MM-DDTHH:MMZ]: <text>). Multiple amends accumulate; the audit trail is visible. If no playbook exists yet for the family, amend creates a notes-only one (so cold-start corrections still land).
PII discipline for amend notes
playbook amend notes are designed to potentially flow upstream as shared knowledge in future versions of the Printing Press. Keep them clean of user-identifying content so the upstream-contribution path stays open without retroactive scrubbing:
Do NOT embed paths to user filesystems, personal API keys or tokens, user email addresses, user GitHub handles, or specific query histories tied to a single user.
Acceptable: endpoint shapes, undocumented field names, API gotchas, observed schema drift, workarounds for CLI surfaces, generalizable pagination or retry tactics.
If a correction is only meaningful with user-specific context, it belongs in a personal note, not in the playbook amend.
Measuring the loop
substack-reader-pp-cli learnings stats reports recall hit rate, teach-to-reuse, playbook resolution rate, and candidate confirm/reject counts from the local learn_events table. Rates are null until they have a denominator; everything stays on this machine. Use it to check whether the loop is earning its keep for this CLI.
Disabling learning
--no-learn on a single command short-circuits both recall and the teach write path. Use for deterministic agent flows or tests that must not be affected by accumulated learnings.
SUBSTACK_READER_NO_LEARN=true in the environment globally disables the pipeline.
Agent Feedback
When you (or the agent) notice something off about this CLI, record it:
substack-reader-pp-cli feedback "the --since flag is inclusive but docs say exclusive"
substack-reader-pp-cli feedback --stdin < notes.txt
substack-reader-pp-cli feedback list --json --limit 10
Entries are stored locally as feedback.jsonl under the resolved data dir. They are never POSTed unless SUBSTACK_READER_FEEDBACK_ENDPOINT is set AND either --send is passed or SUBSTACK_READER_FEEDBACK_AUTO_SEND=true. Default behavior is local-only.
Write what surprised you, not a bug report. Short, specific, one line: that is the part that compounds.
Output Delivery
Every command accepts --deliver <sink>. The output goes to the named sink in addition to (or instead of) stdout, so agents can route command results without hand-piping. Three sinks are supported:
Sink
Effect
stdout
Default; write to stdout only
file:<path>
Atomically write output to <path> (tmp + rename)
webhook:<url>
POST the output body to the URL (application/json or application/x-ndjson when --compact)
Unknown schemes are refused with a structured error naming the supported set. Webhook failures return non-zero and log the URL + HTTP status on stderr.
Named Profiles
A profile is a saved set of flag values, reused across invocations. Use it when a scheduled or recurring agent reuses the same saved flags while providing different input each run.
substack-reader-pp-cli profile save briefing --json
substack-reader-pp-cli --profile briefing categories list
substack-reader-pp-cli profile list --json
substack-reader-pp-cli profile show briefing
substack-reader-pp-cli profile delete briefing --yes
Explicit flags always win over profile values; profile values win over defaults. agent-context lists all available profiles under available_profiles so introspecting agents discover them at runtime.
Exit Codes
Code
Meaning
0
Success
2
Usage error (wrong arguments)
3
Resource not found
5
API error (upstream issue)
7
Rate limited (wait and retry)
10
Config error
Argument Parsing
Parse $ARGUMENTS:
Empty, help, or --help → show substack-reader-pp-cli --help output
Starts with install → ends with mcp → MCP installation; otherwise → see Prerequisites above
Anything else → Direct Use (execute as CLI command with --agent)
MCP Server Installation
Install the MCP server:bash
go install github.com/mvanhorn/printing-press-library/library/media-and-entertainment/substack-reader/cmd/substack-reader-pp-mcp@latest
Register with Claude Code:bash
claude mcp add substack-reader-pp-mcp -- substack-reader-pp-mcp
Verify: claude mcp list
Direct Use
Check if installed: which substack-reader-pp-cli
If not found, offer to install (see Prerequisites at the top of this skill).
Match the user query to the best command from the Unique Capabilities and Command Reference above.
Pp Substack Reader next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
Pp Substack Reader compared with similar skills
Skill
Stars
Used in
Tokens
Auto-check
Licence
Repo updated
Pp Substack Reader this skillmvanhorn/printing-press-library
Advanced content and topic research skill that analyzes trends across Google Analytics, Google Trends, Substack, Medium, Reddit, LinkedIn, X, blogs, podcasts, and YouTube to generate data-driven…
Produce publish-ready Chinese and English writing packages across WeChat articles, Medium/Substack essays, research notes, founder memos, op-eds, speeches, data storytelling, statistical graphics…
Interviews you about objective, audience and context, then designs an end-of-article call to action with copy, placement, A/B test plan and accessibility check.
Analyzes long-form content (transcripts, articles, interviews, webinars, podcast notes) and extracts the 3 best content ideas for any publishing platform — LinkedIn, Substack, blog, newsletter…
The free, offline Trigger phrases: search 1688 for, find a factory on 1688 for, wholesale price on 1688 for, who is the cheapest supplier on 1688 for, compare 1688 suppliers for, use 1688, run 1688.
Inspect known Activity Japan plan IDs or URLs, compare dated prices and sessions, check language-sitemap coverage, and hand off to canonical booking pages.
Read any Substack publication as a local, full-text-searchable corpus — keyless for free posts, your own session for what you subscribe to. Pp Substack Reader is an agent skill from mvanhorn/printing-press-library. Read any Substack publication as a local, full-text-searchable corpus — keyless for free posts, your own session for what you subscribe to.
When should I use Pp Substack Reader?
Pp Substack Reader fits situations like: phrases: archive this Substack; read this Substack post; search my Substack corpus; whats new in my newsletters.
How do I install Pp Substack Reader in Claude Code?
Run `npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a claude-code`. Or copy the skill folder (cli-skills/pp-substack-reader in mvanhorn/printing-press-library) into .claude/skills/pp-substack-reader in your project. Claude Code loads it when a task matches its description.
How do I install Pp Substack Reader in Codex?
Run `npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a codex`. Or copy the skill folder (cli-skills/pp-substack-reader in mvanhorn/printing-press-library) into .agents/skills/pp-substack-reader in your project. Codex loads it when a task matches its description.
Can I use Pp Substack Reader in Cursor, Gemini CLI or GitHub Copilot?
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mvanhorn/printing-press-library --skill pp-substack-reader -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pp-substack-reader, .gemini/skills/pp-substack-reader, .github/skills/pp-substack-reader and .opencode/skills/pp-substack-reader in your project.
What does Pp Substack Reader need to run?
Going by SKILL.md and its folder, Pp Substack Reader needs the command-line tools its instructions call (go, claude and npx). Our summary lists: Node.js. Its frontmatter pre-approves these tools: Read, Bash.
Does Pp Substack Reader access the network?
SKILL.md names 1 domain. In commands or code: substack.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Is Pp Substack Reader safe to install?
Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
What licence does Pp Substack Reader use?
Pp Substack Reader is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
How many tokens does Pp Substack Reader use?
About 7.9k tokens (SKILL.md is roughly 32k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
What are the alternatives to Pp Substack Reader?
Skills that share tags, products or a category with Pp Substack Reader: Content Trend Researcher (alirezarezvani/claude-code-skill-factory, 879 stars), Shakespeare Writing Studio (ZCZHAO-1999/shakespeare-writing-studio, 178 stars), Serenity Reply (leslieyeo/serenity-reply, 136 stars) and End-of-Article CTA Designer (samber/cc-skills, 228 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Who maintains Pp Substack Reader?
mvanhorn (a GitHub user) maintains it in mvanhorn/printing-press-library, which has 2,053 GitHub stars. The repository holds 505 skills in this directory. The repository was last updated on October 6, 2026.