Agent Readiness Audit
indranilbanerjee/digital-marketing-pro
Audit whether AI agents and AI crawlers can actually use a site — robots.txt rules per AI crawler token (OpenAI, Anthropic and Perplexity bots, Google-Extended, Applebot-Extended)…
The first CLI for Scrape-do with Google SERP scraping plus a credit and concurrency governor Trigger phrases: scrape google for, get google search results for, track keyword rank for, scrape this…
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install mvanhorn/printing-press-library pp-scrape-do --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .claude/skills && cp -r skills-src/cli-skills/pp-scrape-do .claude/skills/pp-scrape-do && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .claude/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-doType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install mvanhorn/printing-press-library pp-scrape-do --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .agents/skills && cp -r skills-src/cli-skills/pp-scrape-do .agents/skills/pp-scrape-do && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .agents/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install mvanhorn/printing-press-library pp-scrape-do --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/cli-skills/pp-scrape-do .cursor/skills/pp-scrape-do && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .cursor/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/mvanhorn/printing-press-library.git --path cli-skills/pp-scrape-do--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install mvanhorn/printing-press-library pp-scrape-do --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/cli-skills/pp-scrape-do .gemini/skills/pp-scrape-do && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .gemini/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install mvanhorn/printing-press-library pp-scrape-doInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .github/skills && cp -r skills-src/cli-skills/pp-scrape-do .github/skills/pp-scrape-do && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .github/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install mvanhorn/printing-press-library pp-scrape-do --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/cli-skills/pp-scrape-do .opencode/skills/pp-scrape-do && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "pp-scrape-do" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-scrape-do into .opencode/skills/pp-scrape-do/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-scrape-do", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
pp-scrape-doThe first CLI for Scrape-do with Google SERP scraping plus a credit and concurrency governor Trigger phrases: scrape google for, get google search results for, track keyword rank for, scrape this…
Pp Scrape Do is an agent skill from mvanhorn/printing-press-library. The first CLI for Scrape-do with Google SERP scraping plus a credit and concurrency governor Trigger phrases: scrape google for, get google search results for, track keyword rank for, scrape this url with scrape.do, estimate scrape.do cost for, check my scrape.do credits, use scrape-do, run scrape-do.
Its SKILL.md is about 3.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Data & Analytics, covering Web scraping and Web search. The repository describes itself as: Official library of CLIs generated by the CLI Printing Press. Endorsed, tested, and community-contributed. The licence is Apache-2.0.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 0fdcc7a. It shows what the files ask for, not the result of running them.
Pre-approves these tools, so the agent can use them without asking each time:
ReadBashFrom allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
goclaudenpxFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
linkedin.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
SCRAPEDO_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Pp Scrape Do loads about 3.2k tokens when it runs. Until then it costs about 83 tokens; SKILL.md has 1,400 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
allowed-tools: Read, BashAutomated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from mvanhorn/printing-press-library at commit 0fdcc7a, republished under its Apache-2.0 licence (© mvanhorn). 1,400 words, ~3,234 tokens.
.claude/skills/pp-scrape-do/SKILL.md (or your agent's skills folder).<!-- GENERATED FILE — DO NOT EDIT.
This file is a verbatim mirror of library/developer-tools/scrape-do/SKILL.md,
regenerated post-merge by tools/generate-skills/. Hand-edits here are
silently overwritten on the next regen. Edit the library/ source instead.
See the repository agent guide, section "Generated artifacts: registry.json, cli-skills/". -->
This skill drives the scrape-do-pp-cli binary. You must verify the CLI is installed before invoking any command from this skill. If it is missing, install it first:
npx -y @mvanhorn/printing-press-library install scrape-do --cli-onlyscrape-do-pp-cli --version$PATH. On Windows this is %LOCALAPPDATA%\Programs\PrintingPress\bin; on macOS/Linux it is typically $HOME/.local/bin or a platform-specific location printed by the installer. If you used the Go fallback instead, add $GOPATH/bin (or $HOME/go/bin).If the npx install fails (no Node, offline, etc.), fall back to a direct Go install (requires Go 1.26.6 or newer):
go install github.com/mvanhorn/printing-press-library/library/developer-tools/scrape-do/cmd/scrape-do-pp-cli@latestIf --version reports "command not found" after install, the install step did not put the binary on $PATH. Do not proceed with skill commands until verification succeeds.
Every Scrape.do surface — the core scraper plus the whole Google family (search, maps, news, shopping, flights, hotels, trends) — wrapped with an offline SQLite history. The governor estimates credit cost before every call with cost, debits a local ledger from the authoritative cost header with budget, and gates concurrent requests against your plan's live ceiling so an agent swarm never 429s itself or burns the monthly budget. SERPs become queryable history: drift and movers surface rank changes offline with no re-spend.
Reach for this CLI whenever a task needs Google search results, structured SERP data, or any proxied web scrape via Scrape.do — and especially when several agents share one Scrape.do account. Its governor (cost/budget/batch) keeps concurrent usage inside the plan's limits and the monthly credit budget, and its offline SERP history (drift/movers) answers rank-change questions without re-spending credits. Prefer it over raw curl calls when you care about credit cost, concurrency safety, or comparing results over time.
Do not activate this CLI for requests that require creating, updating, deleting, publishing, commenting, upvoting, inviting, ordering, sending messages, booking, purchasing, or changing remote state. No command mutates remote target state. Note, however, that scrape, google, and batch make billed GET requests that consume Scrape.do account credits — use cost to estimate before spending and budget to track and cap spend.
These capabilities aren't available in any other tool for this API.
drift — Diff a Google query's two most recent stored SERPs and see exactly which results moved, appeared, or dropped — entirely offline, no credits spent.
When an agent needs to know whether a ranking changed since last check, this answers it without re-spending a 10-credit SERP call.
scrape-do-pp-cli drift "best crm software" --jsonmovers — Scan every tracked query's latest-versus-previous SERP snapshot and surface only the queries whose top positions moved past a threshold.
Turns hundreds of tracked keywords into a short 'what changed this week' list an agent can act on.
scrape-do-pp-cli movers --threshold 3 --agentbatch — Dispatch a list of URLs or queries through a shared concurrency lease and per-call credit ledger, auto-retrying only the non-billed 429/502/510 classes and stopping before a credit ceiling.
Lets an agent swarm hammer one account at full speed without 429 storms or blowing the monthly budget.
scrape-do-pp-cli batch --input urls.txt --max-credits 500 --agentbudget — Attribute spend by mode and query-family from a local ledger debited off the authoritative per-call cost header, joined with cached account state to forecast burn-rate against days remaining.
Tells an agent how much budget is left and which workloads are eating it before the account hits a hard 401.
scrape-do-pp-cli budget --agentcost — Print the exact credit cost a request will incur — accounting for render, super-proxy, Google endpoints, and per-domain overrides — before spending a single credit.
Lets an agent compare the cost of cheap vs expensive scrape modes and pick the cheapest path that works.
scrape-do-pp-cli cost --url https://www.linkedin.com/company/example --render --superaccount — Live Scrape.do account state: subscription status, concurrency allowance and headroom, monthly credit cap and remaining credits.
scrape-do-pp-cli account — Fetch live account state: IsActive, ConcurrentRequest, RemainingConcurrentRequest, MaxMonthlyRequestWhen you know what you want to do but not which command does it, ask the CLI directly:
scrape-do-pp-cli which "<capability in your own words>"which resolves a natural-language capability query to the best matching command from this CLI's curated feature index. Exit code 0 means at least one match; exit code 2 means no confident match — fall back to --help or use a narrower query.
scrape-do-pp-cli google search "coffee makers" --agent --select organic_results.position,organic_results.title,organic_results.linkThe SERP JSON is large and deeply nested; --select with dotted paths returns just rank, title, and link so an agent doesn't burn context parsing the full payload.
scrape-do-pp-cli cost --url https://www.linkedin.com/company/example --render --superPrints the expected credits (LinkedIn domain override + render + super proxy) with no API spend, so you can choose the cheapest mode that works.
scrape-do-pp-cli batch --input urls.txt --max-credits 500 --agentDispatches every URL through the shared concurrency lease, retries only the non-billed 429/502/510 classes, and stops before the 500-credit ceiling.
scrape-do-pp-cli movers --threshold 3 --agentCompares each tracked query's two latest stored SERPs and lists only the queries whose top positions moved by 3 or more — entirely offline.
scrape-do-pp-cli sql "SELECT domain, COUNT(*) c FROM serp_organic GROUP BY domain ORDER BY c DESC LIMIT 10"Read-only SQL over the local store — share-of-voice by domain across every stored SERP, with no credit spent.
Scrape.do uses a single API token passed as the token query parameter. Set it as SCRAPEDO_API_KEY in your environment and the CLI never logs the value. The same token drives the core scraper and every Google and Ready-API endpoint.
Run scrape-do-pp-cli doctor to verify setup.
Add --agent to any command. Expands to: --json --compact --no-input --no-color --yes.
Pipeable — JSON on stdout, errors on stderr
Filterable — --select keeps a subset of fields. Dotted paths descend into nested structures; arrays traverse element-wise. Critical for keeping context small on verbose APIs:
scrape-do-pp-cli account --agent --select id,name,statusPreviewable — --dry-run shows the request without sending
Offline-friendly — sync/search commands can use the local SQLite store when available
Non-interactive — never prompts, every input is a flag
Read-only — do not use this CLI for create, update, delete, publish, comment, upvote, invite, order, send, or other mutating requests
Commands that read from the local store or the API wrap output in a provenance envelope:
{
"meta": {"source": "live" | "local", "synced_at": "...", "reason": "..."},
"results": <data>
}Parse .results for data and .meta.source to know whether it's live or local. A human-readable N results (live) summary is printed to stderr only when stdout is a terminal AND no machine-format flag (--json, --csv, --compact, --quiet, --plain, --select) is set — piped/agent consumers and explicit-format runs get pure JSON on stdout.
When you (or the agent) notice something off about this CLI, record it:
scrape-do-pp-cli feedback "the --since flag is inclusive but docs say exclusive"
scrape-do-pp-cli feedback --stdin < notes.txt
scrape-do-pp-cli feedback list --json --limit 10Entries are stored locally at ~/.local/share/scrape-do-pp-cli/feedback.jsonl. They are never POSTed unless SCRAPE_DO_FEEDBACK_ENDPOINT is set AND either --send is passed or SCRAPE_DO_FEEDBACK_AUTO_SEND=true. Default behavior is local-only.
Write what surprised you, not a bug report. Short, specific, one line: that is the part that compounds.
Every command accepts --deliver <sink>. The output goes to the named sink in addition to (or instead of) stdout, so agents can route command results without hand-piping. Three sinks are supported:
| Sink | Effect |
|---|---|
stdout | Default; write to stdout only |
file:<path> | Atomically write output to <path> (tmp + rename) |
webhook:<url> | POST the output body to the URL (application/json or application/x-ndjson when --compact) |
Unknown schemes are refused with a structured error naming the supported set. Webhook failures return non-zero and log the URL + HTTP status on stderr.
A profile is a saved set of flag values, reused across invocations. Use it when a scheduled agent calls the same command every run with the same configuration - HeyGen's "Beacon" pattern.
scrape-do-pp-cli profile save briefing --json
scrape-do-pp-cli --profile briefing account
scrape-do-pp-cli profile list --json
scrape-do-pp-cli profile show briefing
scrape-do-pp-cli profile delete briefing --yesExplicit flags always win over profile values; profile values win over defaults. agent-context lists all available profiles under available_profiles so introspecting agents discover them at runtime.
| Code | Meaning |
|---|---|
| 0 | Success |
| 2 | Usage error (wrong arguments) |
| 3 | Resource not found |
| 4 | Authentication required |
| 5 | API error (upstream issue) |
| 7 | Rate limited (wait and retry) |
| 10 | Config error |
Parse $ARGUMENTS:
help, or --help → show scrape-do-pp-cli --help outputinstall → ends with mcp → MCP installation; otherwise → see Prerequisites above--agent)go install github.com/mvanhorn/printing-press-library/library/developer-tools/scrape-do/cmd/scrape-do-pp-mcp@latestclaude mcp add scrape-do-pp-mcp -- scrape-do-pp-mcpclaude mcp listwhich scrape-do-pp-cli
If not found, offer to install (see Prerequisites at the top of this skill).--agent flag:scrape-do-pp-cli <command> [subcommand] [args] --agentscrape-do-pp-cli <command> --help.© mvanhorn, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in cli-skills/pp-scrape-do of mvanhorn/printing-press-library.
Open the folder on GitHubat commit 0fdcc7a
Pp Scrape Do next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Pp Scrape Do this skillmvanhorn/printing-press-library | 2.1k | — | ~3.2k | Automated safety check: Notes | Apache-2.0 | |
| Agent Readiness Auditindranilbanerjee/digital-marketing-pro | 855 | 1 repos | ~3.9k | Automated safety check: Pass | MIT | |
| Web Search Scraper API Skillbrowser-act/skills | 6.1k | 1 repos | ~1.3k | Automated safety check: Pass | MIT | |
| Keirouter Web Fetchmydisha/keirouter | 147 | — | ~741 | Automated safety check: Pass | MIT | |
| Deepapidavidondrej/skills | 4.1k | — | ~2.5k | Automated safety check: Pass | MIT | |
| Bright Data Best Practicesbrightdata/skills | 264 | 1 repos | ~3.6k | Automated safety check: Pass | MIT |
indranilbanerjee/digital-marketing-pro
Audit whether AI agents and AI crawlers can actually use a site — robots.txt rules per AI crawler token (OpenAI, Anthropic and Perplexity bots, Google-Extended, Applebot-Extended)…
browser-act/skills
This skill helps users automatically extract complete Markdown content from any website via the BrowserAct Web Search Scraper API.
mydisha/keirouter
Fetch URL → markdown / text / HTML via KeiRouter /v1/web/fetch using Firecrawl / Jina Reader / Tavily Extract / Exa Contents.
davidondrej/skills
Use DeepAPI for all web search, deep research, and web scraping (websites, LinkedIn, GitHub, X/Twitter, YouTube, Instagram) instead of built-in search, research, fetch, or browser tools.
brightdata/skills
Build production-ready Bright Data integrations with best practices baked in.
brightdata/skills
Search the web via the Bright Data CLI — bdata search for Google/Bing/Yandex SERP, bdata discover for intent-ranked semantic results.
mvanhorn/printing-press-library
Desktop automation through the real Rust agent-desktop CLI, published in Printing Press through a small bridge.
mvanhorn/printing-press-library
Search, browse, and download Google Fonts from the terminal via the gfonts CLI.
mvanhorn/printing-press-library
The free, offline Trigger phrases: search 1688 for, find a factory on 1688 for, wholesale price on 1688 for, who is the cheapest supplier on 1688 for, compare 1688 suppliers for, use 1688, run 1688.
mvanhorn/printing-press-library
Inspect known Activity Japan plan IDs or URLs, compare dated prices and sessions, check language-sitemap coverage, and hand off to canonical booking pages.
mvanhorn/printing-press-library
Every Admin By Request portal action, plus a local SQLite mirror of audit, events, inventory and requests for ad-hoc...
mvanhorn/printing-press-library
macOS screen capture, window recording, GIF conversion, and agent evidence bundles from the terminal.
Categories
The first CLI for Scrape-do with Google SERP scraping plus a credit and concurrency governor Trigger phrases: scrape google for, get google search results for, track keyword rank for, scrape this…. Pp Scrape Do is an agent skill from mvanhorn/printing-press-library.do credits, use scrape-do, run scrape-do.
Pp Scrape Do fits situations like: phrases: scrape google for; get google search results for; track keyword rank for; scrape this url with scrape.do.
Run `npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a claude-code`. Or copy the skill folder (cli-skills/pp-scrape-do in mvanhorn/printing-press-library) into .claude/skills/pp-scrape-do in your project. Claude Code loads it when a task matches its description.
Run `npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a codex`. Or copy the skill folder (cli-skills/pp-scrape-do in mvanhorn/printing-press-library) into .agents/skills/pp-scrape-do in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mvanhorn/printing-press-library --skill pp-scrape-do -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pp-scrape-do, .gemini/skills/pp-scrape-do, .github/skills/pp-scrape-do and .opencode/skills/pp-scrape-do in your project.
Going by SKILL.md and its folder, Pp Scrape Do needs the command-line tools its instructions call (go, claude and npx) and credentials named SCRAPEDO_API_KEY. Our summary lists: Node.js; A credential in SCRAPEDO_API_KEY. Its frontmatter pre-approves these tools: Read, Bash.
SKILL.md names 1 domain. In commands or code: linkedin.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
Pp Scrape Do is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.2k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Pp Scrape Do: Agent Readiness Audit (indranilbanerjee/digital-marketing-pro, 855 stars), Web Search Scraper API Skill (browser-act/skills, 6.1k stars), Keirouter Web Fetch (mydisha/keirouter, 147 stars) and Deepapi (davidondrej/skills, 4.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
mvanhorn (a GitHub user) maintains it in mvanhorn/printing-press-library, which has 2,053 GitHub stars. The repository holds 505 skills in this directory. The repository was last updated on October 7, 2026.
Source: mvanhorn/printing-press-library on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.