Firecrawl Scrape
firecrawl/skills
Read a known webpage or execute a discovered workflow or data-provider capability.
Printing Press CLI for Firecrawl. An agent skill from mvanhorn/printing-press-library.
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install mvanhorn/printing-press-library pp-firecrawl --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .claude/skills && cp -r skills-src/cli-skills/pp-firecrawl .claude/skills/pp-firecrawl && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .claude/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawlType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install mvanhorn/printing-press-library pp-firecrawl --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .agents/skills && cp -r skills-src/cli-skills/pp-firecrawl .agents/skills/pp-firecrawl && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .agents/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install mvanhorn/printing-press-library pp-firecrawl --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/cli-skills/pp-firecrawl .cursor/skills/pp-firecrawl && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .cursor/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/mvanhorn/printing-press-library.git --path cli-skills/pp-firecrawl--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install mvanhorn/printing-press-library pp-firecrawl --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/cli-skills/pp-firecrawl .gemini/skills/pp-firecrawl && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .gemini/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install mvanhorn/printing-press-library pp-firecrawlInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .github/skills && cp -r skills-src/cli-skills/pp-firecrawl .github/skills/pp-firecrawl && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .github/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install mvanhorn/printing-press-library pp-firecrawl --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/mvanhorn/printing-press-library.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/cli-skills/pp-firecrawl .opencode/skills/pp-firecrawl && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "pp-firecrawl" agent skill from https://github.com/mvanhorn/printing-press-library/tree/main/cli-skills/pp-firecrawl into .opencode/skills/pp-firecrawl/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pp-firecrawl", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
pp-firecrawlPrinting Press CLI for Firecrawl. An agent skill from mvanhorn/printing-press-library.
Pp Firecrawl is an agent skill from mvanhorn/printing-press-library. Printing Press CLI for Firecrawl. API for interacting with Firecrawl services to perform web scraping and crawling tasks.
Its SKILL.md is about 2.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Data & Analytics, covering Web scraping. It works with Firecrawl. The repository describes itself as: Official library of CLIs generated by the CLI Printing Press. Endorsed, tested, and community-contributed. The licence is Apache-2.0.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 0fdcc7a. It shows what the files ask for, not the result of running them.
Pre-approves these tools, so the agent can use them without asking each time:
ReadBashFrom allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
goclaudenpxFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Pp Firecrawl loads about 2.2k tokens when it runs. Until then it costs about 34 tokens; SKILL.md has 900 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
allowed-tools: Read, BashAutomated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from mvanhorn/printing-press-library at commit 0fdcc7a, republished under its Apache-2.0 licence (© mvanhorn). 900 words, ~2,229 tokens.
.claude/skills/pp-firecrawl/SKILL.md (or your agent's skills folder).<!-- GENERATED FILE — DO NOT EDIT.
This file is a verbatim mirror of library/developer-tools/firecrawl/SKILL.md,
regenerated post-merge by tools/generate-skills/. Hand-edits here are
silently overwritten on the next regen. Edit the library/ source instead.
See the repository agent guide, section "Generated artifacts: registry.json, cli-skills/". -->
This skill drives the firecrawl-pp-cli binary. You must verify the CLI is installed before invoking any command from this skill. If it is missing, install it first:
$HOME/.local/bin on macOS/Linux and %LOCALAPPDATA%\Programs\PrintingPress\bin on Windows:npx -y @mvanhorn/printing-press-library install firecrawl --cli-onlyfirecrawl-pp-cli --version$PATH for the agent/runtime that will invoke this skill.If the npx install fails (no Node, offline, etc.), fall back to a direct Go install (requires Go 1.26.6 or newer):
go install github.com/mvanhorn/printing-press-library/library/developer-tools/firecrawl/cmd/firecrawl-pp-cli@latestIf --version reports "command not found" after install, the runtime cannot see the binary directory on $PATH. Do not proceed with skill commands until verification succeeds.
batch — Manage batch
firecrawl-pp-cli batch cancel-scrape — Cancel a batch scrape jobfirecrawl-pp-cli batch get-scrape-errors — Get the errors of a batch scrape jobfirecrawl-pp-cli batch get-scrape-status — Get the status of a batch scrape jobfirecrawl-pp-cli batch scrape-and-extract-from-urls — Scrape multiple URLs and optionally extract information using an LLMcrawl — Manage crawl
firecrawl-pp-cli crawl cancel — Cancel a crawl jobfirecrawl-pp-cli crawl get-active — Get all active crawls for the authenticated teamfirecrawl-pp-cli crawl get-status — Get the status of a crawl jobfirecrawl-pp-cli crawl urls — Crawl multiple URLs based on optionsdeep-research — Manage deep research
firecrawl-pp-cli deep-research get-status — Get the status and results of a deep research operationfirecrawl-pp-cli deep-research start — Start a deep research operation on a queryextract — Manage extract
firecrawl-pp-cli extract data — Extract structured data from pages using LLMsfirecrawl-pp-cli extract get-status — Get the status of an extract jobfirecrawl-search — Manage firecrawl search
firecrawl-pp-cli firecrawl-search — Search and optionally scrape search resultsllmstxt — Manage llmstxt
firecrawl-pp-cli llmstxt generate-llms-txt — Generate LLMs.txt for a websitefirecrawl-pp-cli llmstxt get-llms-txt-status — Get the status and results of an LLMs.txt generation jobmap — Manage map
firecrawl-pp-cli map — Map multiple URLs based on optionsscrape — Manage scrape
firecrawl-pp-cli scrape — Scrape a single URL and optionally extract information using an LLMteam — Manage team
firecrawl-pp-cli team get-credit-usage — Get remaining credits for the authenticated teamfirecrawl-pp-cli team get-token-usage — Get remaining tokens for the authenticated team (Extract only)When you know what you want to do but not which command does it, ask the CLI directly:
firecrawl-pp-cli which "<capability in your own words>"which resolves a natural-language capability query to the best matching command from this CLI's curated feature index. Exit code 0 means at least one match; exit code 2 means no confident match — fall back to --help or use a narrower query.
Store your access token:
firecrawl-pp-cli auth set-token YOUR_TOKEN_HEREOr set FIRECRAWL_BEARER_AUTH as an environment variable.
Run firecrawl-pp-cli doctor to verify setup.
Add --agent to any command. Expands to: --json --compact --no-input --no-color --yes.
Pipeable — JSON on stdout, errors on stderr
Filterable — --select keeps a subset of fields. Dotted paths descend into nested structures; arrays traverse element-wise. Critical for keeping context small on verbose APIs:
firecrawl-pp-cli batch cancel-scrape mock-value --agent --select id,name,statusPreviewable — --dry-run shows the request without sending
Offline-friendly — sync/search commands can use the local SQLite store when available
Non-interactive — never prompts, every input is a flag
Commands that read from the local store or the API wrap output in a provenance envelope:
{
"meta": {"source": "live" | "local", "synced_at": "...", "reason": "..."},
"results": <data>
}Parse .results for data and .meta.source to know whether it's live or local. A human-readable N results (live) summary is printed to stderr only when stdout is a terminal — piped/agent consumers get pure JSON on stdout.
When you (or the agent) notice something off about this CLI, record it:
firecrawl-pp-cli feedback "the --since flag is inclusive but docs say exclusive"
firecrawl-pp-cli feedback --stdin < notes.txt
firecrawl-pp-cli feedback list --json --limit 10Entries are stored locally at ~/.firecrawl-pp-cli/feedback.jsonl. They are never POSTed unless FIRECRAWL_FEEDBACK_ENDPOINT is set AND either --send is passed or FIRECRAWL_FEEDBACK_AUTO_SEND=true. Default behavior is local-only.
Write what surprised you, not a bug report. Short, specific, one line: that is the part that compounds.
Every command accepts --deliver <sink>. The output goes to the named sink in addition to (or instead of) stdout, so agents can route command results without hand-piping. Three sinks are supported:
| Sink | Effect |
|---|---|
stdout | Default; write to stdout only |
file:<path> | Atomically write output to <path> (tmp + rename) |
webhook:<url> | POST the output body to the URL (application/json or application/x-ndjson when --compact) |
Unknown schemes are refused with a structured error naming the supported set. Webhook failures return non-zero and log the URL + HTTP status on stderr.
A profile is a saved set of flag values, reused across invocations. Use it when a scheduled agent calls the same command every run with the same configuration - HeyGen's "Beacon" pattern.
firecrawl-pp-cli profile save briefing --json
firecrawl-pp-cli --profile briefing batch cancel-scrape mock-value
firecrawl-pp-cli profile list --json
firecrawl-pp-cli profile show briefing
firecrawl-pp-cli profile delete briefing --yesExplicit flags always win over profile values; profile values win over defaults. agent-context lists all available profiles under available_profiles so introspecting agents discover them at runtime.
| Code | Meaning |
|---|---|
| 0 | Success |
| 2 | Usage error (wrong arguments) |
| 3 | Resource not found |
| 4 | Authentication required |
| 5 | API error (upstream issue) |
| 7 | Rate limited (wait and retry) |
| 10 | Config error |
Parse $ARGUMENTS:
help, or --help → show firecrawl-pp-cli --help outputinstall → ends with mcp → MCP installation; otherwise → see Prerequisites above--agent)go install github.com/mvanhorn/printing-press-library/library/other/firecrawl-pp-cli/cmd/firecrawl-pp-mcp@latestclaude mcp add firecrawl-pp-mcp -- firecrawl-pp-mcpclaude mcp listwhich firecrawl-pp-cli
If not found, offer to install (see Prerequisites at the top of this skill).--agent flag:firecrawl-pp-cli <command> [subcommand] [args] --agentfirecrawl-pp-cli <command> --help.© mvanhorn, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in cli-skills/pp-firecrawl of mvanhorn/printing-press-library.
Open the folder on GitHubat commit 0fdcc7a
Pp Firecrawl next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Pp Firecrawl this skillmvanhorn/printing-press-library | 2.1k | — | ~2.2k | Automated safety check: Notes | Apache-2.0 | |
| Firecrawl Scrapefirecrawl/skills | 116 | — | ~1.8k | Automated safety check: Pass | ISC | |
| Firecrawl Agentfirecrawl/skills | 116 | — | ~1.2k | Automated safety check: Pass | ISC | |
| Firecrawl Company Directoriesfirecrawl/skills | 116 | — | ~557 | Automated safety check: Pass | ISC | |
| Firecrawl Competitive Intelfirecrawl/skills | 116 | — | ~604 | Automated safety check: Pass | ISC | |
| SEO Firecrawlseranking/seo-skills | 160 | — | ~2.3k | Automated safety check: Pass | MIT |
firecrawl/skills
Read a known webpage or execute a discovered workflow or data-provider capability.
firecrawl/skills
Autonomously navigate websites and extract structured data across pages.
firecrawl/skills
Extract structured company lists from directories with Firecrawl.
firecrawl/skills
Monitor competitor pricing, features, changelogs, dashboards, and product changes with Firecrawl.
seranking/seo-skills
Ad-hoc web scraping, site mapping, and full-site crawling via Firecrawl MCP.
firecrawl/skills
Bulk-extract many pages from one site or section. An agent skill from firecrawl/skills.
mvanhorn/printing-press-library
Desktop automation through the real Rust agent-desktop CLI, published in Printing Press through a small bridge.
mvanhorn/printing-press-library
Search, browse, and download Google Fonts from the terminal via the gfonts CLI.
mvanhorn/printing-press-library
The free, offline Trigger phrases: search 1688 for, find a factory on 1688 for, wholesale price on 1688 for, who is the cheapest supplier on 1688 for, compare 1688 suppliers for, use 1688, run 1688.
mvanhorn/printing-press-library
Inspect known Activity Japan plan IDs or URLs, compare dated prices and sessions, check language-sitemap coverage, and hand off to canonical booking pages.
mvanhorn/printing-press-library
Every Admin By Request portal action, plus a local SQLite mirror of audit, events, inventory and requests for ad-hoc...
mvanhorn/printing-press-library
macOS screen capture, window recording, GIF conversion, and agent evidence bundles from the terminal.
Works with
Categories
Printing Press CLI for Firecrawl. An agent skill from mvanhorn/printing-press-library. Pp Firecrawl is an agent skill from mvanhorn/printing-press-library. Printing Press CLI for Firecrawl.
Pp Firecrawl fits situations like: tasks that involve Web scraping.
Run `npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a claude-code`. Or copy the skill folder (cli-skills/pp-firecrawl in mvanhorn/printing-press-library) into .claude/skills/pp-firecrawl in your project. Claude Code loads it when a task matches its description.
Run `npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a codex`. Or copy the skill folder (cli-skills/pp-firecrawl in mvanhorn/printing-press-library) into .agents/skills/pp-firecrawl in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mvanhorn/printing-press-library --skill pp-firecrawl -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pp-firecrawl, .gemini/skills/pp-firecrawl, .github/skills/pp-firecrawl and .opencode/skills/pp-firecrawl in your project.
Going by SKILL.md and its folder, Pp Firecrawl needs the command-line tools its instructions call (go, claude and npx). Our summary lists: Node.js. Its frontmatter pre-approves these tools: Read, Bash.
SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
Pp Firecrawl is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.2k tokens (SKILL.md is roughly 8.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Pp Firecrawl: Firecrawl Scrape (firecrawl/skills, 116 stars), Firecrawl Agent (firecrawl/skills, 116 stars), Firecrawl Company Directories (firecrawl/skills, 116 stars) and Firecrawl Competitive Intel (firecrawl/skills, 116 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
mvanhorn (a GitHub user) maintains it in mvanhorn/printing-press-library, which has 2,053 GitHub stars. The repository holds 505 skills in this directory. The repository was last updated on October 7, 2026.
Source: mvanhorn/printing-press-library on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.