Agent skill

Firecrawl Scrape

by firecrawl in firecrawl/skills

Read a known webpage or execute a discovered workflow or data-provider capability.

ISCAuto-check passedData & Analytics

Install Firecrawl Scrape

skills CLI
$ npx skills add firecrawl/skills --skill firecrawl-scrape -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install firecrawl/skills firecrawl-scrape --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/firecrawl/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/core/firecrawl-scrape .claude/skills/firecrawl-scrape && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
firecrawl-scrape
GitHub stars
115
Token cost
~1.8k tokens
SKILL.md length
774 words
Files
2 (incl. references)
Skills in repo
29
Repo updated
First seen
Licence
ISC

At a glance

Read a known webpage or execute a discovered workflow or data-provider capability.

  • Structured results once the URL
  • SKILL.md covers Quick start, Find tools, inspect inputs,…, Execution and large results and PDFs and page budgets, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Tool is selected

What it does

Firecrawl Scrape is an agent skill from firecrawl/skills. Read a known webpage or execute a discovered workflow or data-provider capability. Use for page content or structured results once the URL or tool is selected.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/large-results.md`).

It sits in Data & Analytics, covering Web scraping. It works with Firecrawl. The licence is ISC.

When your agent uses it

  • Structured results once the URL
  • Tool is selected

Example prompts

  • “/firecrawl-scrape”

Requirements

  • Node.js
  • Pre-approved tools (allowed-tools): Bash(firecrawl *), Bash(npx firecrawl-cli *)

What it can do on your machine

Read from SKILL.md and the folder at commit d0ffd7c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(firecrawl *)
    • Bash(npx firecrawl-cli *)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Firecrawl Scrape loads about 1.8k tokens when it runs, and up to ~2.7k if it reads all its reference files. Until then it costs about 44 tokens; SKILL.md has 774 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~44
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from firecrawl/skills at commit d0ffd7c, republished under its ISC licence (© firecrawl). 774 words, ~1,805 tokens.

Download SKILL.mdSave it as .claude/skills/firecrawl-scrape/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
firecrawl-scrape
description
Read a known webpage or execute a discovered workflow or data-provider capability. Use for page content or structured results once the URL or tool is selected.
allowed-tools
Bash(firecrawl *), Bash(npx firecrawl-cli *)

firecrawl scrape

Read a URL for page content, or execute a selected provider tool for structured data. Discover tools with search and inspect their inputs with list before execution. Multiple URLs can be scraped concurrently.

For structured datasets, first check for a suitable workflow or data provider using the search skill. Read a known page directly; reuse a selected contract instead of repeating discovery.

Quick start

bash
# Basic markdown extraction
firecrawl scrape "<url>" -o .firecrawl/page.md

# Main content only, no nav/footer
firecrawl scrape "<url>" --only-main-content -o .firecrawl/page.md

# Wait for JS to render, then scrape
firecrawl scrape "<url>" --wait-for 3000 -o .firecrawl/page.md

# Multiple URLs (markdown only; each saved to .firecrawl/; -o is ignored)
firecrawl scrape https://example.com https://example.com/blog https://example.com/docs

# Get markdown and links together
firecrawl scrape "<url>" --format markdown,links -o .firecrawl/page.json

# Ask a question about the page
firecrawl scrape "https://example.com/pricing" --query "What is the enterprise plan price?"

Run firecrawl scrape --help for the full option list.

Done when: the page content or provider result has been checked for errors and inspected in bounded sections to answer the request. Preserve source links and disclose partial results.

Find tools, inspect inputs, and get help

Use the CLI help to check supported options rather than guessing:

bash
firecrawl search --help
firecrawl list --help
firecrawl scrape --help

Domain discovery with --domain-tools returns tool summaries by default. Add --tool-detail full for contracts upfront, or inspect one selected tool with list as shown below. Use --tool-detail compact for only provider, capability and description; inspect by those two IDs with list. Summary remains the default. Prefer full when several related contracts will be needed immediately.

For structured data, search for the task, inspect a matching tool's contract, then execute with the exact input fields it declares:

bash
# Web + domain matching + semantic tools
firecrawl search '<user question>'

# Semantic tools only
firecrawl search alexandria '<user question>'

# Categories → providers → tools → contract
firecrawl list
firecrawl list <category-id> --category
firecrawl list <provider-id>
firecrawl list <provider-id> <capability-id> --pretty

# Execute a tool
firecrawl scrape <provider-id>/<capability-id> --options '<JSON matching the selected contract>'

Normal search includes web results and tool matches; search alexandria searches tools only. list <provider> <capability> --pretty shows the selected contract; use --json for machine-readable output. To browse progressively, use list, then list <category> --category, then list <provider>. Search and list do not execute the selected provider tool. Read only the contracts needed for the task; use returned identifiers rather than guessing them.

Read the expanded contract before building inputs or parsing results:

  • required: true requires that input; each requiresOneOf group requires at least one member, not all of them.
  • Selected-contract inspection already requests examples. Read the singular example.request and example.response when present; an empty request can be valid for tools with optional inputs.
  • response.key identifies the records field inside data.alexandria[i].data; an empty key means that data object itself. Do not assume every provider returns records.
  • Provider pagination differs from catalogue next: use the contract's continuation input and the returned page/cursor, preserve filters, and stop at its exhaustion signal. paginated: true alone does not specify that mapping.

Execution and large results

URL scraping does not execute provider tools automatically. Use exact discovered input fields and resolve record IDs with lookup tools rather than inventing them. Check each data.alexandria[] result for errors, not just the outer success flag.

If the client reports an output/context limit, the upstream request may have succeeded. Preserve the request or scrape ID and recover the retained result before repeating the provider call. For large datasets and PDFs, save output with --json -o when a local filesystem is available and inspect bounded sections with jq or other file tools. Keep stderr separate from JSON stdout; do not merge streams with 2>&1 when piping to a JSON parser. Where remote processing is preferable, use firecrawl scrape firecrawl/bash to select from a retained result. Read large-result recovery for IDs, command examples, expiry, and errors. This is explicit recovery, not automatic overflow detection.

Show full SKILL.md (263 more words)Show less

PDFs and page budgets

PDFs cost 1 credit per parsed page. Use --max-pages (an integer from 1 to 10000) to limit PDF parsing, especially for large or unknown documents:

bash
firecrawl scrape "https://example.com/report.pdf" --max-pages 5 --json -o .firecrawl/report.json

The cap applies to each PDF, not the whole command or total credits. Extra formats and options can add charges. The CLI does not quote page counts or costs before execution. Use JSON output to inspect the returned metadata.numPages (parsed), metadata.totalPages (document total), and metadata.creditsUsed when present; a smaller parsed count means the result is partial.

Tips

  • Prefer plain scrape over --query. Scrape to a file, then use grep, head, or read the markdown directly — you can search and reason over the full content yourself. Use --query only when you want a single targeted answer without saving the page (costs 5 extra credits).
  • Scrape handles static pages and JS-rendered SPAs. Escalate to interact when the page needs interaction (clicks, form fills, pagination) or scrape misses content.
  • Multiple URLs are scraped concurrently — check firecrawl --status for your concurrency limit. This mode saves markdown only and ignores -o; other requested formats are dropped. If markdown wasn't requested, the whole JSON response is written into the .md file.
  • Single format outputs raw content. Multiple formats (e.g., --format markdown,links) output JSON.
  • Always quote URLs — shell interprets ? and & as special characters.
  • Naming convention: .firecrawl/{site}-{path}.md

See also

© firecrawl, ISC. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in skills/core/firecrawl-scrape of firecrawl/skills.

  • SKILL.md
  • references/large-results.md

Open the folder on GitHubat commit d0ffd7c

Compare with similar skills

Firecrawl Scrape next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Firecrawl Scrape compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Firecrawl Scrape this skillfirecrawl/skills115—~1.8kAutomated safety check: PassISC
SEO Firecrawlseranking/seo-skills160—~2.3kAutomated safety check: PassMIT
Keirouter Web Fetchmydisha/keirouter147—~741Automated safety check: PassMIT
Firecrawl Scraperdavila7/claude-code-templates32k7 repos~253Automated safety check: PassMIT
Firecrawlzapier/connectors176—~4.2kAutomated safety check: PassElastic-2.0
Firecrawl Scraperynulihao/AgentSkillOS6171 repos~4.1kAutomated safety check: NotesMIT

Similar skills

  • SEO Firecrawl

    seranking/seo-skills

    Ad-hoc web scraping, site mapping, and full-site crawling via Firecrawl MCP.

    160 GitHub stars~2.3k tokensUpdated 3 mo ago
    Data & AnalyticsAuto-check passed
  • Keirouter Web Fetch

    mydisha/keirouter

    Fetch URL → markdown / text / HTML via KeiRouter /v1/web/fetch using Firecrawl / Jina Reader / Tavily Extract / Exa Contents.

    147 GitHub stars~741 tokensUpdated 26 days ago
    Data & AnalyticsAuto-check passed
  • Firecrawl Scraper

    davila7/claude-code-templates

    Deep web scraping, screenshots, PDF parsing, and website crawling using Firecrawl API

    32k GitHub starsUsed in 7 repos~253 tokens
    Data & AnalyticsAuto-check passed
  • Firecrawl

    zapier/connectors

    Official

    Agent-callable Firecrawl tools — scrape a URL to clean Markdown, crawl a site, search the web, map site URLs, and extract structured data.

    176 GitHub stars~4.2k tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Firecrawl Scraper

    ynulihao/AgentSkillOS

    Complete knowledge domain for Firecrawl v2 API - web scraping and crawling that converts websites into LLM-ready markdown or structured data.

    617 GitHub starsUsed in 1 repo~4.1k tokens
    Data & AnalyticsAuto-check: notes
  • Firecrawl

    Prism-Shadow/penguin-harness

    Search the web and scrape pages into clean markdown with the Firecrawl API — query-based discovery, single-URL extraction including public PDFs, driven by curl with a vault-stored API key.

    2.5k GitHub stars~902 tokensUpdated today
    Data & AnalyticsAuto-check passed

More from firecrawl/skills

All 29 skills in this repo
  • Firecrawl Agent

    firecrawl/skills

    Autonomously navigate websites and extract structured data across pages.

    115 GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Extract structured company lists from directories with Firecrawl.

    115 GitHub stars~557 tokensUpdated today
    Auto-check passed
  • Monitor competitor pricing, features, changelogs, dashboards, and product changes with Firecrawl.

    115 GitHub stars~604 tokensUpdated today
    Auto-check passed
  • Firecrawl Crawl

    firecrawl/skills

    Bulk-extract many pages from one site or section. An agent skill from firecrawl/skills.

    115 GitHub stars~467 tokensUpdated today
    Auto-check passed
  • Pull metrics from analytics dashboards and internal web tools with Firecrawl browser.

    115 GitHub stars~590 tokensUpdated today
    Auto-check passed
  • Firecrawl Deep Research

    firecrawl/skills

    Produce an intensive, cited analytical report: executive summary, multi-angle findings, contrarian views, open questions, and full sources.

    115 GitHub stars~1.4k tokensUpdated today
    Auto-check passed

Works with

Questions about Firecrawl Scrape

What does Firecrawl Scrape do?

Read a known webpage or execute a discovered workflow or data-provider capability. Firecrawl Scrape is an agent skill from firecrawl/skills. Read a known webpage or execute a discovered workflow or data-provider capability.

When should I use Firecrawl Scrape?

Firecrawl Scrape fits situations like: structured results once the URL; tool is selected.

How do I install Firecrawl Scrape in Claude Code?

Run `npx skills add firecrawl/skills --skill firecrawl-scrape -a claude-code`. Or copy the skill folder (skills/core/firecrawl-scrape in firecrawl/skills) into .claude/skills/firecrawl-scrape in your project. Claude Code loads it when a task matches its description.

How do I install Firecrawl Scrape in Codex?

Run `npx skills add firecrawl/skills --skill firecrawl-scrape -a codex`. Or copy the skill folder (skills/core/firecrawl-scrape in firecrawl/skills) into .agents/skills/firecrawl-scrape in your project. Codex loads it when a task matches its description.

Can I use Firecrawl Scrape in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add firecrawl/skills --skill firecrawl-scrape -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/firecrawl-scrape, .gemini/skills/firecrawl-scrape, .github/skills/firecrawl-scrape and .opencode/skills/firecrawl-scrape in your project.

What does Firecrawl Scrape need to run?

SKILL.md names no scripts, command-line tools or credentials: Firecrawl Scrape is instructions for the agent only. Our summary lists: Node.js. Its frontmatter pre-approves these tools: Bash(firecrawl *), Bash(npx firecrawl-cli *).

Does Firecrawl Scrape access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Firecrawl Scrape safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Firecrawl Scrape use?

Firecrawl Scrape is published under the ISC licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Firecrawl Scrape use?

About 1.8k tokens (SKILL.md is roughly 7.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 912 tokens, read only when the agent opens those files.

What are the alternatives to Firecrawl Scrape?

Skills that share tags, products or a category with Firecrawl Scrape: SEO Firecrawl (seranking/seo-skills, 160 stars), Keirouter Web Fetch (mydisha/keirouter, 147 stars), Firecrawl Scraper (davila7/claude-code-templates, 32k stars) and Firecrawl (zapier/connectors, 176 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Firecrawl Scrape?

firecrawl (a GitHub organization) maintains it in firecrawl/skills, which has 115 GitHub stars. The repository holds 29 skills in this directory. The repository was last updated on October 6, 2026.

Source: firecrawl/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.