Agent skill

Firecrawl Page Scrape Integration

by firecrawl in firecrawl/firecrawl

Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL.

ISCAuto-check passedData & Analytics

Install Firecrawl Page Scrape Integration

skills CLI
$ npx skills add firecrawl/firecrawl --skill firecrawl-build-scrape -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install firecrawl/firecrawl firecrawl-build-scrape --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/firecrawl/firecrawl.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/firecrawl-build-scrape .claude/skills/firecrawl-build-scrape && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
firecrawl-build-scrape
GitHub stars
190k
Used in
1 other repo
Token cost
~944 tokens
SKILL.md length
294 words
Files
2 (incl. references)
Skills in repo
5
Repo updated
First seen
Licence
ISC

At a glance

Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL.

  • Fetching the markdown of a known URL for a retrieval pipeline
  • SKILL.md covers Use This When, Default Recommendations, Freshness and Liveness and Common Product Patterns, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Extracting pricing or changelog content from a documentation page

What it does

This skill is for features that start with a known URL and need the content of that one page, for retrieval, summarization, enrichment or monitoring. Defaults are modest: return markdown unless the feature really needs another format, set `onlyMainContent` on article-like pages so navigation and page chrome do not add noise, and add waits or other rendering options only when the page needs them. Richer outputs such as links, screenshots or branding data are requested only when the consumer uses them.

A section on freshness explains that Firecrawl reuses recently indexed content, which makes repeat reads fast. The `maxAge` setting, in milliseconds, bounds how old a reused copy may be, `maxAge: 0` skips the reuse for freshness-critical reads, and `metadata.cacheState` and `metadata.cachedAt` show what was returned. A successful scrape only says what the page returned; whether the thing it describes is still active is a judgment your code makes. If you do not have the URL yet, the skill points to its search sibling, and for content that needs clicks or multi-step navigation, to its interact sibling. Docs pages are listed for Node and TypeScript, Python and Rust.

When your agent uses it

  • Fetching the markdown of a known URL for a retrieval pipeline
  • Extracting pricing or changelog content from a documentation page
  • Monitoring a single page for content changes

Example prompts

  • “Add a function that scrapes a given URL with Firecrawl and returns markdown for our indexer.”
  • “Fetch the main content of https://example.com/pricing without the navigation noise.”
  • “Make the scrape always read a fresh copy instead of a cached one.”

Requirements

  • Access to Firecrawl's `/scrape` endpoint

What it can do on your machine

Read from SKILL.md and the folder at commit e5df790. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • docs.firecrawl.dev

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Firecrawl Page Scrape Integration loads about 944 tokens when it runs, and up to ~1.5k if it reads all its reference files. Until then it costs about 73 tokens; SKILL.md has 294 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~73
When it runs · the whole SKILL.md, loaded when a task matches
~944
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from firecrawl/firecrawl at commit e5df790, republished under its ISC licence (© firecrawl). 294 words, ~944 tokens.

Download SKILL.mdSave it as .claude/skills/firecrawl-build-scrape/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
firecrawl-build-scrape
description
Integrate Firecrawl `/scrape` into product code for single-page extraction. Use when an app already has a URL and needs markdown, HTML, links, screenshots, metadata, or structured page output. Prefer this skill over broader crawl patterns when the feature is page-level.
license
ISC
metadata.author
firecrawl
metadata.version
0.1.0
metadata.homepage
https://www.firecrawl.dev
metadata.source
https://github.com/firecrawl/skills
references
references/freshness-and-liveness.md

Firecrawl Build Scrape

Use this when the application already has the URL and needs content from one page.

Use This When

  • the feature starts from a known URL
  • you need page content for retrieval, summarization, enrichment, or monitoring
  • you want the default extraction primitive before considering /interact

Default Recommendations

  • Return markdown unless the feature truly needs another format.
  • Use onlyMainContent for article-like pages where nav and chrome add noise.
  • Add waits or other rendering options only when the page needs them.

Freshness and Liveness

  • Firecrawl reuses recently indexed content, which is what makes repeat reads of the same URL fast. Set maxAge (milliseconds) to bound how old a reused copy may be, or maxAge: 0 to skip index reuse for a freshness-critical read.
  • Read metadata.cacheState and metadata.cachedAt to see what you actually got.
  • A successful scrape reports what the page returned. Whether the thing the page describes is still active is a source-specific judgment your code makes.
  • See references/freshness-and-liveness.md for the tradeoff, the metadata, and the decision rule.

Common Product Patterns

  • knowledge ingestion from known URLs
  • enrichment from a company, product, or docs page
  • pricing, changelog, and documentation extraction
  • page-level quality checks or monitoring

Escalation Rules

Implementation Notes

  • Keep the integration narrow: one feature, one URL, one extraction contract.
  • Treat /scrape as the default primitive for downstream LLM or indexing pipelines.
  • Request richer formats only when the consumer needs them, such as links, screenshots, or branding data.

Docs (Source of Truth)

Read the source-of-truth page for your project language before writing integration code:

See Also

© firecrawl, ISC. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in skills/firecrawl-build-scrape of firecrawl/firecrawl.

  • SKILL.md
  • references/freshness-and-liveness.md

Open the folder on GitHubat commit e5df790

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in firecrawl/firecrawl, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Firecrawl Page Scrape Integration next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Firecrawl Page Scrape Integration compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Firecrawl Page Scrape Integration this skillfirecrawl/firecrawl190k1 repos~944Automated safety check: PassISC
Oss Bounty Findertinyfish-io/tinyfish-cookbook2.2k—~4.6kAutomated safety check: PassMIT
Apify Actor Developmentapify/agent-skills2.4k—~2.9kAutomated safety check: PassNone
MCP Server Builderanthropics/skills180k63 repos~2.3kAutomated safety check: PassApache-2.0
MCP Server BuildershareAI-lab/learn-claude-code78k5 repos~1.2kAutomated safety check: PassMIT
Skyvern Browser AutomationSkyvern-AI/skyvern23k—~1.9kAutomated safety check: PassAGPL-3.0

Similar skills

  • Oss Bounty Finder

    tinyfish-io/tinyfish-cookbook

    Find paid open-source work, OSS bounties, open source grants, or ways to get paid contributing to open source.

    2.2k GitHub stars~4.6k tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • Apify Actor Development

    apify/agent-skills

    Official

    Creates, changes, debugs and deploys Apify Actors, including their input and output schemas, using the Apify CLI.

    2.4k GitHub stars~2.9k tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • MCP Server Builder

    anthropics/skills

    Official

    Guides the design and implementation of Model Context Protocol servers in TypeScript or Python, from tool naming and error messages to evaluation.

    180k GitHub starsUsed in 63 repos~2.3k tokens
    Agent WorkflowsAuto-check passed
  • MCP Server Builder

    shareAI-lab/learn-claude-code

    Walks through building MCP servers in Python or TypeScript that expose tools, resources and prompts to Claude, with templates, registration and testing.

    78k GitHub starsUsed in 5 repos~1.2k tokens
    Agent WorkflowsAuto-check passed
  • Skyvern Browser Automation

    Skyvern-AI/skyvern

    Automates websites with Skyvern's AI browser agent to fill forms, extract data, download files, log in and run multi-step workflows through SDKs, REST, MCP or a CLI.

    23k GitHub stars~1.9k tokensUpdated today
    Productivity & AutomationAuto-check passed
  • Tavily Search API Integration

    andrewyng/context-hub

    Guides building Tavily integrations for web search, URL extraction, site crawling and AI-assisted research in Python or JavaScript agent and RAG projects.

    14k GitHub stars~1.1k tokensUpdated 4 mo ago
    AI & LLM EngineeringAuto-check passed

More from firecrawl/firecrawl

  • Firecrawl Build Onboarding

    firecrawl/firecrawl

    Gets Firecrawl working in a project: signs you in through the browser, saves FIRECRAWL_API_KEY to .env and picks the first SDK or REST path.

    190k GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check: notes
  • Firecrawl App Integration

    firecrawl/firecrawl

    Adds web search, scraping, structured extraction and browser interaction to application code using Firecrawl's scrape, search and interact endpoints.

    190k GitHub stars~2.1k tokensUpdated today
    Auto-check: notes
  • Guides adding Firecrawl's /interact endpoint to product code for pages that need clicks, forms, pagination or logged-in flows beyond plain scraping.

    190k GitHub starsUsed in 1 repo~731 tokens
    Auto-check passed
  • Firecrawl Search Integration

    firecrawl/firecrawl

    Guidance for adding Firecrawl's /search endpoint to product code and agent workflows when a feature starts from a query rather than a URL.

    190k GitHub starsUsed in 1 repo~1.1k tokens
    Auto-check passed

Questions about Firecrawl Page Scrape Integration

What does Firecrawl Page Scrape Integration do?

Adds Firecrawl's /scrape endpoint to application code to pull markdown, HTML, links, screenshots or structured data from a single known URL. This skill is for features that start with a known URL and need the content of that one page, for retrieval, summarization, enrichment or monitoring. Defaults are modest: return markdown unless the feature really needs another format, set `onlyMainContent` on article-like pages so navigation and page chrome do not add noise, and add waits or other rendering options only when the page needs them.

When should I use Firecrawl Page Scrape Integration?

Firecrawl Page Scrape Integration fits situations like: fetching the markdown of a known URL for a retrieval pipeline; extracting pricing or changelog content from a documentation page; monitoring a single page for content changes.

How do I install Firecrawl Page Scrape Integration in Claude Code?

Run `npx skills add firecrawl/firecrawl --skill firecrawl-build-scrape -a claude-code`. Or copy the skill folder (skills/firecrawl-build-scrape in firecrawl/firecrawl) into .claude/skills/firecrawl-build-scrape in your project. Claude Code loads it when a task matches its description.

How do I install Firecrawl Page Scrape Integration in Codex?

Run `npx skills add firecrawl/firecrawl --skill firecrawl-build-scrape -a codex`. Or copy the skill folder (skills/firecrawl-build-scrape in firecrawl/firecrawl) into .agents/skills/firecrawl-build-scrape in your project. Codex loads it when a task matches its description.

Can I use Firecrawl Page Scrape Integration in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add firecrawl/firecrawl --skill firecrawl-build-scrape -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/firecrawl-build-scrape, .gemini/skills/firecrawl-build-scrape, .github/skills/firecrawl-build-scrape and .opencode/skills/firecrawl-build-scrape in your project.

What does Firecrawl Page Scrape Integration need to run?

SKILL.md names no scripts, command-line tools or credentials: Firecrawl Page Scrape Integration is instructions for the agent only. Our summary lists: Access to Firecrawl's `/scrape` endpoint.

Does Firecrawl Page Scrape Integration access the network?

SKILL.md names 1 domain. As links in the text: docs.firecrawl.dev. This is read from the text; nothing was executed.

Is Firecrawl Page Scrape Integration safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Firecrawl Page Scrape Integration use?

Firecrawl Page Scrape Integration is published under the ISC licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Firecrawl Page Scrape Integration use?

About 944 tokens (SKILL.md is roughly 3.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 587 tokens, read only when the agent opens those files.

What are the alternatives to Firecrawl Page Scrape Integration?

Skills that share tags, products or a category with Firecrawl Page Scrape Integration: Oss Bounty Finder (tinyfish-io/tinyfish-cookbook, 2.2k stars), Apify Actor Development (apify/agent-skills, 2.4k stars), MCP Server Builder (anthropics/skills, 180k stars) and MCP Server Builder (shareAI-lab/learn-claude-code, 78k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Firecrawl Page Scrape Integration?

firecrawl (a GitHub organization) maintains it in firecrawl/firecrawl, which has 189,703 GitHub stars. The repository holds 5 skills in this directory. The repository was last updated on October 9, 2026.

Source: firecrawl/firecrawl on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.