Agent skill

Firecrawl MCP

by LeoYeAI in LeoYeAI/openclaw-master-skills

Auto-generated skill for firecrawl-mcp tools via OneKey Gateway.

MITAuto-check passedData & Analytics

Install Firecrawl MCP

skills CLI
$ npx skills add LeoYeAI/openclaw-master-skills --skill firecrawl-mcp -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install LeoYeAI/openclaw-master-skills firecrawl-mcp --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/LeoYeAI/openclaw-master-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/firecrawl-mcp .claude/skills/firecrawl-mcp && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
firecrawl-mcp
GitHub stars
2.2k
Token cost
~7.4k tokens
SKILL.md length
2,696 words
Files
14 (incl. scripts)
Skills in repo
1,235
Repo updated
First seen
Licence
MIT

At a glance

Auto-generated skill for firecrawl-mcp tools via OneKey Gateway.

  • Works in 4 steps: Add waitFor parameter: Set waitFor: 5000… → Try a different URL: If the URL has a… → Use firecrawl_map to find the correct… → …
  • Tasks that involve Web scraping
  • SKILL.md covers Quick Start, Tools and CLI
  • Runs Python scripts from its folder; calls npx, pip and python3; reaches docs.firecrawl.dev and firecrawl.dev

What it does

Firecrawl MCP is an agent skill from LeoYeAI/openclaw-master-skills. Auto-generated skill for firecrawl-mcp tools via OneKey Gateway.

Its SKILL.md is about 7.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 14 other files, including scripts (for example `_meta.json`, `scripts/firecrawl_agent.py` and `scripts/firecrawl_agent_status.py`).

It sits in Data & Analytics, covering Web scraping and MCP servers. It works with Firecrawl and Model Context Protocol. The repository describes itself as: 🧠 Curated collection of 1209+ best OpenClaw skills — weekly updated by MyClaw.ai. The licence is MIT.

When your agent uses it

  • Tasks that involve Web scraping
  • Tasks that involve MCP servers

Example prompts

  • “/firecrawl-mcp”

Requirements

  • Python 3
  • A credential in YOUR_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Add waitFor parameter: Set waitFor: 5000 to waitFor: 10000 to allow JavaScript to render before extraction
  2. Try a different URL: If the URL has a hash fragment (#section), try the base URL or look for a direct page URL
  3. Use firecrawl_map to find the correct page: Large documentation sites or SPAs often spread content across multiple URLs. Use firecrawl_map…
  4. Use firecrawl_agent: As a last resort for heavily dynamic pages where map+scrape still fails, use the agent which can autonomously…

What it can do on your machine

Read from SKILL.md and the folder at commit e5199b5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 12 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • npx
    • pip
    • python3
    • npm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • docs.firecrawl.dev
    • firecrawl.dev

    Also links to:

    • deepnlp.org
    • producthunt.com
    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Firecrawl MCP loads about 7.4k tokens when it runs. Until then it costs about 20 tokens; SKILL.md has 2,696 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~20
When it runs · the whole SKILL.md, loaded when a task matches
~7.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from LeoYeAI/openclaw-master-skills at commit e5199b5, republished under its MIT licence (© LeoYeAI). 2,696 words, ~7,414 tokens.

Download SKILL.mdSave it as .claude/skills/firecrawl-mcp/SKILL.md (or your agent's skills folder). This skill also uses 13 other files; get the full folder from GitHub.
name
firecrawl-mcp
description
Auto-generated skill for firecrawl-mcp tools via OneKey Gateway.
dependencies.npm
@aiagenta2z/onekey-gateway
dependencies.python
ai-agent-marketplace
installation.npm
npm -g install @aiagenta2z/onekey-gateway
installation.python
pip install ai-agent-marketplace
OneKey Gateway

Use One Access Key to connect to various commercial APIs. Please visit the OneKey Gateway Keys and read the docs OneKey MCP Router Doc and OneKey Gateway Doc.

firecrawl-mcp Skill

Use the OneKey Gateway to access tools for this server via a unified access key.

Quick Start

Set your OneKey access key:

bash
export DEEPNLP_ONEKEY_ROUTER_ACCESS=YOUR_API_KEY

If no key is provided, the scripts fall back to the demo key BETA_TEST_KEY_MARCH_2026. Common settings:

  • unique_id: firecrawl-mcp/firecrawl-mcp
  • api_id: one of the tools listed below

Tools

firecrawl_scrape

Scrape content from a single URL with advanced options. This is the most powerful, fastest and most reliable scraper tool, if available you should always default to using this tool for any web scraping needs.

Best for: Single page content extraction, when you know exactly which page contains the information. Not recommended for: Multiple pages (call scrape multiple times or use crawl), unknown page location (use search). Common mistakes: Using markdown format when extracting specific data points (use JSON instead). Other Features: Use 'branding' format to extract brand identity (colors, fonts, typography, spacing, UI components) for design analysis or style replication.

CRITICAL - Format Selection (you MUST follow this): When the user asks for SPECIFIC data points, you MUST use JSON format with a schema. Only use markdown when the user needs the ENTIRE page content.

Use JSON format when user asks for:

  • Parameters, fields, or specifications (e.g., "get the header parameters", "what are the required fields")
  • Prices, numbers, or structured data (e.g., "extract the pricing", "get the product details")
  • API details, endpoints, or technical specs (e.g., "find the authentication endpoint")
  • Lists of items or properties (e.g., "list the features", "get all the options")
  • Any specific piece of information from a page

Use markdown format ONLY when:

  • User wants to read/summarize an entire article or blog post
  • User needs to see all content on a page without specific extraction
  • User explicitly asks for the full page content

Handling JavaScript-rendered pages (SPAs): If JSON extraction returns empty, minimal, or just navigation content, the page is likely JavaScript-rendered or the content is on a different URL. Try these steps IN ORDER:

  1. Add waitFor parameter: Set waitFor: 5000 to waitFor: 10000 to allow JavaScript to render before extraction
  2. Try a different URL: If the URL has a hash fragment (#section), try the base URL or look for a direct page URL
  3. Use firecrawl_map to find the correct page: Large documentation sites or SPAs often spread content across multiple URLs. Use firecrawl_map with a search parameter to discover the specific page containing your target content, then scrape that URL directly. Example: If scraping "https://docs.example.com/reference" fails to find webhook parameters, use firecrawl_map with {"url": "https://docs.example.com/reference", "search": "webhook"} to find URLs like "/reference/webhook-events", then scrape that specific page.
  4. Use firecrawl_agent: As a last resort for heavily dynamic pages where map+scrape still fails, use the agent which can autonomously navigate and research

Usage Example (JSON format - REQUIRED for specific data extraction):

json
{
  "name": "firecrawl_scrape",
  "arguments": {
    "url": "https://example.com/api-docs",
    "formats": [{
      "type": "json",
      "prompt": "Extract the header parameters for the authentication endpoint",
      "schema": {
        "type": "object",
        "properties": {
          "parameters": {
            "type": "array",
            "items": {
              "type": "object",
              "properties": {
                "name": { "type": "string" },
                "type": { "type": "string" },
                "required": { "type": "boolean" },
                "description": { "type": "string" }
              }
            }
          }
        }
      }
    }]
  }
}

Usage Example (markdown format - ONLY when full content genuinely needed):

json
{
  "name": "firecrawl_scrape",
  "arguments": {
    "url": "https://example.com/article",
    "formats": ["markdown"],
    "onlyMainContent": true
  }
}

Usage Example (branding format - extract brand identity):

json
{
  "name": "firecrawl_scrape",
  "arguments": {
    "url": "https://example.com",
    "formats": ["branding"]
  }
}

Branding format: Extracts comprehensive brand identity (colors, fonts, typography, spacing, logo, UI components) for design analysis or style replication. Performance: Add maxAge parameter for 500% faster scrapes using cached data. Returns: JSON structured data, markdown, branding profile, or other formats as specified.

Parameters:

  • url (string, required):
  • formats (array of object, optional):
  • parsers (array of object, optional):
  • onlyMainContent (boolean, optional):
  • includeTags (array of string, optional):
  • excludeTags (array of string, optional):
  • waitFor (number, optional):
  • actions (array of object, optional):
  • actions[].type (string, required): Values: wait, screenshot, scroll, scrape, click, write, press, executeJavascript, generatePDF
  • actions[].selector (string, optional):
  • actions[].milliseconds (number, optional):
  • actions[].text (string, optional):
  • actions[].key (string, optional):
  • actions[].direction (string, optional): Values: up, down
  • actions[].script (string, optional):
  • actions[].fullPage (boolean, optional):
  • mobile (boolean, optional):
  • skipTlsVerification (boolean, optional):
  • removeBase64Images (boolean, optional):
  • location (object, optional):
  • location.country (string, optional):
  • location.languages (array of string, optional):
  • storeInCache (boolean, optional):
  • zeroDataRetention (boolean, optional):
  • maxAge (number, optional):
  • proxy (string, optional): Values: basic, stealth, enhanced, auto
firecrawl_map

Map a website to discover all indexed URLs on the site.

Best for: Discovering URLs on a website before deciding what to scrape; finding specific sections or pages within a large site; locating the correct page when scrape returns empty or incomplete results. Not recommended for: When you already know which specific URL you need (use scrape); when you need the content of the pages (use scrape after mapping). Common mistakes: Using crawl to discover URLs instead of map; jumping straight to firecrawl_agent when scrape fails instead of using map first to find the right page.

IMPORTANT - Use map before agent: If firecrawl_scrape returns empty, minimal, or irrelevant content, use firecrawl_map with the search parameter to find the specific page URL containing your target content. This is faster and cheaper than using firecrawl_agent. Only use the agent as a last resort after map+scrape fails.

Prompt Example: "Find the webhook documentation page on this API docs site." Usage Example (discover all URLs):

json
{
  "name": "firecrawl_map",
  "arguments": {
    "url": "https://example.com"
  }
}

Usage Example (search for specific content - RECOMMENDED when scrape fails):

json
{
  "name": "firecrawl_map",
  "arguments": {
    "url": "https://docs.example.com/api",
    "search": "webhook events"
  }
}

Returns: Array of URLs found on the site, filtered by search query if provided.

Parameters:

  • url (string, required):
  • search (string, optional):
  • sitemap (string, optional): Values: include, skip, only
  • includeSubdomains (boolean, optional):
  • limit (number, optional):
  • ignoreQueryParameters (boolean, optional):

Search the web and optionally extract content from search results. This is the most powerful web search tool available, and if available you should always default to using this tool for any web search needs.

The query also supports search operators, that you can use if needed to refine the search:

OperatorFunctionalityExamples
""Non-fuzzy matches a string of text"Firecrawl"
-Excludes certain keywords or negates other operators-bad, -site:firecrawl.dev
site:Only returns results from a specified websitesite:firecrawl.dev
inurl:Only returns results that include a word in the URLinurl:firecrawl
allinurl:Only returns results that include multiple words in the URLallinurl:git firecrawl
intitle:Only returns results that include a word in the title of the pageintitle:Firecrawl
allintitle:Only returns results that include multiple words in the title of the pageallintitle:firecrawl playground
related:Only returns results that are related to a specific domainrelated:firecrawl.dev
imagesize:Only returns images with exact dimensionsimagesize:1920x1080
larger:Only returns images larger than specified dimensionslarger:1920x1080

Best for: Finding specific information across multiple websites, when you don't know which website has the information; when you need the most relevant content for a query. Not recommended for: When you need to search the filesystem. When you already know which website to scrape (use scrape); when you need comprehensive coverage of a single website (use map or crawl. Common mistakes: Using crawl or map for open-ended questions (use search instead). Prompt Example: "Find the latest research papers on AI published in 2023." Sources: web, images, news, default to web unless needed images or news. Scrape Options: Only use scrapeOptions when you think it is absolutely necessary. When you do so default to a lower limit to avoid timeouts, 5 or lower. Optimal Workflow: Search first using firecrawl_search without formats, then after fetching the results, use the scrape tool to get the content of the relevantpage(s) that you want to scrape

Usage Example without formats (Preferred):

json
{
  "name": "firecrawl_search",
  "arguments": {
    "query": "top AI companies",
    "limit": 5,
    "sources": [
      { "type": "web" }
    ]
  }
}

Usage Example with formats:

json
{
  "name": "firecrawl_search",
  "arguments": {
    "query": "latest AI research papers 2023",
    "limit": 5,
    "lang": "en",
    "country": "us",
    "sources": [
      { "type": "web" },
      { "type": "images" },
      { "type": "news" }
    ],
    "scrapeOptions": {
      "formats": ["markdown"],
      "onlyMainContent": true
    }
  }
}

Returns: Array of search results (with optional scraped content).

Parameters:

  • query (string, required):
  • limit (number, optional):
  • tbs (string, optional):
  • filter (string, optional):
  • location (string, optional):
  • sources (array of object, optional):
  • sources[].type (string, required): Values: web, images, news
  • scrapeOptions (object, optional):
  • scrapeOptions.formats (array of object, optional):
  • scrapeOptions.parsers (array of object, optional):
  • scrapeOptions.onlyMainContent (boolean, optional):
  • scrapeOptions.includeTags (array of string, optional):
  • scrapeOptions.excludeTags (array of string, optional):
  • scrapeOptions.waitFor (number, optional):
  • scrapeOptions.actions (array of object, optional):
  • scrapeOptions.actions[].type (string, required): Values: wait, screenshot, scroll, scrape, click, write, press, executeJavascript, generatePDF
  • scrapeOptions.actions[].selector (string, optional):
  • scrapeOptions.actions[].milliseconds (number, optional):
  • scrapeOptions.actions[].text (string, optional):
  • scrapeOptions.actions[].key (string, optional):
  • scrapeOptions.actions[].direction (string, optional): Values: up, down
  • scrapeOptions.actions[].script (string, optional):
  • scrapeOptions.actions[].fullPage (boolean, optional):
  • scrapeOptions.mobile (boolean, optional):
  • scrapeOptions.skipTlsVerification (boolean, optional):
  • scrapeOptions.removeBase64Images (boolean, optional):
  • scrapeOptions.location (object, optional):
  • scrapeOptions.location.country (string, optional):
  • scrapeOptions.location.languages (array of string, optional):
  • scrapeOptions.storeInCache (boolean, optional):
  • scrapeOptions.zeroDataRetention (boolean, optional):
  • scrapeOptions.maxAge (number, optional):
  • scrapeOptions.proxy (string, optional): Values: basic, stealth, enhanced, auto
  • enterprise (array of string, optional):
firecrawl_crawl

Starts a crawl job on a website and extracts content from all pages.

Best for: Extracting content from multiple related pages, when you need comprehensive coverage. Not recommended for: Extracting content from a single page (use scrape); when token limits are a concern (use map + batch_scrape); when you need fast results (crawling can be slow). Warning: Crawl responses can be very large and may exceed token limits. Limit the crawl depth and number of pages, or use map + batch_scrape for better control. Common mistakes: Setting limit or maxDiscoveryDepth too high (causes token overflow) or too low (causes missing pages); using crawl for a single page (use scrape instead). Using a /* wildcard is not recommended. Prompt Example: "Get all blog posts from the first two levels of example.com/blog." Usage Example:

json
{
  "name": "firecrawl_crawl",
  "arguments": {
    "url": "https://example.com/blog/*",
    "maxDiscoveryDepth": 5,
    "limit": 20,
    "allowExternalLinks": false,
    "deduplicateSimilarURLs": true,
    "sitemap": "include"
  }
}

Returns: Operation ID for status checking; use firecrawl_check_crawl_status to check progress.

Parameters:

  • url (string, required):
  • prompt (string, optional):
  • excludePaths (array of string, optional):
  • includePaths (array of string, optional):
  • maxDiscoveryDepth (number, optional):
  • sitemap (string, optional): Values: skip, include, only
  • limit (number, optional):
  • allowExternalLinks (boolean, optional):
  • allowSubdomains (boolean, optional):
  • crawlEntireDomain (boolean, optional):
  • delay (number, optional):
  • maxConcurrency (number, optional):
  • webhook (object, optional):
  • deduplicateSimilarURLs (boolean, optional):
  • ignoreQueryParameters (boolean, optional):
  • scrapeOptions (object, optional):
  • scrapeOptions.formats (array of object, optional):
  • scrapeOptions.parsers (array of object, optional):
  • scrapeOptions.onlyMainContent (boolean, optional):
  • scrapeOptions.includeTags (array of string, optional):
  • scrapeOptions.excludeTags (array of string, optional):
  • scrapeOptions.waitFor (number, optional):
  • scrapeOptions.actions (array of object, optional):
  • scrapeOptions.actions[].type (string, required): Values: wait, screenshot, scroll, scrape, click, write, press, executeJavascript, generatePDF
  • scrapeOptions.actions[].selector (string, optional):
  • scrapeOptions.actions[].milliseconds (number, optional):
  • scrapeOptions.actions[].text (string, optional):
  • scrapeOptions.actions[].key (string, optional):
  • scrapeOptions.actions[].direction (string, optional): Values: up, down
  • scrapeOptions.actions[].script (string, optional):
  • scrapeOptions.actions[].fullPage (boolean, optional):
  • scrapeOptions.mobile (boolean, optional):
  • scrapeOptions.skipTlsVerification (boolean, optional):
  • scrapeOptions.removeBase64Images (boolean, optional):
  • scrapeOptions.location (object, optional):
  • scrapeOptions.location.country (string, optional):
  • scrapeOptions.location.languages (array of string, optional):
  • scrapeOptions.storeInCache (boolean, optional):
  • scrapeOptions.zeroDataRetention (boolean, optional):
  • scrapeOptions.maxAge (number, optional):
  • scrapeOptions.proxy (string, optional): Values: basic, stealth, enhanced, auto
Show full SKILL.md (1,052 more words)Show less
firecrawl_check_crawl_status

Check the status of a crawl job.

Usage Example:

json
{
  "name": "firecrawl_check_crawl_status",
  "arguments": {
    "id": "550e8400-e29b-41d4-a716-446655440000"
  }
}

Returns: Status and progress of the crawl job, including results if available.

Parameters:

  • id (string, required):
firecrawl_extract

Extract structured information from web pages using LLM capabilities. Supports both cloud AI and self-hosted LLM extraction.

Best for: Extracting specific structured data like prices, names, details from web pages. Not recommended for: When you need the full content of a page (use scrape); when you're not looking for specific structured data. Arguments:

  • urls: Array of URLs to extract information from
  • prompt: Custom prompt for the LLM extraction
  • schema: JSON schema for structured data extraction
  • allowExternalLinks: Allow extraction from external links
  • enableWebSearch: Enable web search for additional context
  • includeSubdomains: Include subdomains in extraction Prompt Example: "Extract the product name, price, and description from these product pages." Usage Example:
json
{
  "name": "firecrawl_extract",
  "arguments": {
    "urls": ["https://example.com/page1", "https://example.com/page2"],
    "prompt": "Extract product information including name, price, and description",
    "schema": {
      "type": "object",
      "properties": {
        "name": { "type": "string" },
        "price": { "type": "number" },
        "description": { "type": "string" }
      },
      "required": ["name", "price"]
    },
    "allowExternalLinks": false,
    "enableWebSearch": false,
    "includeSubdomains": false
  }
}

Returns: Extracted structured data as defined by your schema.

Parameters:

  • urls (array of string, required):
  • prompt (string, optional):
  • schema (object, optional):
  • allowExternalLinks (boolean, optional):
  • enableWebSearch (boolean, optional):
  • includeSubdomains (boolean, optional):
firecrawl_agent

Autonomous web research agent. This is a separate AI agent layer that independently browses the internet, searches for information, navigates through pages, and extracts structured data based on your query. You describe what you need, and the agent figures out where to find it.

How it works: The agent performs web searches, follows links, reads pages, and gathers data autonomously. This runs asynchronously - it returns a job ID immediately, and you poll firecrawl_agent_status to check when complete and retrieve results.

IMPORTANT - Async workflow with patient polling:

  1. Call firecrawl_agent with your prompt/schema → returns job ID immediately
  2. Poll firecrawl_agent_status with the job ID to check progress
  3. Keep polling for at least 2-3 minutes - agent research typically takes 1-5 minutes for complex queries
  4. Poll every 15-30 seconds until status is "completed" or "failed"
  5. Do NOT give up after just a few polling attempts - the agent needs time to research

Expected wait times:

  • Simple queries with provided URLs: 30 seconds - 1 minute
  • Complex research across multiple sites: 2-5 minutes
  • Deep research tasks: 5+ minutes

Best for: Complex research tasks where you don't know the exact URLs; multi-source data gathering; finding information scattered across the web; extracting data from JavaScript-heavy SPAs that fail with regular scrape. Not recommended for: Simple single-page scraping where you know the URL (use scrape with JSON format instead - faster and cheaper).

Arguments:

  • prompt: Natural language description of the data you want (required, max 10,000 characters)
  • urls: Optional array of URLs to focus the agent on specific pages
  • schema: Optional JSON schema for structured output

Prompt Example: "Find the founders of Firecrawl and their backgrounds" Usage Example (start agent, then poll patiently for results):

json
{
  "name": "firecrawl_agent",
  "arguments": {
    "prompt": "Find the top 5 AI startups founded in 2024 and their funding amounts",
    "schema": {
      "type": "object",
      "properties": {
        "startups": {
          "type": "array",
          "items": {
            "type": "object",
            "properties": {
              "name": { "type": "string" },
              "funding": { "type": "string" },
              "founded": { "type": "string" }
            }
          }
        }
      }
    }
  }
}

Then poll with firecrawl_agent_status every 15-30 seconds for at least 2-3 minutes.

Usage Example (with URLs - agent focuses on specific pages):

json
{
  "name": "firecrawl_agent",
  "arguments": {
    "urls": ["https://docs.firecrawl.dev", "https://firecrawl.dev/pricing"],
    "prompt": "Compare the features and pricing information from these pages"
  }
}

Returns: Job ID for status checking. Use firecrawl_agent_status to poll for results.

Parameters:

  • prompt (string, required):
  • urls (array of string, optional):
  • schema (object, optional):
firecrawl_agent_status

Check the status of an agent job and retrieve results when complete. Use this to poll for results after starting an agent with firecrawl_agent.

IMPORTANT - Be patient with polling:

  • Poll every 15-30 seconds
  • Keep polling for at least 2-3 minutes before considering the request failed
  • Complex research can take 5+ minutes - do not give up early
  • Only stop polling when status is "completed" or "failed"

Usage Example:

json
{
  "name": "firecrawl_agent_status",
  "arguments": {
    "id": "550e8400-e29b-41d4-a716-446655440000"
  }
}

Possible statuses:

  • processing: Agent is still researching - keep polling, do not give up
  • completed: Research finished - response includes the extracted data
  • failed: An error occurred (only stop polling on this status)

Returns: Status, progress, and results (if completed) of the agent job.

Parameters:

  • id (string, required):
firecrawl_browser_create

Create a persistent browser session for code execution via CDP (Chrome DevTools Protocol).

Best for: Running code (Python/JS) that interacts with a live browser page, multi-step browser automation, persistent sessions that survive across multiple tool calls. Not recommended for: Simple page scraping (use firecrawl_scrape instead).

Arguments:

  • ttl: Total session lifetime in seconds (30-3600, optional)
  • activityTtl: Idle timeout in seconds (10-3600, optional)
  • streamWebView: Whether to enable live view streaming (optional)

Usage Example:

json
{
  "name": "firecrawl_browser_create",
  "arguments": {}
}

Returns: Session ID, CDP URL, and live view URL.

Parameters:

  • ttl (number, optional):
  • activityTtl (number, optional):
  • streamWebView (boolean, optional):
firecrawl_browser_execute

Execute code in a browser session. Supports agent-browser commands (bash), Python, or JavaScript.

Best for: Browser automation, navigating pages, clicking elements, extracting data, multi-step browser workflows. Requires: An active browser session (create one with firecrawl_browser_create first).

Arguments:

  • sessionId: The browser session ID (required)
  • code: The code to execute (required)
  • language: "bash", "python", or "node" (optional, defaults to "bash")

Recommended: Use bash with agent-browser commands (pre-installed in every sandbox):

json
{
  "name": "firecrawl_browser_execute",
  "arguments": {
    "sessionId": "session-id-here",
    "code": "agent-browser open https://example.com",
    "language": "bash"
  }
}

Common agent-browser commands:

  • agent-browser open <url> — Navigate to URL
  • agent-browser snapshot — Get accessibility tree with clickable refs (for AI)
  • agent-browser snapshot -i -c — Interactive elements only, compact
  • agent-browser click @e5 — Click element by ref from snapshot
  • agent-browser type @e3 "text" — Type into element
  • agent-browser fill @e3 "text" — Clear and fill element
  • agent-browser get text @e1 — Get text content
  • agent-browser get title — Get page title
  • agent-browser get url — Get current URL
  • agent-browser screenshot [path] — Take screenshot
  • agent-browser scroll down — Scroll page
  • agent-browser wait 2000 — Wait 2 seconds
  • agent-browser --help — Full command reference

For Playwright scripting, use Python (has proper async/await support):

json
{
  "name": "firecrawl_browser_execute",
  "arguments": {
    "sessionId": "session-id-here",
    "code": "await page.goto('https://example.com')\ntitle = await page.title()\nprint(title)",
    "language": "python"
  }
}

Note: Prefer bash (agent-browser) or Python. Returns: Execution result including stdout, stderr, and exit code.

Parameters:

  • sessionId (string, required):
  • code (string, required):
  • language (string, optional): Values: bash, python, node
firecrawl_browser_delete

Destroy a browser session.

Usage Example:

json
{
  "name": "firecrawl_browser_delete",
  "arguments": {
    "sessionId": "session-id-here"
  }
}

Returns: Success confirmation.

Parameters:

  • sessionId (string, required):
firecrawl_browser_list

List browser sessions, optionally filtered by status.

Usage Example:

json
{
  "name": "firecrawl_browser_list",
  "arguments": {
    "status": "active"
  }
}

Returns: Array of browser sessions.

Parameters:

  • status (string, optional): Values: active, destroyed

Usage

CLI

firecrawl_scrape
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_scrape '{"url": "https://example.com/api-docs", "formats": ["json"], "onlyMainContent": true}'
firecrawl_map
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_map '{"query": "webhook"}'
firecrawl_search
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_search '{"query": "product pricing site:example.com"}'
firecrawl_crawl
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_crawl '{"url": "https://example.com"}'
firecrawl_check_crawl_status
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_check_crawl_status '{"crawl_id": "crawl-123"}'
firecrawl_extract
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_extract '{"url": "https://example.com", "selector": "h1"}'
firecrawl_agent
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_agent '{"task": "find pricing", "url": "https://example.com"}'
firecrawl_agent_status
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_agent_status '{"job_id": "job-123"}'
firecrawl_browser_create
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_browser_create '{"profile": "default"}'
firecrawl_browser_execute
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_browser_execute '{"browser_id": "browser-1", "script": "document.title"}'
firecrawl_browser_delete
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_browser_delete '{"browser_id": "browser-1"}'
firecrawl_browser_list
shell
npx onekey agent firecrawl-mcp/firecrawl-mcp firecrawl_browser_list '{}'

Scripts

Each tool has a dedicated script in this folder:

  • skills/firecrawl-mcp/scripts/firecrawl_scrape.py
  • skills/firecrawl-mcp/scripts/firecrawl_map.py
  • skills/firecrawl-mcp/scripts/firecrawl_search.py
  • skills/firecrawl-mcp/scripts/firecrawl_crawl.py
  • skills/firecrawl-mcp/scripts/firecrawl_check_crawl_status.py
  • skills/firecrawl-mcp/scripts/firecrawl_extract.py
  • skills/firecrawl-mcp/scripts/firecrawl_agent.py
  • skills/firecrawl-mcp/scripts/firecrawl_agent_status.py
  • skills/firecrawl-mcp/scripts/firecrawl_browser_create.py
  • skills/firecrawl-mcp/scripts/firecrawl_browser_execute.py
  • skills/firecrawl-mcp/scripts/firecrawl_browser_delete.py
  • skills/firecrawl-mcp/scripts/firecrawl_browser_list.py
Example
bash
python3 scripts/<tool_name>.py --data '{"key": "value"}'

AI Agent Marketplace
Skills Marketplace AI Agent A2Z Deployment
PH AI Agent A2Z Infra
GitHub AI Agent Marketplace

Dependencies

CLI Dependency

Install onekey-gateway from npm

npm install @aiagenta2z/onekey-gateway
Script Dependency

Install the required Python package before running any scripts.

bash
pip install ai-agent-marketplace

Alternatively, install dependencies from the requirements file:

bash
pip install -r requirements.txt

If the package is already installed, skip installation.

Agent rule

Before executing command lines or running any script in the scripts/ directory, ensure the dependencies are installed. Use the onekey CLI as the preferred method to run the skills.

© LeoYeAI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 13 other files (scripts) in skills/firecrawl-mcp of LeoYeAI/openclaw-master-skills.

  • SKILL.md
  • _meta.json
  • scripts/firecrawl_agent.py
  • scripts/firecrawl_agent_status.py
  • scripts/firecrawl_browser_create.py
  • scripts/firecrawl_browser_delete.py
  • scripts/firecrawl_browser_execute.py
  • scripts/firecrawl_browser_list.py
  • scripts/firecrawl_check_crawl_status.py
  • scripts/firecrawl_crawl.py
  • scripts/firecrawl_extract.py
  • scripts/firecrawl_map.py
  • scripts/firecrawl_scrape.py
  • scripts/firecrawl_search.py

Open the folder on GitHubat commit e5199b5

Compare with similar skills

Firecrawl MCP next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Firecrawl MCP compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Firecrawl MCP this skillLeoYeAI/openclaw-master-skills2.2k—~7.4kAutomated safety check: PassMIT
Querying Indonesian Gov Datasuryast/indonesia-gov-apis172—~997Automated safety check: PassMIT
Mysearchskernelx/MySearch-Proxy159—~3kAutomated safety check: NotesNone
SEO Firecrawlseranking/seo-skills161—~2.3kAutomated safety check: PassMIT
Deep Researchaffaan-m/ECC276k2 repos~590Automated safety check: PassMIT
Scraplingforyourhealth111-pixel/Vibe-Skills3.6k—~1.1kAutomated safety check: PassApache-2.0

Similar skills

  • Querying Indonesian Gov Data

    suryast/indonesia-gov-apis

    Query 57 Indonesian government APIs and data sources — BPJPH halal certification, BPOM food safety, OJK financial legality, BPS statistics, BMKG weather/earthquakes, Bank Indonesia exchange rates…

    172 GitHub stars~997 tokensUpdated yesterday
    Data & AnalyticsAuto-check passed
  • Mysearch

    skernelx/MySearch-Proxy

    Install, verify, debug, and use MySearch MCP/Skill. An agent skill from skernelx/MySearch-Proxy.

    159 GitHub stars~3k tokensUpdated 6 mo ago
    Productivity & AutomationAuto-check: notes
  • SEO Firecrawl

    seranking/seo-skills

    Ad-hoc web scraping, site mapping, and full-site crawling via Firecrawl MCP.

    161 GitHub stars~2.3k tokensUpdated 3 mo ago
    Data & AnalyticsAuto-check passed
  • Deep Research

    affaan-m/ECC

    使用firecrawl和exa MCPs进行多源深度研究。搜索网络、综合发现并交付带有来源引用的报告。适用于用户希望对任何主题进行有证据和引用的彻底研究时。

    276k GitHub starsUsed in 2 repos~590 tokens
    Research & ScienceAuto-check passed
  • Scrapling

    foryourhealth111-pixel/Vibe-Skills

    CLI-first web scraping & content extraction with optional MCP server.

    3.6k GitHub stars~1.1k tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed
  • Firecrawl Scrape

    parcadei/Continuous-Claude-v3

    Scrape web pages and extract content via Firecrawl MCP. An agent skill from parcadei/Continuous-Claude-v3.

    3.9k GitHub starsUsed in 1 repo~254 tokens
    Data & AnalyticsAuto-check: notes

More from LeoYeAI/openclaw-master-skills

All 1,235 skills in this repo
  • DevOps Pipeline Management

    LeoYeAI/openclaw-master-skills

    Manages pipelines on a DevOps quality and efficiency platform through its OpenAPI: list workspaces and templates, create, update, run and cancel pipelines, and read run records.

    2.2k GitHub stars~4.2k tokensUpdated 2 mo ago
    Auto-check: notes
  • Feishu Document Collaboration

    LeoYeAI/openclaw-master-skills

    Patches OpenClaw's Feishu extension so an edited document triggers an isolated agent session that reads the doc and replies inline, turning it into a live chat space.

    2.2k GitHub stars~2k tokensUpdated 2 mo ago
    Auto-check passed
  • Files Memory System

    LeoYeAI/openclaw-master-skills

    Multi-context memory management system for OpenClaw agents with group-isolated storage, global shared memory, workspace organization, and group-specific skills isolation.

    2.2k GitHub stars~3.8k tokensUpdated 2 mo ago
    Auto-check passed
  • GEO-Claw AI Visibility Agent

    LeoYeAI/openclaw-master-skills

    Runs a brand's AI-search visibility work end to end: diagnosing how AI platforms represent it, repositioning it, producing AI-optimized content and monitoring ongoing mentions.

    2.2k GitHub stars~4.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Google Workspace CLI

    LeoYeAI/openclaw-master-skills

    Installs and authenticates the gws CLI, then automates Gmail, Drive, Sheets, Calendar, Docs, Chat and Tasks with ready-made recipes, persona bundles and security audits.

    2.2k GitHub stars~2.6k tokensUpdated 2 mo ago
    Auto-check: notes
  • HealthFit Health Advisors

    LeoYeAI/openclaw-master-skills

    Runs four advisor roles, a fitness coach, nutritionist, data analyst and TCM practitioner, to build a health profile and track workouts, diet and wellness over time.

    2.2k GitHub stars~4.4k tokensUpdated 2 mo ago
    Auto-check passed

Questions about Firecrawl MCP

What does Firecrawl MCP do?

Auto-generated skill for firecrawl-mcp tools via OneKey Gateway. Firecrawl MCP is an agent skill from LeoYeAI/openclaw-master-skills. Auto-generated skill for firecrawl-mcp tools via OneKey Gateway.

When should I use Firecrawl MCP?

Firecrawl MCP fits situations like: tasks that involve Web scraping; tasks that involve MCP servers.

How do I install Firecrawl MCP in Claude Code?

Run `npx skills add LeoYeAI/openclaw-master-skills --skill firecrawl-mcp -a claude-code`. Or copy the skill folder (skills/firecrawl-mcp in LeoYeAI/openclaw-master-skills) into .claude/skills/firecrawl-mcp in your project. Claude Code loads it when a task matches its description.

How do I install Firecrawl MCP in Codex?

Run `npx skills add LeoYeAI/openclaw-master-skills --skill firecrawl-mcp -a codex`. Or copy the skill folder (skills/firecrawl-mcp in LeoYeAI/openclaw-master-skills) into .agents/skills/firecrawl-mcp in your project. Codex loads it when a task matches its description.

Can I use Firecrawl MCP in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add LeoYeAI/openclaw-master-skills --skill firecrawl-mcp -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/firecrawl-mcp, .gemini/skills/firecrawl-mcp, .github/skills/firecrawl-mcp and .opencode/skills/firecrawl-mcp in your project.

What does Firecrawl MCP need to run?

Going by SKILL.md and its folder, Firecrawl MCP needs Python for the scripts in its folder and the command-line tools its instructions call (npx, pip, python3 and npm). Our summary lists: Python 3; A credential in YOUR_API_KEY.

Does Firecrawl MCP access the network?

SKILL.md names 5 domains. In commands or code: docs.firecrawl.dev and firecrawl.dev; the agent is likely to contact these when it follows the instructions. As links in the text: deepnlp.org, producthunt.com and github.com. This is read from the text; nothing was executed.

Is Firecrawl MCP safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Firecrawl MCP use?

Firecrawl MCP is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Firecrawl MCP use?

About 7.4k tokens (SKILL.md is roughly 30k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Firecrawl MCP?

Skills that share tags, products or a category with Firecrawl MCP: Querying Indonesian Gov Data (suryast/indonesia-gov-apis, 172 stars), Mysearch (skernelx/MySearch-Proxy, 159 stars), SEO Firecrawl (seranking/seo-skills, 161 stars) and Deep Research (affaan-m/ECC, 276k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Firecrawl MCP?

LeoYeAI (a GitHub user) maintains it in LeoYeAI/openclaw-master-skills, which has 2,161 GitHub stars. The repository holds 1,235 skills in this directory. The repository was last updated on July 20, 2026.

Source: LeoYeAI/openclaw-master-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.