Agent skill

Yao Deepseek Crawler

by yaojingang in yaojingang/yao-geo-skills

A skill your agent uses when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami…

MITAuto-check passedData & Analytics

Install Yao Deepseek Crawler

skills CLI
$ npx skills add yaojingang/yao-geo-skills --skill yao-deepseek-crawler -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yaojingang/yao-geo-skills yao-deepseek-crawler --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yaojingang/yao-geo-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/yao-deepseek-crawler .claude/skills/yao-deepseek-crawler && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
yao-deepseek-crawler
GitHub stars
871
Token cost
~624 tokens
SKILL.md length
264 words
Files
120 (incl. scripts, references)
Skills in repo
24
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami…

  • Works in 7 steps: Read references/user-setup-and-usage.md… → Read… → Read references/report-contract.md for… → …
  • A user provides DeepSeek web AI-search keywords
  • SKILL.md covers Inputs, Workflow and Honest Boundaries
  • Calls node

What it does

Yao Deepseek Crawler is an agent skill from yaojingang/yao-geo-skills. Use when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report. Not for generic website crawling, DeepSeek API chat, SEO writing, or one-off answer generation.

Its SKILL.md is about 620 tokens, which your agent loads only when the skill is triggered. The skill folder holds 124 other files, including scripts and reference files (for example `README.en.md`, `README.md` and `agents/interface.yaml`).

It sits in Data & Analytics, covering Web scraping. It works with DeepSeek. The repository describes itself as: An open-source Skill collection for GEO content and workflows, continuously updated. The licence is MIT.

When your agent uses it

  • A user provides DeepSeek web AI-search keywords
  • Then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report

Example prompts

  • “/yao-deepseek-crawler”

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Read references/user-setup-and-usage.md for install, prerequisites, and user-facing steps.
  2. Read references/deepseek-crawl-workflow.md for crawler setup, preflight, delay, resume, and batch rules.
  3. Read references/report-contract.md for JSON schema, metrics, target/competitor recognition, and report rules.
  4. Run node scripts/preflight.mjs --profile before fresh crawling.
  5. Stage 1: run scripts/deepseek_batch_crawl.mjs with questions, repeat, profile, target entity/type, --safe-random-delay, and output dir.
  6. Stage 2: run scripts/analyze_deepseek_results.py on any crawl JSON with target entity/type, optional brands file, report output dir, and…
  7. Return the raw crawl JSON, structured Markdown, structured Excel workbook, HTML report, summary JSON, semantic-review cache when present…

What it can do on your machine

Read from SKILL.md and the folder at commit d21bfc1. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Yao Deepseek Crawler loads about 624 tokens when it runs, and up to ~9.2k if it reads all its reference files. Until then it costs about 78 tokens; SKILL.md has 264 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~624
With references · SKILL.md plus every file in references/, read only if the agent opens them
~9.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from yaojingang/yao-geo-skills at commit d21bfc1, republished under its MIT licence (© yaojingang). 264 words, ~624 tokens.

Download SKILL.mdSave it as .claude/skills/yao-deepseek-crawler/SKILL.md (or your agent's skills folder). This skill also uses 119 other files; get the full folder from GitHub.
name
yao-deepseek-crawler
description
Use when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report. Not for generic website crawling, DeepSeek API chat, SEO writing, or one-off answer generation.

Yao DeepSeek Crawler

Inputs

Standard inputs: keywords/questions, repeat count, target entity, entity type (人/person, 公司/company, 产品/product), browser profile, and optional output directory. Competitors must match the target type. Reports default to Simplified Chinese with an English summary toggle.

Workflow

  1. Read references/user-setup-and-usage.md for install, prerequisites, and user-facing steps.
  2. Read references/deepseek-crawl-workflow.md for crawler setup, preflight, delay, resume, and batch rules.
  3. Read references/report-contract.md for JSON schema, metrics, target/competitor recognition, and report rules.
  4. Run node scripts/preflight.mjs --profile <profile> before fresh crawling.
  5. Stage 1: run scripts/deepseek_batch_crawl.mjs with questions, repeat, profile, target entity/type, --safe-random-delay, and output dir.
  6. Stage 2: run scripts/analyze_deepseek_results.py on any crawl JSON with target entity/type, optional brands file, report output dir, and semantic review mode. Use --semantic-review auto by default; use --semantic-review required for formal delivery when AI review must pass.
  7. Return the raw crawl JSON, structured Markdown, structured Excel workbook, HTML report, summary JSON, semantic-review cache when present, and failed logs. Reports include AI semantic labels for entity recognition, target-vs-best-3 radar, click-to-reveal bubbles, Chinese source names, clickable citations, title intent, compact treemap, and GEO actions.

Honest Boundaries

  • Do not use for generic website crawling, DeepSeek API chat, SEO copywriting, or one-off answer generation.
  • Reuses local DeepSeek web automation; does not bypass login, CAPTCHA, bot checks, or hidden data.
  • Probability metrics are repeated-sample estimates, not ground truth.
  • Inferred competitors are heuristic unless --semantic-review required passes. AI semantic review is an audit enhancement and never replaces hard-rule gates or answer-body evidence.
  • Review aliases, semantic labels, excluded candidates, and competitor tables before external use.
  • Preserve raw answers, reference titles, URLs, and logs.

© yaojingang, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 119 other files (scripts, references) in skills/yao-deepseek-crawler of yaojingang/yao-geo-skills.

  • SKILL.md
  • .skillignore
  • README.en.md
  • README.md
  • agents/interface.yaml
  • evals/expected_artifacts.json
  • evals/trigger_cases.json
  • examples/nio-nev-deepseek-20260620/README.md
  • examples/nio-nev-deepseek-20260620/batch.log
  • examples/nio-nev-deepseek-20260620/brands.txt
  • examples/nio-nev-deepseek-20260620/deepseek-crawl.json
  • examples/nio-nev-deepseek-20260620/logs/q01-r01.log
  • examples/nio-nev-deepseek-20260620/logs/q01-r02.log
  • examples/nio-nev-deepseek-20260620/logs/q01-r03.log
  • examples/nio-nev-deepseek-20260620/logs/q01-r04.log
  • examples/nio-nev-deepseek-20260620/logs/q01-r05.log
  • … and 104 more

Open the folder on GitHubat commit d21bfc1

Compare with similar skills

Yao Deepseek Crawler next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Yao Deepseek Crawler compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Yao Deepseek Crawler this skillyaojingang/yao-geo-skills871—~624Automated safety check: PassMIT
Copy ResearchVKirill/claude-lane-stack122—~827Automated safety check: PassMIT
Tmuxtrpc-group/trpc-agent-go1.9k23 repos~868Automated safety check: PassApache-2.0
Ketch1broseidon/ketch7021 repos~3.9kAutomated safety check: PassMIT
Crawl4AI Web Scrapingsmallnest/goclaw5991 repos~2.5kAutomated safety check: PassMIT
Boss Zhipin Scrapereatmoreduck/boss-zhipin-scraper1.5k—~2.6kAutomated safety check: PassMIT

Similar skills

  • Copy Research

    VKirill/claude-lane-stack

    Dispatch copy-lead helpers: Tavily, Codex luna/terra, grok/X, OpenCode DeepSeek, Cursor Grok 4.6 medium-fast.

    122 GitHub stars~827 tokensUpdated yesterday
    Productivity & AutomationAuto-check passed
  • Tmux

    trpc-group/trpc-agent-go

    Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.

    1.9k GitHub starsUsed in 23 repos~868 tokens
    Data & AnalyticsAuto-check passed
  • Ketch

    1broseidon/ketch

    Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…

    702 GitHub starsUsed in 1 repo~3.9k tokens
    Data & AnalyticsAuto-check passed
  • Crawl4AI Web Scraping

    smallnest/goclaw

    Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM.

    599 GitHub starsUsed in 1 repo~2.5k tokens
    Data & AnalyticsAuto-check passed
  • Boss Zhipin Scraper

    eatmoreduck/boss-zhipin-scraper

    Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper.

    1.5k GitHub stars~2.6k tokensUpdated 11 days ago
    Data & AnalyticsAuto-check passed
  • Ax

    yusukebe/ax

    Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML.

    719 GitHub starsUsed in 1 repo~918 tokens
    Data & AnalyticsAuto-check passed

More from yaojingang/yao-geo-skills

All 24 skills in this repo
  • Yao Geo Knowledge Base Builder

    yaojingang/yao-geo-skills

    Build evidence-backed GEO brand knowledge bases from official sites, product pages, help centers, white papers, sales materials, media releases, certifications, and trusted third-party sources.

    871 GitHub stars~1.8k tokensUpdated 9 days ago
    Auto-check passed
  • Yao Geo Title Optimizer

    yaojingang/yao-geo-skills

    A skill your agent uses when Chinese content teams need GEO title candidates, scoring, compliance review, or title-to-article mapping for articles, pages, FAQs, comparisons, and topic hubs.

    871 GitHub stars~2k tokensUpdated 9 days ago
    Auto-check passed
  • Yao Geo Tracking

    yaojingang/yao-geo-skills

    Build a company-specific GEO backend tracking plan from a company name plus optional supporting information, using authoritative retrieval anchored on the official website, business-feature…

    871 GitHub stars~1.5k tokensUpdated 9 days ago
    Auto-check passed
  • Yao Geo Execution Roadmap

    yaojingang/yao-geo-skills

    当用户需要把 GEO 全景诊断、机会地图、AI 平台采样结论或品牌事实底座转成 30/60/90 天综合实施方案时使用,输出页面技术、内容矩阵、标题体系、知识库、外部证据、监测闭环、角色分工、验收指标和四格式报告;覆盖 DeepSeek、豆包、千问、Kimi、元宝;不用于从零做全景采样诊断或只写单篇内容。

    871 GitHub stars~646 tokensUpdated 9 days ago
    Auto-check passed
  • Yao Geo Intent Miner

    yaojingang/yao-geo-skills

    A skill your agent uses when a user asks for GEO 意图拓词、AI 搜索意图挖掘、AI 搜索问题集、问题簇、追问链路、查询重写、内容选题库、FAQ 题库、监测 Prompt 库, or AI Intent Miner.

    871 GitHub stars~500 tokensUpdated 9 days ago
    Auto-check passed
  • Yao Geo Page Blueprint

    yaojingang/yao-geo-skills

    A skill your agent uses when the user needs a GEO-friendly page blueprint for a specific product, topic, article, ranking, comparison, FAQ, knowledge-base, or case page, especially when the work…

    871 GitHub stars~789 tokensUpdated 9 days ago
    Auto-check passed

Works with

Questions about Yao Deepseek Crawler

What does Yao Deepseek Crawler do?

A skill your agent uses when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami…. Yao Deepseek Crawler is an agent skill from yaojingang/yao-geo-skills. Use when a user provides DeepSeek web AI-search keywords, repeat count, target entity, and entity type, then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report.

When should I use Yao Deepseek Crawler?

Yao Deepseek Crawler fits situations like: A user provides DeepSeek web AI-search keywords; then needs repeated fresh-window crawls aggregated into JSON plus a Kami HTML GEO report.

How do I install Yao Deepseek Crawler in Claude Code?

Run `npx skills add yaojingang/yao-geo-skills --skill yao-deepseek-crawler -a claude-code`. Or copy the skill folder (skills/yao-deepseek-crawler in yaojingang/yao-geo-skills) into .claude/skills/yao-deepseek-crawler in your project. Claude Code loads it when a task matches its description.

How do I install Yao Deepseek Crawler in Codex?

Run `npx skills add yaojingang/yao-geo-skills --skill yao-deepseek-crawler -a codex`. Or copy the skill folder (skills/yao-deepseek-crawler in yaojingang/yao-geo-skills) into .agents/skills/yao-deepseek-crawler in your project. Codex loads it when a task matches its description.

Can I use Yao Deepseek Crawler in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yaojingang/yao-geo-skills --skill yao-deepseek-crawler -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/yao-deepseek-crawler, .gemini/skills/yao-deepseek-crawler, .github/skills/yao-deepseek-crawler and .opencode/skills/yao-deepseek-crawler in your project.

What does Yao Deepseek Crawler need to run?

Going by SKILL.md and its folder, Yao Deepseek Crawler needs the command-line tools its instructions call (node).

Does Yao Deepseek Crawler access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Yao Deepseek Crawler safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Yao Deepseek Crawler use?

Yao Deepseek Crawler is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Yao Deepseek Crawler use?

About 624 tokens (SKILL.md is roughly 2.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 8.6k tokens, read only when the agent opens those files.

What are the alternatives to Yao Deepseek Crawler?

Skills that share tags, products or a category with Yao Deepseek Crawler: Copy Research (VKirill/claude-lane-stack, 122 stars), Tmux (trpc-group/trpc-agent-go, 1.9k stars), Ketch (1broseidon/ketch, 702 stars) and Crawl4AI Web Scraping (smallnest/goclaw, 599 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Yao Deepseek Crawler?

yaojingang (a GitHub user) maintains it in yaojingang/yao-geo-skills, which has 871 GitHub stars. The repository holds 24 skills in this directory. The repository was last updated on October 1, 2026.

Source: yaojingang/yao-geo-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.