Agent skill

Media Crawler Zhihu

by tsingyuai in tsingyuai/growth-lab

Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs.

Apache-2.0Auto-check passedData & Analytics

Install Media Crawler Zhihu

skills CLI
$ npx skills add tsingyuai/growth-lab --skill media-crawler-zhihu -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install tsingyuai/growth-lab media-crawler-zhihu --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/tsingyuai/growth-lab.git skills-src && mkdir -p .claude/skills && cp -r skills-src/collectors/media-crawler-zhihu .claude/skills/media-crawler-zhihu && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
media-crawler-zhihu
GitHub stars
2k
Token cost
~350 tokens
SKILL.md length
121 words
Files
2
Skills in repo
22
Repo updated
First seen
Licence
Apache-2.0

At a glance

Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs.

  • Expert discourse
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Problem framing
  • Creator research where content type and question context must remain explicit

What it does

Media Crawler Zhihu is an agent skill from tsingyuai/growth-lab. Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs. Use for expert discourse, problem framing, objections, terminology, topic, or creator research where content type and question context must remain explicit.

Its SKILL.md is about 350 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in Data & Analytics, covering Web scraping. The repository describes itself as: An end-to-end growth tool that understands the product, fetch the data it needs, researches the market, executes campaigns, and reviews results to improve the next round of… The licence is Apache-2.0.

When your agent uses it

  • Expert discourse
  • Problem framing
  • Creator research where content type and question context must remain explicit

Example prompts

  • “/media-crawler-zhihu”

What it can do on your machine

Read from SKILL.md and the folder at commit 2d0807c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Media Crawler Zhihu loads about 350 tokens when it runs. Until then it costs about 74 tokens; SKILL.md has 121 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~74
When it runs · the whole SKILL.md, loaded when a task matches
~350

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from tsingyuai/growth-lab at commit 2d0807c, republished under its Apache-2.0 licence (© tsingyuai). 121 words, ~350 tokens.

Download SKILL.mdSave it as .claude/skills/media-crawler-zhihu/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
media-crawler-zhihu
description
Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs. Use for expert discourse, problem framing, objections, terminology, topic, or creator research where content type and question context must remain explicit.

Zhihu collection

Follow media-crawler. Require purpose, keyword/content/creator inputs, sample and date bounds, comments/media scope, and invoking Model Memory destination.

Search with --platform zhihu --type search using problem wording, concept/category and audience/use-case variants. Shortlist by relevance, argument/content-type diversity, recency, author fit and engagement. Configure ZHIHU_SPECIFIED_ID_LIST with full answer, article, or video URLs; configure ZHIHU_CREATOR_URL_LIST with full people URLs. Run detail/creator mode.

Preserve content type. For answers retain question context; for articles retain article title; for video identify unavailable transcript rather than inventing one. Separate author text from comments, deduplicate by canonical content ID, and retain discovery-query provenance. Enable comments/media only for a shortlist and stop on challenge/risk control.

Return content-type breakdown, inputs, raw/copied paths, detail/comment/media coverage, time, commit, exclusions and selection rationale.

© tsingyuai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in collectors/media-crawler-zhihu of tsingyuai/growth-lab.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit 2d0807c

Compare with similar skills

Media Crawler Zhihu next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Media Crawler Zhihu compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Media Crawler Zhihu this skilltsingyuai/growth-lab2k—~350Automated safety check: PassApache-2.0
Tmuxtrpc-group/trpc-agent-go1.9k23 repos~868Automated safety check: PassApache-2.0
Ketch1broseidon/ketch7021 repos~3.9kAutomated safety check: PassMIT
Boss Zhipin Scrapereatmoreduck/boss-zhipin-scraper1.5k—~2.6kAutomated safety check: PassMIT
Crawl4AI Web Scrapingsmallnest/goclaw5991 repos~2.5kAutomated safety check: PassMIT
Axyusukebe/ax7191 repos~918Automated safety check: PassMIT

Similar skills

  • Tmux

    trpc-group/trpc-agent-go

    Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.

    1.9k GitHub starsUsed in 23 repos~868 tokens
    Data & AnalyticsAuto-check passed
  • Ketch

    1broseidon/ketch

    Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…

    702 GitHub starsUsed in 1 repo~3.9k tokens
    Data & AnalyticsAuto-check passed
  • Boss Zhipin Scraper

    eatmoreduck/boss-zhipin-scraper

    Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper.

    1.5k GitHub stars~2.6k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Crawl4AI Web Scraping

    smallnest/goclaw

    Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM.

    599 GitHub starsUsed in 1 repo~2.5k tokens
    Data & AnalyticsAuto-check passed
  • Ax

    yusukebe/ax

    Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML.

    719 GitHub starsUsed in 1 repo~918 tokens
    Data & AnalyticsAuto-check passed
  • Anakinscraper

    Anakin-Inc/anakin

    Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.

    4.5k GitHub stars~859 tokensUpdated 1 mo ago
    Data & AnalyticsAuto-check passed

More from tsingyuai/growth-lab

All 22 skills in this repo
  • Screenshot Assets

    tsingyuai/growth-lab

    Capture authenticated product screenshots through the repository-owned Playwright CDP script and archive them in the invoking loop's Memory.

    2k GitHub stars~583 tokensUpdated 13 days ago
    Auto-check passed
  • Media Crawler

    tsingyuai/growth-lab

    Install, authenticate, configure, operate, and troubleshoot the external MediaCrawler client shared by Douyin, Kuaishou, Bilibili, Weibo, Tieba, and Zhihu collectors.

    2k GitHub stars~731 tokensUpdated 13 days ago
    Auto-check passed
  • Run SEO Page Loop

    tsingyuai/growth-lab

    Run an SEO page observation-action-review loop with persistent Memory by coordinating demand research, page creation, adversarial review, image generation, IndexNow submission, and performance review.

    2k GitHub stars~1.2k tokensUpdated 13 days ago
    Auto-check passed
  • Wechat Article Compose

    tsingyuai/growth-lab

    把一个已确认的选题写成微信公众号长文。锁定复刻锚与用户价值、写出 article.md / wechat.yml / review.md,并通过公众号通用合规检查和人工预览前自检。准备、改写或核验公众号文章时使用;不负责发布,也不产出小红书卡片。

    2k GitHub stars~860 tokensUpdated 13 days ago
    Auto-check passed
  • Wechat Mp Publish

    tsingyuai/growth-lab

    通过微信公众号官方 API 把公众号文章生产单元渲染为微信排版 HTML、生成阅读页预览、上传封面与正文图片并创建草稿;在三重确认下提交发布并查询状态。支持本机直连与固定 IP 远程发布服务两种模式。用户要求预览公众号、同步微信草稿、发布公众号或查询发布状态时使用。

    2k GitHub stars~797 tokensUpdated 13 days ago
    Auto-check: notes
  • Xhs Render Cards

    tsingyuai/growth-lab

    把已批准的小红书草稿、单一分析参考和真实产品素材变成可审查卡片:先完成 DAI 与 image plan,再按确定性、完整效果或可分离图层模式制作,机械验证 PNG 与清单并运行合规检查。精确文字和真实 UI 不交给模型猜测;AI 生图需要单独配置和授权。

    2k GitHub stars~1k tokensUpdated 13 days ago
    Auto-check passed

Questions about Media Crawler Zhihu

What does Media Crawler Zhihu do?

Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs. Media Crawler Zhihu is an agent skill from tsingyuai/growth-lab. Collect Zhihu answers, articles, videos, comments, and creator evidence with MediaCrawler through search and exact URLs.

When should I use Media Crawler Zhihu?

Media Crawler Zhihu fits situations like: expert discourse; problem framing; creator research where content type and question context must remain explicit.

How do I install Media Crawler Zhihu in Claude Code?

Run `npx skills add tsingyuai/growth-lab --skill media-crawler-zhihu -a claude-code`. Or copy the skill folder (collectors/media-crawler-zhihu in tsingyuai/growth-lab) into .claude/skills/media-crawler-zhihu in your project. Claude Code loads it when a task matches its description.

How do I install Media Crawler Zhihu in Codex?

Run `npx skills add tsingyuai/growth-lab --skill media-crawler-zhihu -a codex`. Or copy the skill folder (collectors/media-crawler-zhihu in tsingyuai/growth-lab) into .agents/skills/media-crawler-zhihu in your project. Codex loads it when a task matches its description.

Can I use Media Crawler Zhihu in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add tsingyuai/growth-lab --skill media-crawler-zhihu -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/media-crawler-zhihu, .gemini/skills/media-crawler-zhihu, .github/skills/media-crawler-zhihu and .opencode/skills/media-crawler-zhihu in your project.

What does Media Crawler Zhihu need to run?

SKILL.md names no scripts, command-line tools or credentials: Media Crawler Zhihu is instructions for the agent only.

Does Media Crawler Zhihu access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Media Crawler Zhihu safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Media Crawler Zhihu use?

Media Crawler Zhihu is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Media Crawler Zhihu use?

About 350 tokens (SKILL.md is roughly 1.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Media Crawler Zhihu?

Skills that share tags, products or a category with Media Crawler Zhihu: Tmux (trpc-group/trpc-agent-go, 1.9k stars), Ketch (1broseidon/ketch, 702 stars), Boss Zhipin Scraper (eatmoreduck/boss-zhipin-scraper, 1.5k stars) and Crawl4AI Web Scraping (smallnest/goclaw, 599 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Media Crawler Zhihu?

tsingyuai (a GitHub organization) maintains it in tsingyuai/growth-lab, which has 2,000 GitHub stars. The repository holds 22 skills in this directory. The repository was last updated on September 28, 2026.

Source: tsingyuai/growth-lab on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.