Read URLs and PDFs
tw93/Waza
Fetches web pages and PDFs and returns a source-grounded summary, clean Markdown, quotes or citations, routing each kind of link to a suitable fetch method.
Converts a URL or PDF into clean Markdown, with dedicated routes for WeChat articles, Feishu docs, arXiv papers and login-gated pages, before any summary or rewrite.
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxy --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
Claude Code skills documentation · loads skills from .claude/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .claude/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxy --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .agents/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxy --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .cursor/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxy --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .gemini/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxyInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .github/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install joeseesun/qiaomu-markdown-proxy qiaomu-markdown-proxy --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "qiaomu-markdown-proxy" agent skill from https://github.com/joeseesun/qiaomu-markdown-proxy/tree/main into .opencode/skills/qiaomu-markdown-proxy/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qiaomu-markdown-proxy", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
qiaomu-markdown-proxyConverts a URL or PDF into clean Markdown, with dedicated routes for WeChat articles, Feishu docs, arXiv papers and login-gated pages, before any summary or rewrite.
Whenever a user gives a link to read, this skill goes first and hands the Markdown to whatever comes next, such as a summary, translation, article or podcast script. It sorts each URL by type: WeChat article links use `scripts/fetch_weixin.sh`, which tries a proxy and falls back to Playwright on a verification page; Feishu and Lark docs use `scripts/fetch_feishu.py`, which needs Feishu API authentication; arXiv links use `scripts/extract_tex.py` to pull sections, figures and formulas from the LaTeX source; PDFs use `scripts/extract_pdf.sh`; everything else goes through `scripts/fetch.sh` with a cascade of proxy services.
It is not used when you already pasted the full text, for YouTube links (a separate download skill handles those) or for ordinary web search. After fetching, it shows the title, author and source platform. A request that only asks for extraction saves the page to `~/Downloads` as a Markdown file with YAML front matter (title, author, date, URL, source), unless you say preview only; a read-then-write request continues straight into the downstream task.
3 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit ab56e0a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/, which the agent can run.
Shell commands in SKILL.md call:
bashpython3curlFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
r.jina.aix.commp.weixin.qq.comxxx.feishu.cnarxiv.orgFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
FEISHU_APP_SECRETFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
URL to Markdown Fetcher loads about 1.4k tokens when it runs, and up to ~2.3k if it reads all its reference files. Until then it costs about 131 tokens; SKILL.md has 280 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from joeseesun/qiaomu-markdown-proxy at commit ab56e0a, republished under its MIT licence (© joeseesun). 280 words, ~1,363 tokens.
.claude/skills/qiaomu-markdown-proxy/SKILL.md (or your agent's skills folder). This skill also uses 25 other files; get the full folder from GitHub.将任意 URL 转为干净的 Markdown。支持需要登录的页面、PDF、专有平台。
mp.weixin.qq.com 和飞书文档属于专用路由。不要先试普通网页打开或通用内容解析器。qiaomu-youtube-download,普通搜索问题交给搜索工具。收到 URL 后,先判断类型,不同类型走不同通道:
| URL Pattern | Route To | Reason |
|---|---|---|
mp.weixin.qq.com | scripts/fetch_weixin.sh | 先代理,验证码页自动回退 Playwright |
feishu.cn/docx/ feishu.cn/wiki/ larksuite.com/docx/ | scripts/fetch_feishu.py | 需飞书 API 认证 |
youtube.com youtu.be | qiaomu-youtube-download skill | YouTube 有专用工具链 |
huggingface.co/papers/ | 提取 arXiv ID → scripts/extract_tex.py | HuggingFace 论文页实际是 arXiv 镜像,先找到 arXiv 链接再走 LaTeX 提取 |
arxiv.org/abs/ arxiv.org/pdf/ | scripts/extract_tex.py | 从 LaTeX 源码提取结构化内容 (章节/图表/公式) |
.pdf (URL or local path) | scripts/extract_pdf.sh | PDF 专用提取 |
| All other URLs | scripts/fetch.sh | 代理级联自动 fallback |
if URL contains "mp.weixin.qq.com":
→ bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch_weixin.sh "URL"
→ Done
if URL contains "feishu.cn/docx/" or "feishu.cn/wiki/" or "larksuite.com/docx/":
→ python3 ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch_feishu.py "URL"
→ Done
if URL contains "huggingface.co/papers/":
→ First fetch the page (WebFetch) to find the arXiv URL
→ Then python3 ~/.agents/skills/qiaomu-markdown-proxy/scripts/extract_tex.py "{arxiv_url}"
→ Done
if URL contains "arxiv.org/abs/" or "arxiv.org/pdf/":
→ python3 ~/.agents/skills/qiaomu-markdown-proxy/scripts/extract_tex.py "URL"
→ Done
if URL contains "youtube.com" or "youtu.be":
→ Call qiaomu-youtube-download skill
→ Done
if URL ends with ".pdf" or is local PDF path:
if remote URL:
→ Try: curl -sL "https://r.jina.ai/{url}"
→ If fails: download + extract_pdf.sh
if local path:
→ bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/extract_pdf.sh "PATH"
→ Done
else:
→ bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch.sh "URL"
→ DoneAfter fetching, show to user:
Title: {title}
Author: {author} (if available)
Source: {platform} (公众号 / 飞书文档 / 网页 / PDF)
URL: {original_url}
Summary
{3-5 sentence summary}
Content
{full Markdown, truncated at 200 lines if long}Composite request(“读取后写稿/总结/分析”):把提取结果直接交给下游任务,在同一轮继续;除非用户要求,不必额外保存源文件。
Extraction-only request(“只读取/转 Markdown”):保存到 ~/Downloads/{title}.md,使用 YAML frontmatter。
Filename: use article title, remove special characters.
Format: YAML frontmatter (title, author, date, url, source) + Markdown body.
Tell the user the saved path.
Skip saving if the user says “just preview” or “don't save”.
Only stop after extraction when extraction was the complete request. If the user asked for a downstream deliverable, continue to that deliverable.
bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch.sh "https://example.com/article"bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch.sh "https://x.com/username/status/1234567890"bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch_weixin.sh "https://mp.weixin.qq.com/s/abc123"python3 ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch_feishu.py "https://xxx.feishu.cn/docx/xxxxxxxx"python3 ~/.agents/skills/qiaomu-markdown-proxy/scripts/extract_tex.py "https://arxiv.org/abs/1706.03762"curl -sL "https://r.jina.ai/https://example.com/paper.pdf"bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/extract_pdf.sh "/path/to/paper.pdf"bash ~/.agents/skills/qiaomu-markdown-proxy/scripts/fetch.sh "https://example.com" "http://127.0.0.1:7890"fetch.sh handles proxy cascade with automatic fallbackuv is available, it runs them in an isolated environmentpython3 -m playwright install chromium if absentFEISHU_APP_ID + FEISHU_APP_SECRET env varsreferences/methods.md© joeseesun, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 25 other files (scripts, references, assets) in the repository root of joeseesun/qiaomu-markdown-proxy.
Open the folder on GitHubat commit ab56e0a
URL to Markdown Fetcher next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| URL to Markdown Fetcher this skilljoeseesun/qiaomu-markdown-proxy | 509 | — | ~1.4k | Automated safety check: Pass | MIT | |
| Read URLs and PDFstw93/Waza | 7.2k | — | ~1.8k | Automated safety check: Pass | MIT | |
| Web To Markdownrookie-ricardo/erduo-skills | 935 | — | ~894 | Automated safety check: Pass | MIT | |
| Clean Content FetchLeoYeAI/openclaw-master-skills | 2.2k | — | ~574 | Automated safety check: Pass | MIT | |
| Huashu Markdown Publishing Pipelinealchaincyf/huashu-md-html | 907 | — | ~4.8k | Automated safety check: Pass | MIT | |
| Feishu Docs Exporter and Writerleemysw/feishu-docx | 268 | — | ~1.8k | Automated safety check: Pass | MIT |
tw93/Waza
Fetches web pages and PDFs and returns a source-grounded summary, clean Markdown, quotes or citations, routing each kind of link to a suitable fetch method.
rookie-ricardo/erduo-skills
Convert a web URL into cleaned Markdown with deterministic routing.
LeoYeAI/openclaw-master-skills
获取干净、可读的网页正文内容,适合现代网页、博客、新闻、公告和微信公众号文章抓取;支持网页正文提取、内容清洗、去噪、Markdown 输出,适用于普通 fetch 效果不佳、页面噪音较多或动态渲染干扰的场景。Clean content fetch for modern web pages, article extraction, WeChat article capture, content…
alchaincyf/huashu-md-html
Converts files and web pages into clean Markdown, then turns Markdown into polished HTML, Word, PDF and EPUB using four templates.
leemysw/feishu-docx
Exports Feishu and Lark documents, sheets, bitables and wikis to Markdown, and creates, appends to or manages cloud docs from the command line.
joeseesun/qiaomu-epub-book-generator
Generate EPUB ebooks from Markdown files with SVG→PNG conversion, remote/local image embedding, compressed illustrations, and WeChat Reading compatibility.
Converts a URL or PDF into clean Markdown, with dedicated routes for WeChat articles, Feishu docs, arXiv papers and login-gated pages, before any summary or rewrite. Whenever a user gives a link to read, this skill goes first and hands the Markdown to whatever comes next, such as a summary, translation, article or podcast script.sh` with a cascade of proxy services.
URL to Markdown Fetcher fits situations like: reading a WeChat public-account article or a Feishu document from its link; converting a webpage or PDF into Markdown before summarizing or translating it; pulling the structured content of an arXiv paper from its LaTeX source; fetching a post from a site that requires login, such as X.
Run `npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a claude-code`. Or copy the skill folder (the joeseesun/qiaomu-markdown-proxy repository) into .claude/skills/qiaomu-markdown-proxy in your project. Claude Code loads it when a task matches its description.
Run `npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a codex`. Or copy the skill folder (the joeseesun/qiaomu-markdown-proxy repository) into .agents/skills/qiaomu-markdown-proxy in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add joeseesun/qiaomu-markdown-proxy --skill qiaomu-markdown-proxy -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/qiaomu-markdown-proxy, .gemini/skills/qiaomu-markdown-proxy, .github/skills/qiaomu-markdown-proxy and .opencode/skills/qiaomu-markdown-proxy in your project.
Going by SKILL.md and its folder, URL to Markdown Fetcher needs the command-line tools its instructions call (bash, python3 and curl) and credentials named FEISHU_APP_SECRET. Our summary lists: Bash, to run the fetch scripts; Feishu API credentials, for Feishu and Lark documents; Playwright, for the WeChat fallback.
SKILL.md names 5 domains. In commands or code: r.jina.ai, x.com, mp.weixin.qq.com, xxx.feishu.cn and arxiv.org; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
URL to Markdown Fetcher is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.4k tokens (SKILL.md is roughly 5.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 914 tokens, read only when the agent opens those files.
Skills that share tags, products or a category with URL to Markdown Fetcher: Read URLs and PDFs (tw93/Waza, 7.2k stars), Web To Markdown (rookie-ricardo/erduo-skills, 935 stars), Clean Content Fetch (LeoYeAI/openclaw-master-skills, 2.2k stars) and Huashu Markdown Publishing Pipeline (alchaincyf/huashu-md-html, 907 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
joeseesun (a GitHub user) maintains it in joeseesun/qiaomu-markdown-proxy, which has 509 GitHub stars. The repository was last updated on August 5, 2026.
Source: joeseesun/qiaomu-markdown-proxy on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.