Mineru
Nebutra/MinerU-Skill
An AI-Native skill for parsing PDF / Office / image files into clean Markdown with MinerU — a fast, zero-config document parser for AI agents.
PDF 数据提取工具。当用户提到"PDF 提取"、"PDF 转 Markdown"、"PDF 解析"、"提取 PDF 内容"、"PDF 转 JSON"、"RAG PDF"时使用。OpenDataLoader PDF 是目前基准测试第一的 PDF 解析器,支持本地模式(快速、确定)和混合 AI 模式(复杂表格、扫描件、公式),输出 Markdown、JSON(带边界框)、HTML。适用于需要从…
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install chujianyun/skills opendataloader-pdf --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/opendataloader-pdf .claude/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .claude/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdfType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install chujianyun/skills opendataloader-pdf --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/opendataloader-pdf .agents/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .agents/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install chujianyun/skills opendataloader-pdf --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/opendataloader-pdf .cursor/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .cursor/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/chujianyun/skills.git --path skills/opendataloader-pdf--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install chujianyun/skills opendataloader-pdf --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/opendataloader-pdf .gemini/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .gemini/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install chujianyun/skills opendataloader-pdfInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/opendataloader-pdf .github/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .github/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add chujianyun/skills --skill opendataloader-pdf -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install chujianyun/skills opendataloader-pdf --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/chujianyun/skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/opendataloader-pdf .opencode/skills/opendataloader-pdf && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "opendataloader-pdf" agent skill from https://github.com/chujianyun/skills/tree/main/skills/opendataloader-pdf into .opencode/skills/opendataloader-pdf/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "opendataloader-pdf", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
opendataloader-pdfPDF 数据提取工具。当用户提到"PDF 提取"、"PDF 转 Markdown"、"PDF 解析"、"提取 PDF 内容"、"PDF 转 JSON"、"RAG PDF"时使用。OpenDataLoader PDF 是目前基准测试第一的 PDF 解析器,支持本地模式(快速、确定)和混合 AI 模式(复杂表格、扫描件、公式),输出 Markdown、JSON(带边界框)、HTML。适用于需要从…
Opendataloader PDF is an agent skill from chujianyun/skills. PDF 数据提取工具。当用户提到"PDF 提取"、"PDF 转 Markdown"、"PDF 解析"、"提取 PDF 内容"、"PDF 转 JSON"、"RAG PDF"时使用。OpenDataLoader PDF 是目前基准测试第一的 PDF 解析器,支持本地模式(快速、确定)和混合 AI 模式(复杂表格、扫描件、公式),输出 Markdown、JSON(带边界框)、HTML。适用于需要从 PDF 提取结构化数据用于 RAG/LLM pipeline,或需要批量处理 PDF 文档的场景。
Its SKILL.md is about 830 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Documents & Office, covering PDF, Retrieval-augmented generation and Markdown. It works with Python and npm. The repository describes itself as: WuMing's Claude Skills.
Read from SKILL.md and the folder at commit 50f90c1. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pipnpmFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
github.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Opendataloader PDF loads about 827 tokens when it runs. Until then it costs about 67 tokens; SKILL.md has 174 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 174 words (~827 tokens).
“PDF 解析器 · 基准测试第一 · RAG/LLM 数据提取利器”
Just SKILL.md in skills/opendataloader-pdf of chujianyun/skills.
Open the folder on GitHubat commit 50f90c1
Opendataloader PDF next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Opendataloader PDF this skillchujianyun/skills | 740 | — | ~827 | Automated safety check: Pass | Custom licence | |
| MineruNebutra/MinerU-Skill | 122 | — | ~1.4k | Automated safety check: Pass | MIT | |
| Markdown Exporterbowenliang123/markdown-exporter | 272 | 1 repos | ~5.3k | Automated safety check: Pass | Apache-2.0 | |
| Markdown ConverterTeam-Commonly/commonly | 1.4k | — | ~557 | Automated safety check: Pass | Apache-2.0 | |
| Markdropshoryasethia/markdrop | 211 | — | ~1.4k | Automated safety check: Notes | GPL-3.0 | |
| Slidev CLI Skilldskilld-dev/vue-ecosystem-skills | 181 | — | ~1.1k | Automated safety check: Pass | MIT |
Nebutra/MinerU-Skill
An AI-Native skill for parsing PDF / Office / image files into clean Markdown with MinerU — a fast, zero-config document parser for AI agents.
bowenliang123/markdown-exporter
Convert Markdown text to DOCX, PPTX, XLSX, PDF, PNG, SVG, HTML, IPYNB, MD, CSV, JSON, JSONL, XML files, and extract code blocks in Markdown to Python, Bash,JS and etc files.
Team-Commonly/commonly
Convert binary documents (PDF, DOCX, XLSX, PPTX, HTML, EPUB, images) to clean LLM-friendly Markdown using Microsoft's markitdown Python tool.
shoryasethia/markdrop
Professional AI skill and usage instructions for the Markdrop package, a Python tool for converting PDFs to Markdown/HTML with AI-powered image/table descriptions.
skilld-dev/vue-ecosystem-skills
Build, present, and ship Slidev decks with @slidev/cli. An agent skill from skilld-dev/vue-ecosystem-skills.
cosmicstack-labs/mercury-agent-skills
Convert Markdown to publication-quality PDF with reportlab — CJK/Latin mixed text, themes, cover pages, watermarks, callouts, formulas, and interactive theme selection
chujianyun/skills
将公开或用户有权访问的 Wiki 完整转换为面向 Agent 的离线知识 Skill,并按原 Wiki 层级保存 Markdown、生成检索索引和逐文档哈希清单,支持无变化不落盘的手动或自动增量更新。当用户要求把 Wiki、文档站或帮助中心做成 Skill、同步已有 Wiki Skill、保持文档目录树或设置 Wiki Skill 自动更新时使用;不用于只摘要单篇网页或绕过登录、付费墙和访问控制。
chujianyun/skills
为照片、身份证、护照、学位证、毕业证、资格证、营业执照等证件或证书扫描件及 PDF 添加本地文字水印。用户提出照片加版权水印、身份证或学历证件添加“仅限某用途”水印、资质文件批量加水印、生成水印预览,或需要保持头像、二维码、印章等区域可辨认时使用。支持 JPEG、PNG、WebP 和 PDF,默认先预览、保留原件并清除图片元数据。不用于去除水印、伪造或篡改证件内容,也不用于仅设计 Logo…
chujianyun/skills
GitHub 源码解读助手。适用于用户提供 GitHub 仓库链接,并希望解读源码、理解原理、分析架构、生成学习报告或快速上手文档时使用。会在 working 目录下生成源码解读和快速上手两份文档。默认先交付初稿,不自动复查;如果用户明确同意,再安排后续复查。不适用于仅克隆仓库或只要一句简介的场景。
chujianyun/skills
LlamaIndex 官方用户文档离线知识库,用于检索并回答 LlamaIndex Python 框架的安装、RAG、数据加载、索引、检索与查询、Agent、Workflow、模型、Embedding、向量库、评估、可观测性、部署、LlamaCloud 和 LlamaParse 等问题,也可生成有文档依据的示例代码与排障建议。当用户提到…
chujianyun/skills
论文解读助手。适用于用户发送 arXiv 论文链接,并希望下载论文、解读论文、生成读书笔记、做论文拆解或输出详细报告时使用。会在工作目录创建论文文件夹、下载 PDF 与 TeX Source(如有)、生成中文 Markdown 报告。默认先交付初稿,不自动复查;如果用户明确同意,再安排后续复查。不适用于只要简短推荐语的情况。
chujianyun/skills
千问AI平台(Qianwen AI Platform / DashScope)官方文档离线知识库,用于检索并回答模型选择、API Key、OpenAI 兼容接口、DashScope SDK、文本与多模态生成、图像/视频/语音、Realtime API、Embedding、Reranking、Function Calling、MCP、批量调用、计费、Token Plan、API/SDK/CLI…
Categories
PDF 数据提取工具。当用户提到"PDF 提取"、"PDF 转 Markdown"、"PDF 解析"、"提取 PDF 内容"、"PDF 转 JSON"、"RAG PDF"时使用。OpenDataLoader PDF 是目前基准测试第一的 PDF 解析器,支持本地模式(快速、确定)和混合 AI 模式(复杂表格、扫描件、公式),输出 Markdown、JSON(带边界框)、HTML。适用于需要从…. Opendataloader PDF is an agent skill from chujianyun/skills.
Opendataloader PDF fits situations like: tasks that involve PDF; tasks that involve Retrieval-augmented generation; tasks that involve Markdown.
Run `npx skills add chujianyun/skills --skill opendataloader-pdf -a claude-code`. Or copy the skill folder (skills/opendataloader-pdf in chujianyun/skills) into .claude/skills/opendataloader-pdf in your project. Claude Code loads it when a task matches its description.
Run `npx skills add chujianyun/skills --skill opendataloader-pdf -a codex`. Or copy the skill folder (skills/opendataloader-pdf in chujianyun/skills) into .agents/skills/opendataloader-pdf in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add chujianyun/skills --skill opendataloader-pdf -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/opendataloader-pdf, .gemini/skills/opendataloader-pdf, .github/skills/opendataloader-pdf and .opencode/skills/opendataloader-pdf in your project.
Going by SKILL.md and its folder, Opendataloader PDF needs the command-line tools its instructions call (pip and npm). Our summary lists: Python 3; Node.js.
SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Opendataloader PDF has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.
About 827 tokens (SKILL.md is roughly 3.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Opendataloader PDF: Mineru (Nebutra/MinerU-Skill, 122 stars), Markdown Exporter (bowenliang123/markdown-exporter, 272 stars), Markdown Converter (Team-Commonly/commonly, 1.4k stars) and Markdrop (shoryasethia/markdrop, 211 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
chujianyun (a GitHub user) maintains it in chujianyun/skills, which has 740 GitHub stars. The repository holds 35 skills in this directory. The repository was last updated on October 9, 2026.
Source: chujianyun/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.