Agent skill

Paddleocr Doc Parsing

by freestylefly in freestylefly/canghe-skills

Advanced document parsing with PaddleOCR. An agent skill from freestylefly/canghe-skills.

No licenceAuto-check: notesDocuments & Office

Install Paddleocr Doc Parsing

skills CLI
$ npx skills add freestylefly/canghe-skills --skill paddleocr-doc-parsing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install freestylefly/canghe-skills paddleocr-doc-parsing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/freestylefly/canghe-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/paddleocr-doc-parsing .claude/skills/paddleocr-doc-parsing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
paddleocr-doc-parsing
GitHub stars
461
Token cost
~2.8k tokens
SKILL.md length
1,151 words
Files
11 (incl. scripts, references)
Skills in repo
19
Repo updated
First seen
Licence
None found

At a glance

Advanced document parsing with PaddleOCR. An agent skill from freestylefly/canghe-skills.

  • Works in 5 steps: ONLY use PaddleOCR Document Parsing API… → NEVER parse documents directly - Do NOT… → NEVER offer alternatives - Do NOT… → …
  • Tasks that involve Document parsing
  • SKILL.md covers When to Use This Skill, How to Use This Skill, Important Notes and Reference Documentation, plus 1 more section
  • Runs Python scripts from its folder; calls python; reaches xxx.paddleocr.com and paddleocr.com; needs PADDLEOCR_ACCESS_TOKEN

What it does

Paddleocr Doc Parsing is an agent skill from freestylefly/canghe-skills. Advanced document parsing with PaddleOCR. Returns complete document structure including text, tables, formulas, charts, and layout information. The AI agent extracts relevant content based on user needs.

Its SKILL.md is about 2.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files, including scripts and reference files (for example `_meta.json`, `references/output_schema.md` and `scripts/configure.py`).

It sits in Documents & Office, covering Document parsing. The repository describes itself as: 苍何的技能skills仓库,搜集好用的 skills,辅助提效.

When your agent uses it

  • Tasks that involve Document parsing

Example prompts

  • “/paddleocr-doc-parsing”

Requirements

  • Python 3
  • A credential in PADDLEOCR_ACCESS_TOKEN

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. ONLY use PaddleOCR Document Parsing API - Execute the script python scripts/vl_caller.py
  2. NEVER parse documents directly - Do NOT parse documents yourself
  3. NEVER offer alternatives - Do NOT suggest "I can try to analyze it" or similar
  4. IF API fails - Display the error message and STOP immediately
  5. NO fallback methods - Do NOT attempt document parsing any other way

What it can do on your machine

Read from SKILL.md and the folder at commit dd0bf35. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 8 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • xxx.paddleocr.com
    • paddleocr.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • PADDLEOCR_ACCESS_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Paddleocr Doc Parsing loads about 2.8k tokens when it runs, and up to ~3.5k if it reads all its reference files. Until then it costs about 56 tokens; SKILL.md has 1,151 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~56
When it runs · the whole SKILL.md, loaded when a task matches
~2.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:236
    ot run `configure.py` or create a local `.env` file by default if the skill is installed under a host application direct

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 1,151 words (~2,798 tokens).

name
paddleocr-doc-parsing

Read the full SKILL.md on GitHub

Files

SKILL.md and 10 other files (scripts, references) in skills/paddleocr-doc-parsing of freestylefly/canghe-skills.

  • SKILL.md
  • _meta.json
  • references/output_schema.md
  • scripts/configure.py
  • scripts/lib.py
  • scripts/optimize_file.py
  • scripts/requirements-optimize.txt
  • scripts/requirements.txt
  • scripts/smoke_test.py
  • scripts/split_pdf.py
  • scripts/vl_caller.py

Open the folder on GitHubat commit dd0bf35

Compare with similar skills

Paddleocr Doc Parsing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Paddleocr Doc Parsing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Paddleocr Doc Parsing this skillfreestylefly/canghe-skills461—~2.8kAutomated safety check: NotesNone
MarkitdownImCa0/just-laws78214 repos~3.2kAutomated safety check: NotesMIT
DOCX ToolkitXiaomiMiMo/MiMo-Code14k—~2.4kAutomated safety check: PassApache-2.0
Huashu Markdown Publishing Pipelinealchaincyf/huashu-md-html907—~4.8kAutomated safety check: PassMIT
Markitdownjimmc414/Kosmos5952 repos~1.7kAutomated safety check: PassNone
Liteparsebastani-inc/atomic850—~1.4kAutomated safety check: PassMIT

Similar skills

  • Markitdown

    ImCa0/just-laws

    Convert files and office documents to Markdown. An agent skill from ImCa0/just-laws.

    782 GitHub starsUsed in 14 repos~3.2k tokens
    Documents & OfficeAuto-check: notes
  • DOCX Toolkit

    XiaomiMiMo/MiMo-Code

    Produces, edits and reads Microsoft Word files through python-docx and lxml, with a decision table for picking the lightest workflow for a given task.

    14k GitHub stars~2.4k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Huashu Markdown Publishing Pipeline

    alchaincyf/huashu-md-html

    Converts files and web pages into clean Markdown, then turns Markdown into polished HTML, Word, PDF and EPUB using four templates.

    907 GitHub stars~4.8k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed
  • Markitdown

    jimmc414/Kosmos

    Convert various file formats (PDF, Office documents, images, audio, web content, structured data) to Markdown optimized for LLM processing.

    595 GitHub starsUsed in 2 repos~1.7k tokens
    Documents & OfficeAuto-check passed
  • Liteparse

    bastani-inc/atomic

    A skill your agent uses whenever a task involves a document file (PDF, DOCX, PPTX, XLSX, or image) and you need to read it or pull text, tables, or specific values out of it — to answer a question…

    850 GitHub stars~1.4k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Beautiful Article

    ConardLi/garden-skills

    Turns a URL, PDF, DOCX, Markdown file, text or screenshots into a designed, shareable single-file HTML article through a staged review workflow.

    13k GitHub stars~4.7k tokensUpdated 2 mo ago
    Documents & OfficeAuto-check passed

More from freestylefly/canghe-skills

All 19 skills in this repo
  • Canghe Comic

    freestylefly/canghe-skills

    Knowledge comic creator supporting multiple art styles and tones.

    461 GitHub starsUsed in 8 repos~3.2k tokens
    Auto-check passed
  • Canghe Post To Wechat

    freestylefly/canghe-skills

    Posts content to WeChat Official Account (微信公众号) via API or Chrome CDP.

    461 GitHub starsUsed in 3 repos~3.4k tokens
    Auto-check: notes
  • Canghe Post To X

    freestylefly/canghe-skills

    Posts content and articles to X (Twitter). An agent skill from freestylefly/canghe-skills.

    461 GitHub starsUsed in 4 repos~1.7k tokens
    Auto-check: warnings
  • Canghe URL To Markdown

    freestylefly/canghe-skills

    Fetch any URL and convert to markdown using Chrome CDP. An agent skill from freestylefly/canghe-skills.

    461 GitHub starsUsed in 4 repos~1.1k tokens
    Auto-check passed
  • Canghe Xhs Images

    freestylefly/canghe-skills

    Generates Xiaohongshu (Little Red Book) infographic series with 10 visual styles and 8 layouts.

    461 GitHub starsUsed in 4 repos~4.9k tokens
    Auto-check passed
  • Canghe Markdown To HTML

    freestylefly/canghe-skills

    Converts Markdown to styled HTML with WeChat-compatible themes.

    461 GitHub starsUsed in 2 repos~1.6k tokens
    Auto-check passed

Questions about Paddleocr Doc Parsing

What does Paddleocr Doc Parsing do?

Advanced document parsing with PaddleOCR. An agent skill from freestylefly/canghe-skills. Paddleocr Doc Parsing is an agent skill from freestylefly/canghe-skills. Advanced document parsing with PaddleOCR.

When should I use Paddleocr Doc Parsing?

Paddleocr Doc Parsing fits situations like: tasks that involve Document parsing.

How do I install Paddleocr Doc Parsing in Claude Code?

Run `npx skills add freestylefly/canghe-skills --skill paddleocr-doc-parsing -a claude-code`. Or copy the skill folder (skills/paddleocr-doc-parsing in freestylefly/canghe-skills) into .claude/skills/paddleocr-doc-parsing in your project. Claude Code loads it when a task matches its description.

How do I install Paddleocr Doc Parsing in Codex?

Run `npx skills add freestylefly/canghe-skills --skill paddleocr-doc-parsing -a codex`. Or copy the skill folder (skills/paddleocr-doc-parsing in freestylefly/canghe-skills) into .agents/skills/paddleocr-doc-parsing in your project. Codex loads it when a task matches its description.

Can I use Paddleocr Doc Parsing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add freestylefly/canghe-skills --skill paddleocr-doc-parsing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/paddleocr-doc-parsing, .gemini/skills/paddleocr-doc-parsing, .github/skills/paddleocr-doc-parsing and .opencode/skills/paddleocr-doc-parsing in your project.

What does Paddleocr Doc Parsing need to run?

Going by SKILL.md and its folder, Paddleocr Doc Parsing needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named PADDLEOCR_ACCESS_TOKEN. Our summary lists: Python 3; A credential in PADDLEOCR_ACCESS_TOKEN.

Does Paddleocr Doc Parsing access the network?

SKILL.md names 2 domains. In commands or code: xxx.paddleocr.com and paddleocr.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Paddleocr Doc Parsing safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Paddleocr Doc Parsing use?

No licence was found for Paddleocr Doc Parsing or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Paddleocr Doc Parsing use?

About 2.8k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 671 tokens, read only when the agent opens those files.

What are the alternatives to Paddleocr Doc Parsing?

Skills that share tags, products or a category with Paddleocr Doc Parsing: Markitdown (ImCa0/just-laws, 782 stars), DOCX Toolkit (XiaomiMiMo/MiMo-Code, 14k stars), Huashu Markdown Publishing Pipeline (alchaincyf/huashu-md-html, 907 stars) and Markitdown (jimmc414/Kosmos, 595 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Paddleocr Doc Parsing?

freestylefly (a GitHub user) maintains it in freestylefly/canghe-skills, which has 461 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on June 8, 2026.

Source: freestylefly/canghe-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.