PDF Processing
anthropics/skills
Handles everyday PDF jobs in Python and on the command line: extract text and tables, merge, split, rotate, watermark, fill forms, encrypt and OCR.
Deterministic PDF operations through bundled scripts: extract text and tables, merge files or page ranges, split by range, fill form fields and build PDFs from data.
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install TokenRhythm/opensquilla pdf-toolkit --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .claude/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .claude/skills/pdf-toolkit && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .claude/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkitType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install TokenRhythm/opensquilla pdf-toolkit --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .agents/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .agents/skills/pdf-toolkit && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .agents/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install TokenRhythm/opensquilla pdf-toolkit --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .cursor/skills/pdf-toolkit && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .cursor/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/TokenRhythm/opensquilla.git --path src/opensquilla/skills/bundled/pdf-toolkit--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install TokenRhythm/opensquilla pdf-toolkit --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .gemini/skills/pdf-toolkit && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .gemini/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install TokenRhythm/opensquilla pdf-toolkitInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .github/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .github/skills/pdf-toolkit && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .github/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install TokenRhythm/opensquilla pdf-toolkit --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/TokenRhythm/opensquilla.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/src/opensquilla/skills/bundled/pdf-toolkit .opencode/skills/pdf-toolkit && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "pdf-toolkit" agent skill from https://github.com/TokenRhythm/opensquilla/tree/main/src/opensquilla/skills/bundled/pdf-toolkit into .opencode/skills/pdf-toolkit/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "pdf-toolkit", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
pdf-toolkitDeterministic PDF operations through bundled scripts: extract text and tables, merge files or page ranges, split by range, fill form fields and build PDFs from data.
A set of structural PDF operations for jobs where you know exactly what you want done. The agent picks a bundled script by goal: `extract.py` for text and tables, `merge.py` to combine files or page ranges, `split.py` to cut a PDF by page ranges, `form_fill.py` to fill text form fields, and an inline reportlab snippet to build a new PDF from data. For a natural-language rewrite the agent drafts the new content first and then uses these operations to produce the file.
Extraction uses pdfplumber, which keeps column layout better than naive extraction, with `--tables-strategy` to switch table detection between lines, text and explicit modes, and `--json` for structured output. Merge accepts file names or a JSON manifest with 1-based page ranges per file, and split writes one numbered file per range. Scanned PDFs are out of scope because no OCR engine is included, and the skill points to a sibling OCR skill. Work stays in a restricted workspace, and the finished PDF is published as an artifact.
Read from SKILL.md and the folder at commit 4494195. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 4 files in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
pythonFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
PDF Toolkit loads about 1.9k tokens when it runs, and up to ~3.4k if it reads all its reference files. Until then it costs about 112 tokens; SKILL.md has 646 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from TokenRhythm/opensquilla at commit 4494195, republished under its Apache-2.0 licence (© TokenRhythm). 646 words, ~1,914 tokens.
.claude/skills/pdf-toolkit/SKILL.md (or your agent's skills folder). This skill also uses 7 other files; get the full folder from GitHub.Deterministic, structural PDF operations. Use this skill for programmatic work where you know exactly what you want done. For a natural-language rewrite, first draft the replacement content with ordinary reasoning, then use the explicit extract/generate/merge operations here to create the final PDF.
Use the inline Python examples with execute_code in a restricted channel.
Keep files in the active workspace and call publish_artifact with the
finished PDF. Shell commands below require an available exec_command;
they are optional shortcuts, not a reason to request host execution. If the
sandbox or a required library is unavailable, report the limitation rather
than retrying outside the sandbox.
| Goal | Script |
|---|---|
| Get text or tables out of a PDF | extract.py |
| Combine pages from multiple PDFs | merge.py |
| Split a PDF by page ranges | split.py |
Fill /Tx form fields in a PDF | form_fill.py |
| Build a new PDF from data | inline reportlab snippet, see Path C below |
python {baseDir}/scripts/extract.py /path/to/doc.pdf --jsonOutput:
{
"pages": 12,
"metadata": {"title": "...", "author": "..."},
"text": [
{"page": 1, "content": "..."},
{"page": 2, "content": "..."}
],
"tables": [
{"page": 3, "rows": [["..."], ["..."]]}
]
}Text uses pdfplumber (already in default dependencies) which preserves
column layout better than naive PDF text extraction. Tables use
pdfplumber.extract_tables() with default settings; for tricky layouts
pass --tables-strategy lines|text|explicit to switch detection mode.
For OCR (scanned PDFs), this skill does not include Tesseract — use the sibling skill that wraps an OCR engine (out of scope here).
Merge full files:
python {baseDir}/scripts/merge.py a.pdf b.pdf c.pdf --out combined.pdfOr merge specific page ranges with the manifest form:
python {baseDir}/scripts/merge.py manifest.json --out combined.pdfmanifest.json:
[
{"file": "a.pdf", "pages": "1-3"},
{"file": "b.pdf", "pages": "5,7,9-11"},
{"file": "c.pdf"}
]Page ranges are 1-based, comma-separated, hyphen for ranges. Omit pages to
include the whole file. Splits use the same syntax in reverse:
python {baseDir}/scripts/split.py input.pdf --pages "1-3,7,10-12" --out output_dir/Each range writes one output file: output_dir/input_001.pdf,
output_dir/input_002.pdf, …
python {baseDir}/scripts/form_fill.py form.pdf data.json --out filled.pdfdata.json maps field name → string value:
{
"applicant_name": "Wei E.",
"submission_date": "2026-05-06",
"agreed": "Yes"
}The script discovers fields via pypdf.PdfReader.get_fields() and updates
them with update_page_form_field_values(). Fields not present in the JSON
are left untouched. Run with --list-fields to enumerate the form's fields
without filling.
Caveats:
/Btn checkbox fields take the export value (often Yes, On, or 1)
rather than true — inspect with --list-fields to discover.--clear-signatures if that is intended.Use reportlab directly when you need a new PDF:
from reportlab.pdfgen import canvas
from reportlab.lib.pagesizes import LETTER
from pathlib import Path
c = canvas.Canvas(str(Path("out.pdf")), pagesize=LETTER)
c.setFont("Helvetica-Bold", 18)
c.drawString(72, 720, "Q3 Review")
c.setFont("Helvetica", 11)
c.drawString(72, 696, "Revenue grew 18% year over year.")
c.showPage()
c.save()For Chinese text, register a CJK font instead of Helvetica. ReportLab's CID font works without downloading fonts or reading a user font directory:
from reportlab.pdfbase import pdfmetrics
from reportlab.pdfbase.cidfonts import UnicodeCIDFont
from reportlab.pdfgen import canvas
pdfmetrics.registerFont(UnicodeCIDFont("STSong-Light"))
c = canvas.Canvas("report.pdf")
c.setFont("STSong-Light", 14)
c.drawString(72, 760, "季度报告:收入增长")
c.save()CID fonts rely on PDF reader CJK support. When an embedded font is required,
use a licensed font already available inside the workspace or permitted system
font roots. Validate the text with pypdf before publishing.
For tables, headers/footers, and multi-column layouts, switch to
reportlab.platypus (SimpleDocTemplate, Paragraph, Table,
PageBreak). See references/reportlab.md.
This public entry deliberately keeps the PDF mutation step deterministic. For requests such as "make the title shorter", inspect the source page, draft the replacement text with the model, and generate a new document with the reviewed content. Do not claim that arbitrary in-place page rewriting is available.
| Symptom | Cause | Fix |
|---|---|---|
| Extracted text is empty | Scanned PDF, no text layer | OCR is out of scope; use a separate OCR skill |
| Garbled characters in extract | PDF uses a custom font encoding | Try pdfplumber.open(path, laparams={...}) with char_margin adjustments |
| Merged PDF is huge | Underlying PDFs include large embedded fonts | Subset fonts via pypdf compress_content_streams() |
| Form fill silently no-ops | Field name in JSON does not match PDF field name | Run with --list-fields first to see exact names |
| Pages out of order after split | Range overlap collapsed unexpectedly | Use disjoint ranges, e.g. 1-3,4-6 not 1-5,3-6 |
© TokenRhythm, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 7 other files (scripts, references) in src/opensquilla/skills/bundled/pdf-toolkit of TokenRhythm/opensquilla.
Open the folder on GitHubat commit 4494195
PDF Toolkit next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| PDF Toolkit this skillTokenRhythm/opensquilla | 7.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | |
| PDF Processinganthropics/skills | 180k | 48 repos | ~2k | Automated safety check: Pass | Proprietary | |
| PDF Processing with PythonHKUDS/DeepTutor | 41k | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | |
| PDF ToolkitXiaomiMiMo/MiMo-Code | 14k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | |
| PDF Generation, Forms and Extractionpipeshub-ai/pipeshub-ai | 3.8k | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | |
| PDF Processing Guideagentscope-ai/QwenPaw | 35k | — | ~1.8k | Automated safety check: Pass | Proprietary |
anthropics/skills
Handles everyday PDF jobs in Python and on the command line: extract text and tables, merge, split, rotate, watermark, fill forms, encrypt and OCR.
HKUDS/DeepTutor
Reads, extracts from, creates, merges, splits, watermarks, encrypts and fills PDF files with pdfplumber, pypdf and reportlab inside a Python sandbox.
XiaomiMiMo/MiMo-Code
Reads, transforms, composes and fills PDFs with Python scripts for extraction, merging, watermarking, encryption, OCR and form filling.
pipeshub-ai/pipeshub-ai
Picks the right library for generating a new PDF, filling an existing PDF form, or extracting text and tables, defaulting to Node where possible.
agentscope-ai/QwenPaw
Handles PDF tasks with Python libraries and command-line tools: extract text and tables, merge, split, rotate, create, fill forms and more.
telagod/code-abyss
Picks the right Python library or CLI tool for a PDF task, text and table extraction, merging, splitting, OCR, watermarking or form filling, and points to a matching recipe.
TokenRhythm/opensquilla
Runs multi-round research in three stages with a persisted state file, evidence tracking and a long-form report with per-claim citations.
TokenRhythm/opensquilla
Inspects, edits in place or creates Word .docx files with bundled Python scripts, keeping existing styles intact when content changes.
TokenRhythm/opensquilla
Guides semantic, accessible HTML work: pages, forms, media and HTML5 APIs, plus how to deliver a runnable webpage project with a preview.
TokenRhythm/opensquilla
Reads, edits in place, or creates PowerPoint .pptx decks, picking one of three paths based on what tools and files are available.
TokenRhythm/opensquilla
Inspects, edits in place or creates Microsoft Excel .xlsx workbooks with openpyxl, treating each cell as a typed number, string, datetime or formula value.
TokenRhythm/opensquilla
Hands a self-contained coding task to Codex, Claude Code, OpenCode or Pi as a non-interactive background process, using OpenSquilla's exec_command and process tools.
Categories
Deterministic PDF operations through bundled scripts: extract text and tables, merge files or page ranges, split by range, fill form fields and build PDFs from data. A set of structural PDF operations for jobs where you know exactly what you want done.py` to fill text form fields, and an inline reportlab snippet to build a new PDF from data.
PDF Toolkit fits situations like: pulling tables or text out of a report PDF; combining several PDFs or selected page ranges into one file; splitting a PDF by page ranges; filling a form PDF's text fields from data.
Run `npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a claude-code`. Or copy the skill folder (src/opensquilla/skills/bundled/pdf-toolkit in TokenRhythm/opensquilla) into .claude/skills/pdf-toolkit in your project. Claude Code loads it when a task matches its description.
Run `npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a codex`. Or copy the skill folder (src/opensquilla/skills/bundled/pdf-toolkit in TokenRhythm/opensquilla) into .agents/skills/pdf-toolkit in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add TokenRhythm/opensquilla --skill pdf-toolkit -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pdf-toolkit, .gemini/skills/pdf-toolkit, .github/skills/pdf-toolkit and .opencode/skills/pdf-toolkit in your project.
Going by SKILL.md and its folder, PDF Toolkit needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python with pypdf, pdfplumber and reportlab.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
PDF Toolkit is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.5k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with PDF Toolkit: PDF Processing (anthropics/skills, 180k stars), PDF Processing with Python (HKUDS/DeepTutor, 41k stars), PDF Toolkit (XiaomiMiMo/MiMo-Code, 14k stars) and PDF Generation, Forms and Extraction (pipeshub-ai/pipeshub-ai, 3.8k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
TokenRhythm (a GitHub organization) maintains it in TokenRhythm/opensquilla, which has 7,087 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on October 4, 2026.
Source: TokenRhythm/opensquilla on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.