Agent skill

PDF Processing Openai

by lawve-ai in lawve-ai/awesome-legal-skills

Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control.

Apache-2.0Auto-check: notesDocuments & Office

Install PDF Processing Openai

skills CLI
$ npx skills add lawve-ai/awesome-legal-skills --skill pdf-processing-openai -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install lawve-ai/awesome-legal-skills pdf-processing-openai --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/lawve-ai/awesome-legal-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/pdf-editor-openai .claude/skills/pdf-processing-openai && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
pdf-processing-openai
GitHub stars
842
Token cost
~684 tokens
SKILL.md length
269 words
Files
3
Skills in repo
154
Repo updated
First seen
Licence
Apache-2.0

At a glance

Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control.

  • Works in 4 steps: Prefer visual review: render PDF pages… → Use reportlab to generate PDFs when… → Use pdfplumber (or pypdf) for text… → …
  • Work with PDFs (.pdf files) for:
  • SKILL.md covers When to use, Workflow, Temp and output conventions and Dependencies (install if…, plus 4 more sections
  • Calls uv, python3 and brew

What it does

PDF Processing Openai is an agent skill from lawve-ai/awesome-legal-skills. Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control. Use to work with PDFs (.pdf files) for: (1) Reading or extracting content from existing PDFs, (2) Creating new PDF documents with professional formatting, (3) Generating reports, documents, or layouts that require precise typography and design, or any other PDF reading or generation tasks.

Its SKILL.md is about 680 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `README.md`).

It sits in Documents & Office, covering PDF. It works with OpenAI and pypdf. The repository describes itself as: A curated list of awesome Agent Skills for automating legal work. The licence is Apache-2.0.

When your agent uses it

  • Work with PDFs (.pdf files) for:
  • Extracting content from existing PDFs
  • Creating new PDF documents with professional formatting
  • Generating reports

Example prompts

  • “/pdf-processing-openai”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Prefer visual review: render PDF pages to PNGs and inspect them.
  2. Use reportlab to generate PDFs when creating new documents.
  3. Use pdfplumber (or pypdf) for text extraction and quick checks; do not rely on it for layout fidelity.
  4. After each meaningful update, re-render pages and verify alignment, spacing, and legibility.

What it can do on your machine

Read from SKILL.md and the folder at commit 045f738. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • uv
    • python3
    • brew
    • apt-get
    • pdftoppm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

PDF Processing Openai loads about 684 tokens when it runs. Until then it costs about 101 tokens; SKILL.md has 269 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~101
When it runs · the whole SKILL.md, loaded when a task matches
~684

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteRuns commands with sudoSKILL.md:47
    sudo apt-get install -y poppler-utils

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from lawve-ai/awesome-legal-skills at commit 045f738, republished under its Apache-2.0 licence (© lawve-ai). 269 words, ~684 tokens.

Download SKILL.mdSave it as .claude/skills/pdf-processing-openai/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
pdf-processing-openai
description
Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control. Use to work with PDFs (.pdf files) for: (1) Reading or extracting content from existing PDFs, (2) Creating new PDF documents with professional formatting, (3) Generating reports, documents, or layouts that require precise typography and design, or any other PDF reading or generation tasks.
metadata.author
OpenAI
metadata.license
Apache-2.0
metadata.version
2026.01.30

PDF Skill

When to use

  • Read or review PDF content where layout and visuals matter.
  • Create PDFs programmatically with reliable formatting.
  • Validate final rendering before delivery.

Workflow

  1. Prefer visual review: render PDF pages to PNGs and inspect them.
    • Use pdftoppm if available.
    • If unavailable, install Poppler or ask the user to review the output locally.
  2. Use reportlab to generate PDFs when creating new documents.
  3. Use pdfplumber (or pypdf) for text extraction and quick checks; do not rely on it for layout fidelity.
  4. After each meaningful update, re-render pages and verify alignment, spacing, and legibility.

Temp and output conventions

  • Use tmp/pdfs/ for intermediate files; delete when done.
  • Write final artifacts under output/pdf/ when working in this repo.
  • Keep filenames stable and descriptive.

Dependencies (install if missing)

Prefer uv for dependency management.

Python packages:

uv pip install reportlab pdfplumber pypdf

If uv is unavailable:

python3 -m pip install reportlab pdfplumber pypdf

System tools (for rendering):

# macOS (Homebrew)
brew install poppler

# Ubuntu/Debian
sudo apt-get install -y poppler-utils

If installation isn't possible in this environment, tell the user which dependency is missing and how to install it locally.

Environment

No required environment variables.

Rendering command

pdftoppm -png $INPUT_PDF $OUTPUT_PREFIX

Quality expectations

  • Maintain polished visual design: consistent typography, spacing, margins, and section hierarchy.
  • Avoid rendering issues: clipped text, overlapping elements, broken tables, black squares, or unreadable glyphs.
  • Charts, tables, and images must be sharp, aligned, and clearly labeled.
  • Use ASCII hyphens only. Avoid U+2011 (non-breaking hyphen) and other Unicode dashes.
  • Citations and references must be human-readable; never leave tool tokens or placeholder strings.

Final checks

  • Do not deliver until the latest PNG inspection shows zero visual or formatting defects.
  • Confirm headers/footers, page numbering, and section transitions look polished.
  • Keep intermediate files organized or remove them after final approval.

© lawve-ai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files in skills/pdf-editor-openai of lawve-ai/awesome-legal-skills.

  • SKILL.md
  • LICENSE.txt
  • README.md

Open the folder on GitHubat commit 045f738

Compare with similar skills

PDF Processing Openai next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

PDF Processing Openai compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
PDF Processing Openai this skilllawve-ai/awesome-legal-skills842—~684Automated safety check: NotesApache-2.0
Openai PDFtrailofbits/skills-curated5139 repos~669Automated safety check: NotesCC-BY-SA-4.0
PDF Text Replaceinstavm/coderunner893—~536Automated safety check: PassApache-2.0
Lov Any2pdflovstudio/any2pdf211—~2.4kAutomated safety check: NotesMIT
PDF ToolkitXiaomiMiMo/MiMo-Code14k—~1.7kAutomated safety check: PassApache-2.0
PDF Generation, Forms and Extractionpipeshub-ai/pipeshub-ai3.8k—~2.9kAutomated safety check: PassApache-2.0

Similar skills

  • Openai PDF

    trailofbits/skills-curated

    Official

    A skill your agent uses when tasks involve reading, creating, or reviewing PDF files where rendering and layout matter; prefer visual checks by rendering pages (Poppler) and use Python tools such as…

    513 GitHub starsUsed in 9 repos~669 tokens
    Documents & OfficeAuto-check: notes
  • PDF Text Replace

    instavm/coderunner

    Replace text in fillable PDF forms by updating form field values.

    893 GitHub stars~536 tokensUpdated 3 days ago
    Documents & OfficeAuto-check passed
  • Lov Any2pdf

    lovstudio/any2pdf

    Convert Markdown documents to professionally typeset PDF files with reportlab.

    211 GitHub stars~2.4k tokensUpdated 2 mo ago
    Documents & OfficeAuto-check: notes
  • PDF Toolkit

    XiaomiMiMo/MiMo-Code

    Reads, transforms, composes and fills PDFs with Python scripts for extraction, merging, watermarking, encryption, OCR and form filling.

    14k GitHub stars~1.7k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Picks the right library for generating a new PDF, filling an existing PDF form, or extracting text and tables, defaulting to Node where possible.

    3.8k GitHub stars~2.9k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Markdown Paged Guide

    BigStrongSun/ccswitchmulti

    Render Markdown manuals into polished per-page PNG images and PDF deliverables.

    121 GitHub stars~589 tokensUpdated 5 days ago
    Documents & OfficeAuto-check passed

More from lawve-ai/awesome-legal-skills

All 154 skills in this repo
  • Customs Trade Law Onur Kafkas

    lawve-ai/awesome-legal-skills

    U.S. An agent skill from lawve-ai/awesome-legal-skills.

    842 GitHub stars~4.1k tokensUpdated 7 days ago
    Auto-check passed
  • Eu Data Act Oliver Schmidt Prietz

    lawve-ai/awesome-legal-skills

    Practitioner skill for advising on EU Regulation 2023/2854 (Data Act).

    842 GitHub stars~3.9k tokensUpdated 7 days ago
    Auto-check passed
  • Litigation Deadline Calendar

    lawve-ai/awesome-legal-skills

    Calendar litigation and arbitration deadlines from a scheduling order.

    842 GitHub stars~4.3k tokensUpdated 7 days ago
    Auto-check passed
  • Outlook Emails Lawvable

    lawve-ai/awesome-legal-skills

    Read, search, and download emails and attachments from Microsoft Outlook via OAuth2.

    842 GitHub stars~672 tokensUpdated 7 days ago
    Auto-check passed
  • Ambiguity Report

    lawve-ai/awesome-legal-skills

    Turn an interpretive-ambiguity audit of a legal text — contract, statute, regulation, or judicial opinion — into a polished deliverable.

    842 GitHub stars~4.6k tokensUpdated 7 days ago
    Auto-check passed
  • Az Eu Website Privacy Audit

    lawve-ai/awesome-legal-skills

    Audits a website for compliance with Azerbaijan's Law on Personal Data No.

    842 GitHub stars~3.8k tokensUpdated 7 days ago
    Auto-check passed

Works with

Questions about PDF Processing Openai

What does PDF Processing Openai do?

Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control. PDF Processing Openai is an agent skill from lawve-ai/awesome-legal-skills. Toolkit for comprehensive PDF reading, reviwing, and creation with visual quality control.

When should I use PDF Processing Openai?

PDF Processing Openai fits situations like: work with PDFs (.pdf files) for:; extracting content from existing PDFs; creating new PDF documents with professional formatting; generating reports.

How do I install PDF Processing Openai in Claude Code?

Run `npx skills add lawve-ai/awesome-legal-skills --skill pdf-processing-openai -a claude-code`. Or copy the skill folder (skills/pdf-editor-openai in lawve-ai/awesome-legal-skills) into .claude/skills/pdf-processing-openai in your project. Claude Code loads it when a task matches its description.

How do I install PDF Processing Openai in Codex?

Run `npx skills add lawve-ai/awesome-legal-skills --skill pdf-processing-openai -a codex`. Or copy the skill folder (skills/pdf-editor-openai in lawve-ai/awesome-legal-skills) into .agents/skills/pdf-processing-openai in your project. Codex loads it when a task matches its description.

Can I use PDF Processing Openai in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add lawve-ai/awesome-legal-skills --skill pdf-processing-openai -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pdf-processing-openai, .gemini/skills/pdf-processing-openai, .github/skills/pdf-processing-openai and .opencode/skills/pdf-processing-openai in your project.

What does PDF Processing Openai need to run?

Going by SKILL.md and its folder, PDF Processing Openai needs the command-line tools its instructions call (uv, python3, brew, apt-get and pdftoppm). Our summary lists: Python 3.

Does PDF Processing Openai access the network?

SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is PDF Processing Openai safe to install?

Our automated static check of SKILL.md found notes only (runs commands with sudo), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does PDF Processing Openai use?

PDF Processing Openai is published under the Apache-2.0 licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does PDF Processing Openai use?

About 684 tokens (SKILL.md is roughly 2.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to PDF Processing Openai?

Skills that share tags, products or a category with PDF Processing Openai: Openai PDF (trailofbits/skills-curated, 513 stars), PDF Text Replace (instavm/coderunner, 893 stars), Lov Any2pdf (lovstudio/any2pdf, 211 stars) and PDF Toolkit (XiaomiMiMo/MiMo-Code, 14k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains PDF Processing Openai?

lawve-ai (a GitHub organization) maintains it in lawve-ai/awesome-legal-skills, which has 842 GitHub stars. The repository holds 154 skills in this directory. The repository was last updated on October 2, 2026.

Source: lawve-ai/awesome-legal-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.