Agent skill

PDF Organizer

by cat-xierluo in cat-xierluo/legal-skills

当需要整理法律 PDF 时使用:检测文字层,生成页面索引、整理草稿和下游交接文件,按内容拆分、合并或直接重命名 OCR 后双层扫描件并规范命名;支持跨源页引用拼合与页覆盖审计(孤儿页/重复引用);可做旋转与倾斜校正,不做 OCR 或压缩。

MITAuto-check passedDocuments & Office

Install PDF Organizer

skills CLI
$ npx skills add cat-xierluo/legal-skills --skill pdf-organizer -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install cat-xierluo/legal-skills pdf-organizer --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/cat-xierluo/legal-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/pdf-organizer .claude/skills/pdf-organizer && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
pdf-organizer
GitHub stars
720
Token cost
~2.3k tokens
SKILL.md length
452 words
Files
11 (incl. scripts, references)
Skills in repo
62
Repo updated
First seen
Licence
MIT

At a glance

当需要整理法律 PDF 时使用:检测文字层,生成页面索引、整理草稿和下游交接文件,按内容拆分、合并或直接重命名 OCR 后双层扫描件并规范命名;支持跨源页引用拼合与页覆盖审计(孤儿页/重复引用);可做旋转与倾斜校正,不做 OCR 或压缩。

  • Works in 3 steps: 清点与预处理判断 → 判断拆分还是合并 → 命名规则
  • Tasks that involve PDF
  • SKILL.md covers 定位, 与其他技能配合, 依赖 and 输入/输出, plus 5 more sections
  • Runs Python scripts from its folder; calls python3 and brew

What it does

PDF Organizer is an agent skill from cat-xierluo/legal-skills. 当需要整理法律 PDF 时使用:检测文字层,生成页面索引、整理草稿和下游交接文件,按内容拆分、合并或直接重命名 OCR 后双层扫描件并规范命名;支持跨源页引用拼合与页覆盖审计(孤儿页/重复引用);可做旋转与倾斜校正,不做 OCR 或压缩。

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 13 other files, including scripts and reference files (for example `CHANGELOG.md`, `references/manifest-contract.md` and `references/organize-manifest-schema.md`).

It sits in Documents & Office, covering PDF. It works with Python. The licence is MIT.

When your agent uses it

  • Tasks that involve PDF

Example prompts

  • “/pdf-organizer”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. 清点与预处理判断
  2. 判断拆分还是合并
  3. 命名规则

What it can do on your machine

Read from SKILL.md and the folder at commit db2c58c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 4 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3
    • brew

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

PDF Organizer loads about 2.3k tokens when it runs, and up to ~6.3k if it reads all its reference files. Until then it costs about 33 tokens; SKILL.md has 452 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~33
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~6.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from cat-xierluo/legal-skills at commit db2c58c, republished under its MIT licence (© cat-xierluo). 452 words, ~2,283 tokens.

Download SKILL.mdSave it as .claude/skills/pdf-organizer/SKILL.md (or your agent's skills folder). This skill also uses 10 other files; get the full folder from GitHub.
name
pdf-organizer
description
当需要整理法律 PDF 时使用:检测文字层,生成页面索引、整理草稿和下游交接文件,按内容拆分、合并或直接重命名 OCR 后双层扫描件并规范命名;支持跨源页引用拼合与页覆盖审计(孤儿页/重复引用);可做旋转与倾斜校正,不做 OCR 或压缩。
homepage
https://github.com/cat-xierluo/legal-skills
author
杨卫薪律师(微信ywxlaw)
version
0.6.1
license
MIT

PDF Organizer

定位

本技能处理“PDF 文件和法律文书逻辑不一致”的问题。它关注的不是普通文件拼接,而是按文书内容把 PDF 整理成律师实际会使用的文件。

核心职责:

  1. 按内容把整份扫描/OCR PDF 拆成多份独立文书。
  2. 按内容把被拆散的同一份文书合并回一个 PDF。
  3. 对不需要拆分或合并的 PDF,直接根据内容识别结果规范命名。
  4. 根据标题、主体、相对方、案由、日期等要素生成规范文件名。
  5. 生成页面级检查索引和 manifest 草稿,供 AI/人工复核。
  6. 以页引用集合统一表达拆分、复制、合并与跨源拼合,执行前可审计页覆盖(孤儿页/重复引用页)。
  7. 生成下游 handoff.json,附页级来源引用,供案件材料整理、合同审查、诉讼分析等 Skill 协同使用。
  8. 在整理前后做轻量页面方向处理,例如 90/180/270 度旋转;倾斜校正作为可选预处理能力。
  9. 将 PDF 每页标准化为 A4 尺寸:横向页面自动适配 A4 横版(842×595 pt),竖向页面自动适配 A4 竖版(595×842 pt),等比缩放居中。

本技能不做 OCR、不压缩 PDF,也不替代案件材料整体归类。若 PDF 还没有可检索文字层,先使用 OCR 工具;若需要复杂裁边、压缩、批量 OCR 或高质量图像预处理,优先使用 PDF Processor。

与其他技能配合

  • OCR 前处理、双层 PDF、压缩、复杂裁边等工作交给 PDF Processor。
  • 复杂版面识别、PDF 转 Markdown 或需要结构化 OCR 结果时,先使用 OCR 解析工具。
  • 本技能输出的整理后 PDF 可以继续交给材料整理、案件分析或文书生成技能。

依赖

系统依赖
依赖安装方式
python3macOS 通常已内置
ocrmypdf可选,仅 deskew 倾斜校正需要;macOS: brew install ocrmypdf
Python 包
包名用途安装命令
pypdf拆分、合并、复制、页面旋转python3 -m pip install -r scripts/requirements.txt

只做清单规划时不需要安装依赖;执行拆分、合并或旋转时需要 pypdf。倾斜校正依赖 ocrmypdf 命令行工具,未安装时应提示用户改用 PDF Processor 或先安装。

输入/输出

输入
  • 整份 OCR 后 PDF:例如一份 16 页的扫描合并件。
  • 已拆散的多个 PDF:例如同一份起诉状被拆成 第 1-2 页.pdf、第 3-4 页.pdf。
  • 独立 PDF 文件:整份 PDF 已经是一份完整文书,不需要拆分或合并,只需要根据内容规范命名。
  • 可选 OCR Markdown 或文字片段:用于辅助判断标题、边界和合并关系。
  • 可选命名偏好:是否保留日期、是否写案由、是否使用简称。
输出
  • 整理后的 PDF 文件。交付目录默认只放最终 PDF。
  • 过程文件统一归档到 archive/YYYYMMDD_HHMMSS_{来源名}/。

archive 默认包含:

  • organize_manifest.input.json:本次输入清单副本。
  • organize_manifest.resolved.json:实际输出路径、页码、状态和命名依据。
  • organize_report.md:给人工复核的摘要报告。
  • handoff.json:面向下游 Skill 的交接文件。
  • run_meta.json:运行元数据。

工作流程

1. 清点与预处理判断
  1. 确认输入是整份 PDF、多个碎片 PDF、独立 PDF 文件(只需重命名),还是以上几类的混合。
  2. 读取页数,建立页码与文件的对应关系。
  3. 先做文字层检测,确认 PDF 是 OCR 后双层 PDF;没有可检索文字层时,停止内容拆分/合并判断,提示用户先走 PDF Processor 生成双层 PDF。
  4. 运行页面检查,提取每页文字量、疑似标题、页码、日期、主体候选和边界信号。
  5. 可先生成 organize_manifest.json 草稿,再由 AI 根据页面证据和命名规则复核。
  6. 判断是否需要先做页面方向处理:
    • 页面整体横竖方向错误:使用 90/180/270 度旋转。
    • 轻微倾斜影响阅读或 OCR:可用 deskew,但复杂图像问题优先交给 PDF Processor。
  7. 不修改原始文件,最终 PDF 写入输出目录,manifest、报告、交接文件和运行记录写入 Skill archive。
2. 判断拆分还是合并

拆分场景:

  • 一份 PDF 中包含多份文书。
  • 页面顶部出现新的文书标题。
  • 页码从“第1页共N页”或 1/N 重新开始。
  • 出现新的落款、函号、案号、收件法院、合同编号或签署页。
  • 版式从正文续页切换为新的表单、封面或附件。

合并场景:

  • 多个 PDF 是同一份文书的连续页。
  • 标题相同,页码连续,例如第1页共2页、第2页共2页。
  • 后一文件明显是前一文件的续页、签署页、付款页或附件页。
  • 文件名只是机械页码,不能代表独立文书。

重命名场景:

  • 整份 PDF 就是一份完整文书,不需要拆分也不需要合并。
  • 当前文件名只有页码、日期或扫描仪默认名,不能反映文书内容。
  • 在 manifest 中使用 input_file 指定来源,用 suggested_filename 给出规范文件名。
  • 如果同时需要旋转或倾斜校正,可以在同一条 segment 中加上 rotate 或 deskew。

如果边界不确定,不要强行合并或拆开;在 manifest 中标记 needs_review: true,并在文件名中保留页码提示。

常见法律文书的细分识别规则见 references/recognition-rules.md。

3. 命名规则

默认文件名格式:

text
文书名称 关键主体 补充区分.pdf

规则:

  1. 默认不添加序号;只有用户明确要求排序编号,或同名文件无法通过内容区分时,才添加序号。
  2. 文书名称优先使用首页标题,不使用“第 1-2 页”这类临时名。
  3. 关键主体优先写客户/委托人、相对方和案由/法律关系。
  4. 我方律所、承办律师、常见出具机构通常不写入文件名,因为这些信息对用户来说已知且重复。
  5. 日期能稳定识别时用 YYYYMMDD;日期缺失或 OCR 可疑时省略。
  6. 文件名内部默认用空格分隔,不使用下划线。
  7. 同名文件由脚本自动追加 1、 2,不覆盖已有文件。
  8. 低置信度文件使用 待确认 页码.pdf 或 疑似文书名称 页码.pdf。

示例:

text
专项法律服务合同 北京青柏教育咨询有限公司.pdf
委托代理合同 张家宁与杭州叠影科技有限公司 著作权权属侵权纠纷.pdf
授权委托书 张家宁.pdf
律师事务所函 张家宁诉杭州叠影科技有限公司.pdf
民事起诉状 张家宁诉杭州叠影科技有限公司 普通版.pdf
民事起诉状 张家宁诉杭州叠影科技有限公司 要素式.pdf
Show full SKILL.md (178 more words)Show less

Manifest

执行前先生成 organize_manifest.json。字段说明见 references/organize-manifest-schema.md。

所有 segment 在执行前都会被编译为「文件 + 页码」的页引用集合:拆分、复制、合并、跨源拼合都只是引用运算,manifest 即引用表,物理 PDF 在确认执行前一个字节不动。执行器只消费引用表;整文件单源段保留字节级复制。

拆分示例
json
{
  "source_pdf": "/path/to/ocr.pdf",
  "output_dir": "/path/to/output",
  "segments": [
    {
      "id": "D001",
      "pages": "1-2",
      "suggested_filename": "专项法律服务合同 北京青柏教育咨询有限公司.pdf",
      "title": "专项法律服务合同",
      "confidence": "high",
      "needs_review": false,
      "evidence": "第 1 页标题为专项法律服务合同,页脚显示第1页共2页。"
    }
  ]
}
合并示例
json
{
  "output_dir": "/path/to/output",
  "segments": [
    {
      "id": "D001",
      "source_items": [
        {"file": "/path/to/起诉状 第1-2页.pdf"},
        {"file": "/path/to/起诉状 第3-4页.pdf"}
      ],
      "suggested_filename": "民事起诉状 张家宁诉杭州叠影科技有限公司 普通版.pdf",
      "title": "民事起诉状",
      "confidence": "high",
      "needs_review": false,
      "evidence": "两份 PDF 标题一致且页码连续,第二份为同一份起诉状续页。"
    }
  ]
}
跨源页引用拼合示例

从多份来源 PDF 各取若干页拼成一份新材料(如证据组合卷),使用 refs 字段:

json
{
  "output_dir": "/path/to/output",
  "segments": [
    {
      "id": "D001",
      "refs": [
        {"file": "/path/to/起诉状.pdf", "pages": "5"},
        {"file": "/path/to/证据卷.pdf", "pages": "1-3"},
        {"file": "/path/to/补充说明.pdf"}
      ],
      "suggested_filename": "证据组合 原告提交.pdf",
      "confidence": "medium",
      "needs_review": true,
      "evidence": "起诉状第5页列明的证据与证据卷第1-3页对应。"
    }
  ]
}

refs 按数组顺序拼合;每项 pages 可省略(默认整份)。segment 的有效 id 必须唯一;省略时生成 D001、D002 等,显式编号也不能与这些默认值冲突。

方向处理示例
json
{
  "output_dir": "/path/to/output",
  "segments": [
    {
      "id": "D001",
      "input_file": "/path/to/横向扫描.pdf",
      "rotate": 90,
      "suggested_filename": "授权委托书 张家宁.pdf",
      "confidence": "medium",
      "needs_review": true
    }
  ]
}
重命名示例

不需要拆分或合并,只根据内容识别结果规范命名:

json
{
  "output_dir": "/path/to/output",
  "segments": [
    {
      "id": "D001",
      "input_file": "/path/to/scan_20260531_001.pdf",
      "suggested_filename": "民事起诉状 张家宁诉杭州叠影科技有限公司 普通版.pdf",
      "title": "民事起诉状",
      "document_type": "民事起诉状",
      "confidence": "high",
      "needs_review": false,
      "evidence": "首页标题为民事起诉状,正文提及张家宁与杭州叠影科技有限公司著作权权属侵权纠纷。"
    },
    {
      "id": "D002",
      "input_file": "/path/to/scan_20260531_002.pdf",
      "suggested_filename": "授权委托书 张家宁.pdf",
      "title": "授权委托书",
      "document_type": "授权委托书",
      "confidence": "high",
      "needs_review": false,
      "evidence": "首页标题为授权委托书,委托人张家宁。"
    }
  ]
}

执行脚本

首次执行拆分、合并或旋转时,先安装依赖:

bash
python3 -m pip install -r scripts/requirements.txt

预览计划:

bash
python3 scripts/pdf_organizer.py --manifest organize_manifest.json --dry-run

只做页引用编译与覆盖审计(不写任何 PDF),检查文件存在、页码越界、孤儿页与重复引用页:

bash
python3 scripts/pdf_organizer.py --validate-manifest organize_manifest.json

覆盖审计模式(off / warn / strict):默认 warn(孤儿页/重复引用页提示不拦截);strict 时发现孤儿页或重复引用页直接终止,一个 PDF 都不写。可命令行 --coverage-check strict 或 manifest 顶层 "coverage_check": "strict" 指定:

bash
python3 scripts/pdf_organizer.py --manifest organize_manifest.json --coverage-check strict

拆分整份扫描件时推荐 strict:它在完整引用表编译成功后,检查已引用来源的每一页是否被且仅被引用一次,避免漏页或多切。重复有效 ID、段编译失败、覆盖异常或审计异常均在写 PDF 前阻断;单独验证遇到审计异常也返回非零退出码。

这是执行前检查,不是整批事务:warn / off 模式保留逐段处理,某段失败时此前成功文件可能已写出;所有模式在后续写入、旋转或倾斜校正失败时也可能留下文件,退出码非零不代表输出目录为空。交付前必须核对 resolved manifest 的逐段 status、实际文件和报告,失败段不得当作完成。详见 清单契约与回归边界。

只检测 PDF 是否有可检索文字层:

bash
python3 scripts/pdf_organizer.py --check-text-layer "/path/to/file.pdf"

生成页面检查索引:

bash
python3 scripts/pdf_organizer.py \
  --inspect "/path/to/ocr.pdf" \
  --inspect-output "/path/to/page_inspection.json"

生成 manifest 草稿:

bash
python3 scripts/pdf_organizer.py \
  --suggest-manifest "/path/to/ocr.pdf" \
  --output-dir "/path/to/organized" \
  --manifest-output "/path/to/organize_manifest.json"
A4 标准化

将一个或多个 PDF 的每页标准化为 A4 尺寸,横向页面→A4 横版,竖向页面→A4 竖版:

bash
# 原地覆盖
python3 scripts/pdf_organizer.py --normalize-a4 file1.pdf file2.pdf

# 输出到指定目录
python3 scripts/pdf_organizer.py --normalize-a4 file1.pdf file2.pdf --normalize-output-dir /path/to/output

草稿只作为复核起点。脚本会保守使用 待确认,不要跳过人工/AI 复核直接执行。

确认后执行:

bash
python3 scripts/pdf_organizer.py --manifest organize_manifest.json

常用覆盖参数:

bash
python3 scripts/pdf_organizer.py \
  --manifest organize_manifest.json \
  --source "/path/to/ocr.pdf" \
  --output-dir "/path/to/organized" \
  --archive-root "/path/to/archive-root"

脚本默认以 strict 模式检测来源 PDF 的文字层;检测不到可检索文字层时会停止执行。只有在纯旋转、复制等不依赖内容判断的场景,才可在 manifest 中设置 "text_check": "off" 或命令行使用 --text-check off。

脚本只向输出目录写入最终 PDF,不修改源 PDF;manifest、resolved JSON、报告、handoff.json 和元数据写入 archive。

下游交接

每次正式执行都会在 archive 中生成 handoff.json。下游 Skill 优先读取该文件,而不是重新猜测文件名和文书类型。

来源数组按输出页顺序排列,仅合并连续的同源页组;同一个 file 可以在列表里重复出现,下游不得按文件名重新分组、排序或去重。pages 保持整数数组类型,包含原始顺序和重复页。

每份输出文书都带页级溯源:source_refs(归一化页引用 [{file, pages}])和 source_refs_label(人读标签如 起诉状.pdf P5 + 证据卷.pdf P1-3),下游引用任何关键信息时可回溯到具体来源页。

suggested_downstream 按文书类别给出路由标签,不绑定具体 Skill 名称:

  • 合同、协议 → 合同审查。
  • 起诉状、判决书、裁定书、答辩状、申请书、证据目录、庭审笔录 → 诉讼分析。
  • 授权委托书、律师事务所函和其他辅助材料 → 材料整理。
  • 低置信度或需复核 → 回到本 Skill 复核。
  • 未命中以上类别的 → 材料整理。

具体由哪个 Skill 消费,由下游根据当前可用的 Skill 和案件上下文自行决定。

交付检查

完成后检查:

  1. 输出 PDF 数量与 manifest segment 数一致。
  2. 拆分文件页数与 pages 对应;合并文件页数等于来源页数之和。
  3. 覆盖审计无未解释的孤儿页或重复引用页;有则说明漏页、多切或有意复用,需向用户说明。
  4. archive 中的 organize_report.md 没有 low 且 needs_review: true 的未处理项;如果有,明确提示用户人工复核。
  5. 原始 PDF 和原始碎片 PDF 未被覆盖或移动。
  6. 命名依据能回溯到 OCR 文本或页面内容,没有臆测日期、案号或主体。

© cat-xierluo, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 10 other files (scripts, references) in skills/pdf-organizer of cat-xierluo/legal-skills.

  • SKILL.md
  • CHANGELOG.md
  • LICENSE.txt
  • archive/.gitkeep
  • references/manifest-contract.md
  • references/organize-manifest-schema.md
  • references/recognition-rules.md
  • scripts/pdf_organizer.py
  • scripts/pdf_page_refs.py
  • scripts/requirements.txt
  • scripts/test_manifest_contract.py

Open the folder on GitHubat commit db2c58c

Compare with similar skills

PDF Organizer next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

PDF Organizer compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
PDF Organizer this skillcat-xierluo/legal-skills720—~2.3kAutomated safety check: PassMIT
Instrument Data To Allotropeaws-samples/amazon-bedrock-agents-healthcare-lifesciences2742 repos~2.7kAutomated safety check: PassApache-2.0
Software Certificate SkillIvanCodesDev/software-certificate-skill156—~1.6kAutomated safety check: PassMIT
Doc Cleanernotoriouslab/doc-cleaner309—~712Automated safety check: PassMIT
MineruNebutra/MinerU-Skill123—~504Automated safety check: PassMIT
Office To Mdshuyu-labs/WebCode278—~1kAutomated safety check: NotesCustom licence

Similar skills

  • Instrument Data To Allotrope

    aws-samples/amazon-bedrock-agents-healthcare-lifesciences

    Official

    Convert laboratory instrument output files (PDF, CSV, Excel, TXT) to Allotrope Simple Model (ASM) JSON format or flattened 2D CSV.

    274 GitHub starsUsed in 2 repos~2.7k tokens
    Documents & OfficeAuto-check passed
  • Software Certificate Skill

    IvanCodesDev/software-certificate-skill

    面向普通用户,从真实软件项目全自动生成中国软件著作权申请资料:一次收集登记事实,自动分析业务、选择可追溯源码、取得真实界面证据,生成申请表信息、规范黑白灰操作手册、代码前后30页或全部材料及真实 DOCX/PDF;内部验证、渲染、哈希与备份只进入系统临时运行区,项目最终仅保留正式资料。适配 Codex、Claude…

    156 GitHub stars~1.6k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed
  • Doc Cleaner

    notoriouslab/doc-cleaner

    Convert PDF, DOCX, XLSX, and text files to clean, structured Markdown.

    309 GitHub stars~712 tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed
  • Mineru

    Nebutra/MinerU-Skill

    An AI-Native skill for parsing PDF / Office / image files into Markdown with MinerU — a fast, zero-config document parser for AI agents.

    123 GitHub stars~504 tokensUpdated 17 days ago
    Documents & OfficeAuto-check passed
  • Office To Md

    shuyu-labs/WebCode

    Convert Office documents (Word, Excel, PowerPoint, PDF) to Markdown format.

    278 GitHub stars~1k tokensUpdated 3 mo ago
    Documents & OfficeAuto-check: notes
  • Cc Streaming Export Safety

    doccker/cc-use-exp

    当实现用户驱动的大文件导出或批量序列化(Excel/CSV/JSON/JSONL/PDF,数据量未知或超过 1 万行/10 MB)时触发;普通小文件下载、静态资源下载、非导出 Writer/Report 类不触发。防止 OOM、临时文件残留、同步导出阻塞 HTTP 线程和表格公式注入。

    1.1k GitHub stars~2.2k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed

More from cat-xierluo/legal-skills

All 62 skills in this repo
  • Elements-Style Complaint Generator

    cat-xierluo/legal-skills

    Converts a lawyer's ordinary complaint or a described case into the Supreme People's Court's elements-style Word template, with layout checks on the result.

    721 GitHub stars~2.6k tokensUpdated yesterday
    Auto-check: notes
  • Lecture Performance Review

    cat-xierluo/legal-skills

    Analyzes raw lecture transcripts for verbal tics, pacing, time use and promise follow-through, with optional slide-by-slide comparison and cross-session tracking.

    721 GitHub stars~2.2k tokensUpdated yesterday
    Auto-check passed
  • De-AI Polish for Chinese Articles

    cat-xierluo/legal-skills

    Detects and rewrites machine-sounding patterns in the body text of Chinese articles while keeping the author's facts, headings and legal terms intact.

    721 GitHub stars~2.9k tokensUpdated yesterday
    Auto-check passed
  • GitHub Star Manager

    cat-xierluo/legal-skills

    Finds GitHub projects mentioned in articles or screenshots and stars them, tracks updates to your starred repos, and builds an HTML dashboard to browse them.

    721 GitHub stars~2k tokensUpdated yesterday
    Auto-check: notes
  • Legal Harness Initializer

    cat-xierluo/legal-skills

    Sets up or incrementally updates AGENTS.md and CLAUDE.md for legal professionals, with a minimal safety baseline and a check that a new session loads and follows the rules.

    721 GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Moot Court Simulation Builder

    cat-xierluo/legal-skills

    Chinese-language skill that organizes a case file into a multi-role mock trial with judge, parties and clerk, producing a transcript, issue review and a to-strengthen list.

    721 GitHub stars~1.4k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about PDF Organizer

What does PDF Organizer do?

当需要整理法律 PDF 时使用:检测文字层,生成页面索引、整理草稿和下游交接文件,按内容拆分、合并或直接重命名 OCR 后双层扫描件并规范命名;支持跨源页引用拼合与页覆盖审计(孤儿页/重复引用);可做旋转与倾斜校正,不做 OCR 或压缩。. PDF Organizer is an agent skill from cat-xierluo/legal-skills.

When should I use PDF Organizer?

PDF Organizer fits situations like: tasks that involve PDF.

How do I install PDF Organizer in Claude Code?

Run `npx skills add cat-xierluo/legal-skills --skill pdf-organizer -a claude-code`. Or copy the skill folder (skills/pdf-organizer in cat-xierluo/legal-skills) into .claude/skills/pdf-organizer in your project. Claude Code loads it when a task matches its description.

How do I install PDF Organizer in Codex?

Run `npx skills add cat-xierluo/legal-skills --skill pdf-organizer -a codex`. Or copy the skill folder (skills/pdf-organizer in cat-xierluo/legal-skills) into .agents/skills/pdf-organizer in your project. Codex loads it when a task matches its description.

Can I use PDF Organizer in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add cat-xierluo/legal-skills --skill pdf-organizer -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/pdf-organizer, .gemini/skills/pdf-organizer, .github/skills/pdf-organizer and .opencode/skills/pdf-organizer in your project.

What does PDF Organizer need to run?

Going by SKILL.md and its folder, PDF Organizer needs Python for the scripts in its folder and the command-line tools its instructions call (python3 and brew). Our summary lists: Python 3.

Does PDF Organizer access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is PDF Organizer safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does PDF Organizer use?

PDF Organizer is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does PDF Organizer use?

About 2.3k tokens (SKILL.md is roughly 9.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4k tokens, read only when the agent opens those files.

What are the alternatives to PDF Organizer?

Skills that share tags, products or a category with PDF Organizer: Instrument Data To Allotrope (aws-samples/amazon-bedrock-agents-healthcare-lifesciences, 274 stars), Software Certificate Skill (IvanCodesDev/software-certificate-skill, 156 stars), Doc Cleaner (notoriouslab/doc-cleaner, 309 stars) and Mineru (Nebutra/MinerU-Skill, 123 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains PDF Organizer?

cat-xierluo (a GitHub user) maintains it in cat-xierluo/legal-skills, which has 720 GitHub stars. The repository holds 62 skills in this directory. The repository was last updated on October 10, 2026.

Source: cat-xierluo/legal-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.