Markitdown
ImCa0/just-laws
Convert files and office documents to Markdown. An agent skill from ImCa0/just-laws.
Parse local document files into LLM-ready content on the daemon itself — PDFs → Markdown / structured JSON (with bounding boxes) / page screenshots via the bundled lit CLI.
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Prismer-AI/PrismerCloud liteparse --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .claude/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .claude/skills/liteparse && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .claude/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparseType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Prismer-AI/PrismerCloud liteparse --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .agents/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .agents/skills/liteparse && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .agents/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Prismer-AI/PrismerCloud liteparse --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .cursor/skills/liteparse && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .cursor/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Prismer-AI/PrismerCloud.git --path sdk/cloud/catalog/skills/liteparse--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Prismer-AI/PrismerCloud liteparse --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .gemini/skills/liteparse && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .gemini/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Prismer-AI/PrismerCloud liteparseInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .github/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .github/skills/liteparse && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .github/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Prismer-AI/PrismerCloud --skill liteparse -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Prismer-AI/PrismerCloud liteparse --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Prismer-AI/PrismerCloud.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/sdk/cloud/catalog/skills/liteparse .opencode/skills/liteparse && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "liteparse" agent skill from https://github.com/Prismer-AI/PrismerCloud/tree/main/sdk/cloud/catalog/skills/liteparse into .opencode/skills/liteparse/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "liteparse", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
liteparseParse local document files into LLM-ready content on the daemon itself — PDFs → Markdown / structured JSON (with bounding boxes) / page screenshots via the bundled lit CLI.
Liteparse is an agent skill from Prismer-AI/PrismerCloud. Parse local document files into LLM-ready content on the daemon itself — PDFs → Markdown / structured JSON (with bounding boxes) / page screenshots via the bundled lit CLI. PDF works out of the box, offline, zero external deps (Tesseract + PDFium are bundled): digital PDFs use a near-instant native text path (--no-ocr), scanned PDFs fall back to bundled OCR. Image files (PNG/JPG) additionally need ImageMagick, and Office files (DOCX/XLSX/PPTX) need LibreOffice — install those on demand only when required. Use…
Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Documents & Office, covering Document parsing, PDF and Word documents. It works with LibreOffice, Microsoft Excel, Microsoft PowerPoint and Microsoft Word. The licence is MIT.
4 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit e5d9444. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
apt-getbrewcurlnpmnpxFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
github.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Liteparse loads about 2.1k tokens when it runs. Until then it costs about 186 tokens; SKILL.md has 900 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Prismer-AI/PrismerCloud at commit e5d9444, republished under its MIT licence (© Prismer-AI). 900 words, ~2,133 tokens.
.claude/skills/liteparse/SKILL.md (or your agent's skills folder).LiteParse turns document files into text the model can read — on the
daemon machine, offline, zero cloud dependency. It ships as the
@llamaindex/liteparse npm package (CLI: lit) with PDFium and
Tesseract bundled inside the package — the digital-PDF and scanned-OCR
paths need no external system dependencies.
Upstream:
run-llama/liteparse
- skill
run-llama/llamaparse-agent-skills. Apache-2.0 (core) / MIT (skill), © LlamaIndex. Vendored as a Prismer built-in.
| Tier | Inputs | External dep | Status |
|---|---|---|---|
| 1 — PDF | digital + scanned PDF | none (Tesseract + PDFium bundled) | Default, out of the box — lit is baked into the daemon image |
| 2 — Images | PNG / JPG / TIFF / … | ImageMagick | Install on demand |
| 3 — Office | DOCX / XLSX / PPTX / ODT | LibreOffice (~1GB) | Install on demand |
Only Tier 1 is guaranteed present. Tiers 2 and 3 pull large system dependencies that are deliberately not baked into the daemon image — install them yourself, only when the actual input requires it (see below). Do not assume they are already installed.
lit is available on the daemon. Sanity-check once, quietly:
command -v lit && lit --version # baked into the imageIf (and only if) it is somehow missing, install the package locally — it is ~38MB with the OCR/render binaries bundled and installs in a few seconds:
npm i @llamaindex/liteparse && npx lit --versionDo not rely on the CLI "self-installing on first use" — that path is unreliable (npm permission / cache conflicts). Either it is baked in, or you install the package explicitly as above.
--no-ocr. PDFium
extracts the text directly — near-instant (~0.35s/page), no OCR.
This is the fast, cheap path; prefer it whenever the PDF is
computer-generated.If unsure which kind you have, run --no-ocr first: empty/garbled output
means it's a scan → re-run without --no-ocr to OCR it.
lit parse report.pdf --no-ocr # digital PDF → fast native text
lit parse scan.pdf # scanned PDF → bundled OCR
lit parse report.pdf --format markdown -o out.md # Markdown output to a file
lit parse report.pdf --format json -o out.json # structured JSON + bounding boxes
lit parse report.pdf --target-pages "1-5,10,15-20" # page subset (huge speedup on OCR path)--format accepts markdown (default) or json (adds per-block
bounding boxes). Other useful flags: -o/--output, --target-pages "1-5,10", --dpi <n> (render DPI, default 150 — bump to 300 for small
scanned text), --ocr-language eng+chi_sim (Tesseract lang codes for the
OCR path), --password '****' (encrypted PDFs), -q/--quiet.
lit screenshot report.pdf -o ./shots # all pages → PNG
lit screenshot report.pdf --target-pages "1,3,5" -o ./shots
lit screenshot report.pdf --dpi 300 -o ./shots # high-resParsing standalone image files (PNG/JPG/TIFF/…) needs ImageMagick, which is not in the daemon image. Install it the first time you actually have an image input:
apt-get update && apt-get install -y imagemagick # Debian/Ubuntu (daemon image)
brew install imagemagick # macOS (local dev)
lit parse diagram.png --format markdownIf ImageMagick can't be installed (no network / no apt), report
ImageMagick unavailable and stop — don't fabricate the image's contents.
DOCX / XLSX / PPTX / ODT conversion needs LibreOffice (~1GB), which is intentionally not baked into the daemon image. Install it only when you genuinely need to parse an Office file (it's a big download — don't do it speculatively):
apt-get update && apt-get install -y libreoffice # Debian/Ubuntu (daemon image)
brew install --cask libreoffice # macOS (local dev)
lit parse deck.pptx --format markdownIf LibreOffice can't be installed, say so and ask the user for a PDF export of the document instead of guessing.
A shared cloud parse service for heavy formats (offloading Office / hi-res OCR to a hosted backend so agents don't install 1GB deps) is a future TODO — it is not running today. Don't route to a cloud OCR endpoint.
ingestTwo skills, clean split by source location:
liteparse (this skill) — local files on disk. PDF / image /
Office files already on the machine → Markdown / JSON / screenshots,
fully offline. This is the only local document-parsing path.ingest (sibling skill) — web URLs and search. cloud load /
cloud search fetch and compress remote pages. liteparse cannot
fetch URLs; if you only have a URL, either curl -sL <url> -o file
then parse the local copy with liteparse, or route to ingest.Decision rule: local file → liteparse. Web page / search → ingest.
assets skill). URL-only → curl -sL <url> -o <name> first, then
parse the local copy (or use ingest).lit parse <f> --no-ocr (fast native).lit parse <f> (OCR); set
--ocr-language, bump --dpi 300 for small text, and use
--target-pages to avoid OCR'ing pages you don't need.--format json.lit screenshot, then read
the PNG.lit returned. Empty or low-confidence page → say so; never guess.$PRISMER_ARTIFACTS_DIR, then
explicitly run cloud deliver <abs-path> for the current reply or
cloud task attach <abs-path> for a task. Auto-scan is OFF; writing the
file alone is not delivery. See office-artifacts SKILL.md §Delivery contract.Parsed <filename>: <N> pages, format=<markdown|json>, path=<native|ocr> — then answer the user's actual question, citing page
numbers. Don't dump the whole parsed body into chat unless asked.Rendered <N> page screenshot(s) → <dir>, then read
the relevant ones.Forbidden unless lit actually ran with exit 0 and produced output:
"parsed the document", "the PDF says…", "extracted the table". If lit is
unavailable, a required dependency (ImageMagick / LibreOffice) can't be
installed, or the parse failed, completion text must start with 无法解析 文档 (reason) / Cannot parse document (reason) — do not substitute
guessed content.
© Prismer-AI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in sdk/cloud/catalog/skills/liteparse of Prismer-AI/PrismerCloud.
Open the folder on GitHubat commit e5d9444
Liteparse next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Liteparse this skillPrismer-AI/PrismerCloud | 1.6k | — | ~2.1k | Automated safety check: Pass | MIT | |
| MarkitdownImCa0/just-laws | 781 | 14 repos | ~3.2k | Automated safety check: Notes | MIT | |
| Markitdownjimmc414/Kosmos | 595 | 2 repos | ~1.7k | Automated safety check: Pass | None | |
| Nutrient Document Processingaffaan-m/ECC | 276k | 4 repos | ~1.5k | Automated safety check: Pass | MIT | |
| Document Converterwentorai/Research-Claw | 858 | — | ~1.3k | Automated safety check: Pass | Custom licence | |
| To MarkdownMathews-Tom/armory | 329 | — | ~2k | Automated safety check: Pass | MIT |
ImCa0/just-laws
Convert files and office documents to Markdown. An agent skill from ImCa0/just-laws.
jimmc414/Kosmos
Convert various file formats (PDF, Office documents, images, audio, web content, structured data) to Markdown optimized for LLM processing.
affaan-m/ECC
Process, convert, OCR, extract, redact, sign, and fill documents using the Nutrient DWS API.
wentorai/Research-Claw
Convert Office documents (PPTX, DOCX, XLSX, PDF, HTML, CSV, JSON, XML, images) to Markdown using Microsoft MarkItDown.
Mathews-Tom/armory
Convert any file or URL to clean Markdown: PDF, DOCX, XLSX, PPTX, HTML, images (OCR), audio, CSV, YouTube.
xu-xiang/everything-claude-code-zh
使用 Nutrient DWS API 处理、转换、OCR、提取、脱敏、签名和填充文档。支持 PDF、DOCX、XLSX、PPTX、HTML 和图像。
Prismer-AI/PrismerCloud
Gives an agent account-scoped access to Gmail, Calendar, Drive, Contacts, Docs and Sheets through the gws CLI or a bundled Python client.
Prismer-AI/PrismerCloud
Walks an agent through creating, importing, editing, validating, testing and publishing Prismer Skills with a fixed workflow and bundled scripts.
Prismer-AI/PrismerCloud
Operates a mailbox from the terminal with the external Himalaya CLI over IMAP, SMTP, Notmuch or Sendmail, separate from any built-in email gateway adapter.
Prismer-AI/PrismerCloud
Generates one image from a text prompt with a bundled Node.js helper and delivers it once as the attachment to the current Prismer reply.
Prismer-AI/PrismerCloud
Produces 3Blue1Brown-style explainer animations with Manim Community Edition for math, algorithms, equations and architecture diagrams, with planning and rendering references.
Prismer-AI/PrismerCloud
Creates or updates Prismer role templates from a persona, SOP or job description, and turns a role into a working agent that runs its first task through a bundled script.
Categories
Parse local document files into LLM-ready content on the daemon itself — PDFs → Markdown / structured JSON (with bounding boxes) / page screenshots via the bundled lit CLI. Liteparse is an agent skill from Prismer-AI/PrismerCloud. Parse local document files into LLM-ready content on the daemon itself — PDFs → Markdown / structured JSON (with bounding boxes) / page screenshots via the bundled lit CLI.
Liteparse fits situations like: the user attaches; points to a local file that must be read before reasoning; asks to extract text / tables / page images from a file on disk.
Run `npx skills add Prismer-AI/PrismerCloud --skill liteparse -a claude-code`. Or copy the skill folder (sdk/cloud/catalog/skills/liteparse in Prismer-AI/PrismerCloud) into .claude/skills/liteparse in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Prismer-AI/PrismerCloud --skill liteparse -a codex`. Or copy the skill folder (sdk/cloud/catalog/skills/liteparse in Prismer-AI/PrismerCloud) into .agents/skills/liteparse in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Prismer-AI/PrismerCloud --skill liteparse -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/liteparse, .gemini/skills/liteparse, .github/skills/liteparse and .opencode/skills/liteparse in your project.
Going by SKILL.md and its folder, Liteparse needs the command-line tools its instructions call (apt-get, brew, curl, npm and npx). Our summary lists: Node.js.
SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Liteparse is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.1k tokens (SKILL.md is roughly 8.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Liteparse: Markitdown (ImCa0/just-laws, 781 stars), Markitdown (jimmc414/Kosmos, 595 stars), Nutrient Document Processing (affaan-m/ECC, 276k stars) and Document Converter (wentorai/Research-Claw, 858 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Prismer-AI (a GitHub organization) maintains it in Prismer-AI/PrismerCloud, which has 1,555 GitHub stars. The repository holds 88 skills in this directory. The repository was last updated on September 30, 2026.
Source: Prismer-AI/PrismerCloud on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.