Agent skill

Paper Image Extractor

by LigphiDonk in LigphiDonk/Oh-my--paper

Extracts figures from a research paper, preferring the arXiv source package for original-quality images and falling back to PDF extraction.

MITAuto-check passedResearch & Science

Install Paper Image Extractor

skills CLI
$ npx skills add LigphiDonk/Oh-my--paper --skill paper-image-extractor -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install LigphiDonk/Oh-my--paper paper-image-extractor --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/LigphiDonk/Oh-my--paper.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/paper-image-extractor .claude/skills/paper-image-extractor && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
paper-image-extractor
GitHub stars
738
Used in
1 other repo
Token cost
~810 tokens
SKILL.md length
279 words
Files
2 (incl. scripts)
Skills in repo
27
Repo updated
First seen
Licence
MIT

At a glance

Extracts figures from a research paper, preferring the arXiv source package for original-quality images and falling back to PDF extraction.

  • Works in 4 steps: Download source:… → Extract and look for pics/, figures/,… → Copy image files to output directory → …
  • Collecting figures from an arXiv paper for a literature review or slides
  • SKILL.md covers Canonical Summary, Trigger Rules, Resource Use Rules and Execution Contract, plus 4 more sections
  • Runs Python scripts from its folder; calls python; reaches arxiv.org

What it does

This skill pulls the figures out of a paper using a three-tier strategy. The preferred route downloads the arXiv source package, looks in directories such as pics, figures, fig, images and img, copies the image files and converts PDF figures to PNG. If that fails, scripts/extract_images.py extracts figures from the PDF, and as a last resort embedded image objects are taken from the compiled PDF with PyMuPDF.

Images go to the output directory you specify, together with an index.md that lists each image's metadata and a source label: arxiv-source, pdf-figure or pdf-extraction. The skill needs Python 3.8 or newer, PyMuPDF and requests, plus network access to arXiv. It says to resolve paths from its own folder, to explain any blocker and fall back to manual steps when a dependency is missing, and to save outputs in the project workspace rather than the skill folder.

When your agent uses it

  • Collecting figures from an arXiv paper for a literature review or slides
  • Getting original-quality images instead of screenshots of a PDF
  • Building an indexed set of figures with source labels

Example prompts

  • “Extract every figure from the Attention Is All You Need paper into ./figures.”
  • “Pull the figures out of this paper's PDF, since the arXiv source is not available.”
  • “Make an index.md of the extracted images with their source labels.”

Requirements

  • Python 3.8 or newer with PyMuPDF and requests
  • Network access to arXiv

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Download source: https://arxiv.org/e-print/[PAPER_ID]
  2. Extract and look for pics/, figures/, fig/, images/, img/ directories
  3. Copy image files to output directory
  4. Convert PDF figures to PNG

What it can do on your machine

Read from SKILL.md and the folder at commit 6baece9. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • arxiv.org

    Also links to:

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Paper Image Extractor loads about 810 tokens when it runs. Until then it costs about 27 tokens; SKILL.md has 279 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~27
When it runs · the whole SKILL.md, loaded when a task matches
~810

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from LigphiDonk/Oh-my--paper at commit 6baece9, republished under its MIT licence (© LigphiDonk). 279 words, ~810 tokens.

Download SKILL.mdSave it as .claude/skills/paper-image-extractor/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
paper-image-extractor
description
Extract figures from papers — prioritizes arXiv source package for high-quality images
id
paper-image-extractor
version
1.0.0
stages
publication
tools
read_file, search_project, write_file, run_terminal
summary
Extract figures from papers — prioritizes arXiv source package for high-quality images
primaryIntent
research
intents
research
capabilities
multimodal
domains
cs-ai, vision
keywords
paper-image-extractor, image extraction, multimodal, cs-ai, vision, paper, image, extractor, extract, figures, from, papers

paper-image-extractor

Canonical Summary

Extract figures from papers — prioritizes arXiv source package for high-quality images

Trigger Rules

Use this skill when the user request matches its research workflow scope. Prefer the bundled resources instead of recreating templates or reference material. Keep outputs traceable to project files, citations, scripts, or upstream evidence.

Resource Use Rules

  • Treat scripts/ as optional helpers. Run them only when their dependencies are available, keep outputs in the project workspace, and explain a manual fallback if execution is blocked.

Execution Contract

  • Resolve every relative path from this skill directory first.
  • Prefer inspection before mutation when invoking bundled scripts.
  • If a required runtime, CLI, credential, or API is unavailable, explain the blocker and continue with the best manual fallback instead of silently skipping the step.
  • Do not write generated artifacts back into the skill directory; save them inside the active project workspace.

Upstream Instructions

You are the Paper Image Extractor for Dr. Claw.

Goal

Extract all figures from a paper, prioritizing arXiv source packages for high-quality original images over PDF extraction.

Extraction Strategy (3-tier priority)

Priority 1: arXiv Source Package (Best)

  1. Download source: https://arxiv.org/e-print/[PAPER_ID]
  2. Extract and look for pics/, figures/, fig/, images/, img/ directories
  3. Copy image files to output directory
  4. Convert PDF figures to PNG

Priority 2: PDF Figure Extraction (Fallback)

bash
python scripts/extract_images.py "[PAPER_ID]" "[OUTPUT_DIR]" "[INDEX_PATH]"

Priority 3: Direct PDF Image Extraction (Last Resort)

Extract embedded image objects from the compiled PDF using PyMuPDF.

Output

  • Images saved to specified output directory
  • index.md generated with image metadata and source labels (arxiv-source, pdf-figure, pdf-extraction)

Scripts

  • scripts/extract_images.py — Main extraction script with 3-tier strategy

Dependencies

  • Python 3.8+, PyMuPDF (fitz), requests
  • Network access (arXiv)

Based on evil-read-arxiv — an automated paper reading workflow. MIT License.

© LigphiDonk, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/paper-image-extractor of LigphiDonk/Oh-my--paper.

  • SKILL.md
  • scripts/extract_images.py

Open the folder on GitHubat commit 6baece9

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in LigphiDonk/Oh-my--paper, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Paper Image Extractor next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Paper Image Extractor compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Paper Image Extractor this skillLigphiDonk/Oh-my--paper7381 repos~810Automated safety check: PassMIT
Ref Downloaderltczding-gif/ref-downloader139—~5.9kAutomated safety check: PassMIT
Paper and Web Source Analyzerjuliye2025/evil-read-arxiv1.7k—~715Automated safety check: PassNone
Paper Figure Extractorjuliye2025/evil-read-arxiv1.7k—~298Automated safety check: PassNone
Papers Skillsickn33/agentic-awesome-skills47k1 repos~2.1kAutomated safety check: PassMIT
PDF Processinganthropics/skills180k48 repos~2kAutomated safety check: PassProprietary

Similar skills

  • Ref Downloader

    ltczding-gif/ref-downloader

    A skill your agent uses when the user asks to batch-download academic PDFs with ref-downloader — either ALL references of one paper (Mode A: DOI or PDF input), OR a custom batch of papers (Mode B…

    139 GitHub stars~5.9k tokensUpdated 4 mo ago
    Research & ScienceAuto-check passed
  • Paper and Web Source Analyzer

    juliye2025/evil-read-arxiv

    Analyzes arXiv papers, PDFs, project pages and blogs and writes an evidence-backed Obsidian note with images and a knowledge-graph entry.

    1.7k GitHub stars~715 tokensUpdated 24 days ago
    Research & ScienceAuto-check passed
  • Paper Figure Extractor

    juliye2025/evil-read-arxiv

    Pulls architecture, method and result figures from an arXiv paper or PDF into an Obsidian vault and writes an index of them.

    1.7k GitHub stars~298 tokensUpdated 24 days ago
    Research & ScienceAuto-check passed
  • Papers Skill

    sickn33/agentic-awesome-skills

    Skill for academic research workflows: search Semantic Scholar (200M+ papers), inspect citations, download arXiv PDFs, and extract PDF text.

    47k GitHub starsUsed in 1 repo~2.1k tokens
    Research & ScienceAuto-check passed
  • PDF Processing

    anthropics/skills

    Official

    Handles everyday PDF jobs in Python and on the command line: extract text and tables, merge, split, rotate, watermark, fill forms, encrypt and OCR.

    180k GitHub starsUsed in 48 repos~2k tokens
    Documents & OfficeAuto-check passed
  • PDF Processing Guide

    shareAI-lab/learn-claude-code

    Gives the agent command-line and Python recipes for reading, creating, merging and splitting PDF files, plus tips for large and scanned documents.

    78k GitHub starsUsed in 5 repos~646 tokens
    Documents & OfficeAuto-check passed

More from LigphiDonk/Oh-my--paper

All 27 skills in this repo
  • Preprint Search on bioRxiv

    LigphiDonk/Oh-my--paper

    Searches bioRxiv life sciences preprints by keyword, author, date range or category with a Python script, returning JSON metadata and optional PDF downloads.

    738 GitHub starsUsed in 12 repos~3.7k tokens
    Auto-check passed
  • Literature PDF OCR Library Builder

    LigphiDonk/Oh-my--paper

    Searches and downloads legally accessible academic PDFs, OCRs them to Markdown, and organizes the results into a traceable, AI-readable literature library.

    738 GitHub stars~1.1k tokensUpdated 5 mo ago
    Auto-check passed
  • Inno Code Survey

    LigphiDonk/Oh-my--paper

    Finds and clones missing code repositories for a chosen research idea, then writes a survey that maps academic concepts to their implementations.

    738 GitHub stars~3.6k tokensUpdated 5 mo ago
    Auto-check passed
  • Turns experimental data such as CSV, JSON or TensorBoard logs into statistical significance tests, visualizations and a drafted Results section.

    738 GitHub stars~3k tokensUpdated 5 mo ago
    Auto-check passed
  • Citation Verification Guide

    LigphiDonk/Oh-my--paper

    Lays out principles for catching fake, mismatched, or inconsistently formatted citations in academic writing, checked through live web search.

    738 GitHub stars~2.2k tokensUpdated 5 mo ago
    Auto-check passed
  • Single-Cell Initial Analysis

    LigphiDonk/Oh-my--paper

    Runs a seven-step quality-control and exploration pipeline on scRNA-seq, CyTOF or flow cytometry data and writes a plain-language report of what it found.

    738 GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check passed

Works with

Questions about Paper Image Extractor

What does Paper Image Extractor do?

Extracts figures from a research paper, preferring the arXiv source package for original-quality images and falling back to PDF extraction. This skill pulls the figures out of a paper using a three-tier strategy. The preferred route downloads the arXiv source package, looks in directories such as pics, figures, fig, images and img, copies the image files and converts PDF figures to PNG.

When should I use Paper Image Extractor?

Paper Image Extractor fits situations like: collecting figures from an arXiv paper for a literature review or slides; getting original-quality images instead of screenshots of a PDF; building an indexed set of figures with source labels.

How do I install Paper Image Extractor in Claude Code?

Run `npx skills add LigphiDonk/Oh-my--paper --skill paper-image-extractor -a claude-code`. Or copy the skill folder (skills/paper-image-extractor in LigphiDonk/Oh-my--paper) into .claude/skills/paper-image-extractor in your project. Claude Code loads it when a task matches its description.

How do I install Paper Image Extractor in Codex?

Run `npx skills add LigphiDonk/Oh-my--paper --skill paper-image-extractor -a codex`. Or copy the skill folder (skills/paper-image-extractor in LigphiDonk/Oh-my--paper) into .agents/skills/paper-image-extractor in your project. Codex loads it when a task matches its description.

Can I use Paper Image Extractor in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add LigphiDonk/Oh-my--paper --skill paper-image-extractor -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/paper-image-extractor, .gemini/skills/paper-image-extractor, .github/skills/paper-image-extractor and .opencode/skills/paper-image-extractor in your project.

What does Paper Image Extractor need to run?

Going by SKILL.md and its folder, Paper Image Extractor needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.8 or newer with PyMuPDF and requests; Network access to arXiv.

Does Paper Image Extractor access the network?

SKILL.md names 2 domains. In commands or code: arxiv.org; the agent is likely to contact it when it follows the instructions. As links in the text: github.com. This is read from the text; nothing was executed.

Is Paper Image Extractor safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Paper Image Extractor use?

Paper Image Extractor is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Paper Image Extractor use?

About 810 tokens (SKILL.md is roughly 3.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Paper Image Extractor?

Skills that share tags, products or a category with Paper Image Extractor: Ref Downloader (ltczding-gif/ref-downloader, 139 stars), Paper and Web Source Analyzer (juliye2025/evil-read-arxiv, 1.7k stars), Paper Figure Extractor (juliye2025/evil-read-arxiv, 1.7k stars) and Papers Skill (sickn33/agentic-awesome-skills, 47k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Paper Image Extractor?

LigphiDonk (a GitHub user) maintains it in LigphiDonk/Oh-my--paper, which has 738 GitHub stars. The repository holds 27 skills in this directory. The repository was last updated on April 15, 2026.

Source: LigphiDonk/Oh-my--paper on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.