Agent skill

Extract Slide Text

by pamelafox in pamelafox/presentation-skills

Extract text from each page of a PDF into a markdown file using pdftotext.

MITAuto-check passedDocuments & Office

Install Extract Slide Text

skills CLI
$ npx skills add pamelafox/presentation-skills --skill extract-slide-text -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install pamelafox/presentation-skills extract-slide-text --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/pamelafox/presentation-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/extract-slide-text .claude/skills/extract-slide-text && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
extract-slide-text
GitHub stars
125
Token cost
~440 tokens
SKILL.md length
139 words
Files
2
Skills in repo
14
Repo updated
First seen
Licence
MIT

At a glance

Extract text from each page of a PDF into a markdown file using pdftotext.

  • : extract text from PDF slides
  • SKILL.md covers Arguments, Output format, Why this matters and Prerequisites
  • Runs Python scripts from its folder; calls uv, brew and apt-get
  • Get slide text content

What it does

Extract Slide Text is an agent skill from pamelafox/presentation-skills. Extract text from each page of a PDF into a markdown file using pdftotext. Produces slideascii.md with a heading, image reference, and extracted text per slide. USE FOR: extract text from PDF slides, get slide text content, PDF to markdown text, slideascii.md.

Its SKILL.md is about 440 tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `extract_slide_text.py`).

It sits in Documents & Office, covering Slides and decks and PDF. The repository describes itself as: Skills for AI agents to process presentations - helpful for teachers and speakers. The licence is MIT.

When your agent uses it

  • : extract text from PDF slides
  • Get slide text content
  • PDF to markdown text

Example prompts

  • “/extract-slide-text”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit 2b809b3. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships script files (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv
    • brew
    • apt-get

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Extract Slide Text loads about 440 tokens when it runs. Until then it costs about 70 tokens; SKILL.md has 139 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~70
When it runs · the whole SKILL.md, loaded when a task matches
~440

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from pamelafox/presentation-skills at commit 2b809b3, republished under its MIT licence (© pamelafox). 139 words, ~440 tokens.

Download SKILL.mdSave it as .claude/skills/extract-slide-text/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
extract-slide-text
description
Extract text from each page of a PDF into a markdown file using pdftotext. Produces slide_ascii.md with a heading, image reference, and extracted text per slide. USE FOR: extract text from PDF slides, get slide text content, PDF to markdown text, slide_ascii.md.
argument-hint
<pdf_path> <output_path> [images_dir]

Extract slide text from PDF

Run the extract_slide_text.py script to extract the text content of each PDF page into a structured markdown file:

bash
uv run .agents/skills/extract-slide-text/extract_slide_text.py <pdf_path> <output_path> [images_dir]

Arguments

  • pdf_path (required): Path to the PDF file.
  • output_path (required): Path to write the output markdown file
  • images_dir (optional): Path to the slide images directory. Used to generate correct relative image references. Defaults to slide_images/.

Output format

A markdown file with one section per slide:

markdown
## Slide 1

![Slide 1](slide_images/slide_1.png)

\```
Extracted text content from slide 1
\```

## Slide 2

![Slide 2](slide_images/slide_2.png)

\```
Extracted text content from slide 2
\```

Pages with no extractable text (e.g., full-bleed images) show (no extractable text).

Why this matters

PDF text extraction is deterministic — it produces ground-truth slide content without relying on vision models. This prevents misidentification of embedded screenshots or demo captures as actual slide content, a common failure mode when using only image-based slide analysis.

Prerequisites

Poppler utilities must be installed (provides the pdftotext command):

  • macOS: brew install poppler
  • Ubuntu: apt-get install poppler-utils

© pamelafox, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .agents/skills/extract-slide-text of pamelafox/presentation-skills.

  • SKILL.md
  • extract_slide_text.py

Open the folder on GitHubat commit 2b809b3

Compare with similar skills

Extract Slide Text next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Extract Slide Text compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Extract Slide Text this skillpamelafox/presentation-skills125—~440Automated safety check: PassMIT
Paper Deckzsyggg/paper-craft-skills1.3k—~1.4kAutomated safety check: PassNone
Paper2slidesQuZhan51496/paper2anything469—~3.8kAutomated safety check: NotesApache-2.0
Li CarouselJakeschincariol/linkedin-agent-skill1.6k—~715Automated safety check: PassMIT
Ky Markdown RebuilderKyrieCheungYep/ky-markdown-rebuilder117—~5.7kAutomated safety check: PassNone
PDF ReadingWide-Moat/open-computer-use1261 repos~2.7kAutomated safety check: PassProprietary

Similar skills

  • Paper Deck

    zsyggg/paper-craft-skills

    将论文、技术文章或知识内容制作成高真实感的 AIGC 幻灯片。先做叙事结构和逐页视觉导演,再调用生图模型生成每一页 16:9 slide image,最后合成为 PPTX/PDF。适合论文汇报、组会、公开课、技术分享、商业化研究展示;当用户提到“论文PPT”“AI生成PPT”“不像AI的PPT”“高质感幻灯片”“逐页生图PPT”时使用。

    1.3k GitHub stars~1.4k tokensUpdated 4 mo ago
    Documents & OfficeAuto-check passed
  • Paper2slides

    QuZhan51496/paper2anything

    Turn an academic paper PDF into a presentation deck (.pptx) end-to-end.

    469 GitHub stars~3.8k tokensUpdated 2 mo ago
    Documents & OfficeAuto-check: notes
  • Li Carousel

    Jakeschincariol/linkedin-agent-skill

    Build a LinkedIn document post (carousel) - slide-by-slide copy, the cover that earns the swipe, and the PDF to upload.

    1.6k GitHub stars~715 tokensUpdated 22 days ago
    Documents & OfficeAuto-check passed
  • Ky Markdown Rebuilder

    KyrieCheungYep/ky-markdown-rebuilder

    Rebuild visual documents into reliable Markdown by combining text extraction with page or screenshot alignment.

    117 GitHub stars~5.7k tokensUpdated 3 mo ago
    Documents & OfficeAuto-check passed
  • PDF Reading

    Wide-Moat/open-computer-use

    A skill your agent uses when you need to read, inspect, or extract content from PDF files — especially when file content is NOT in your context and you need to read it from disk.

    126 GitHub starsUsed in 1 repo~2.7k tokens
    Documents & OfficeAuto-check passed
  • Sci HTML

    ShZhao27208/Aut_Sci_Write

    Generate academic presentation-style HTML slide decks and browser reports from PDFs, structured text, Markdown, paper summaries, outlines, or research notes.

    209 GitHub stars~1.4k tokensUpdated 1 mo ago
    Documents & OfficeAuto-check: notes

More from pamelafox/presentation-skills

All 14 skills in this repo
  • Generate Images Mai

    pamelafox/presentation-skills

    Generate or edit bitmap images with Microsoft MAI-Image-2.5 through the Azure AI image APIs.

    125 GitHub stars~1.5k tokensUpdated 1 mo ago
    Auto-check: notes
  • Make Revealjs Presentation

    pamelafox/presentation-skills

    Create or update a RevealJS HTML presentation using the repository's bundled slide template.

    125 GitHub stars~1.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Capture Video Frames

    pamelafox/presentation-skills

    Capture frames from a YouTube video at a regular interval, produce a manifest mapping filenames to timestamps, and describe each frame with an LLM.

    125 GitHub stars~1.7k tokensUpdated 1 mo ago
    Auto-check passed
  • Fetch Slides

    pamelafox/presentation-skills

    Fetch presentation slides from a URL and convert them to PDF.

    125 GitHub stars~528 tokensUpdated 1 mo ago
    Auto-check passed
  • Generate Writeup

    pamelafox/presentation-skills

    Generate an annotated blog-style write-up from a presentation's slides and video recording.

    125 GitHub stars~2.2k tokensUpdated 1 mo ago
    Auto-check passed
  • Outline Slides

    pamelafox/presentation-skills

    Generate a numbered outline of presentation slides with one-sentence summaries.

    125 GitHub stars~462 tokensUpdated 1 mo ago
    Auto-check passed

Questions about Extract Slide Text

What does Extract Slide Text do?

Extract text from each page of a PDF into a markdown file using pdftotext. Extract Slide Text is an agent skill from pamelafox/presentation-skills. Extract text from each page of a PDF into a markdown file using pdftotext.

When should I use Extract Slide Text?

Extract Slide Text fits situations like: : extract text from PDF slides; get slide text content; PDF to markdown text.

How do I install Extract Slide Text in Claude Code?

Run `npx skills add pamelafox/presentation-skills --skill extract-slide-text -a claude-code`. Or copy the skill folder (.agents/skills/extract-slide-text in pamelafox/presentation-skills) into .claude/skills/extract-slide-text in your project. Claude Code loads it when a task matches its description.

How do I install Extract Slide Text in Codex?

Run `npx skills add pamelafox/presentation-skills --skill extract-slide-text -a codex`. Or copy the skill folder (.agents/skills/extract-slide-text in pamelafox/presentation-skills) into .agents/skills/extract-slide-text in your project. Codex loads it when a task matches its description.

Can I use Extract Slide Text in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add pamelafox/presentation-skills --skill extract-slide-text -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/extract-slide-text, .gemini/skills/extract-slide-text, .github/skills/extract-slide-text and .opencode/skills/extract-slide-text in your project.

What does Extract Slide Text need to run?

Going by SKILL.md and its folder, Extract Slide Text needs Python for the scripts in its folder and the command-line tools its instructions call (uv, brew and apt-get). Our summary lists: Python 3.

Does Extract Slide Text access the network?

SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Extract Slide Text safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Extract Slide Text use?

Extract Slide Text is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Extract Slide Text use?

About 440 tokens (SKILL.md is roughly 1.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Extract Slide Text?

Skills that share tags, products or a category with Extract Slide Text: Paper Deck (zsyggg/paper-craft-skills, 1.3k stars), Paper2slides (QuZhan51496/paper2anything, 469 stars), Li Carousel (Jakeschincariol/linkedin-agent-skill, 1.6k stars) and Ky Markdown Rebuilder (KyrieCheungYep/ky-markdown-rebuilder, 117 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Extract Slide Text?

pamelafox (a GitHub user) maintains it in pamelafox/presentation-skills, which has 125 GitHub stars. The repository holds 14 skills in this directory. The repository was last updated on September 2, 2026.

Source: pamelafox/presentation-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.