Batch
asgeirtj/system_prompts_leaks
Research and plan a large-scale change, then execute it in parallel across 5–30 isolated worktree agents that each open a PR.
A skill your agent uses when extracting from many files at once with shared config, bounded parallelism, per-file overrides, and error recovery.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extraction --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .claude/skills/batch-extraction && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .claude/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extractionType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extraction --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .agents/skills/batch-extraction && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .agents/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extraction --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .cursor/skills/batch-extraction && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .cursor/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/hashgraph-online/awesome-codex-plugins.git --path plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extraction --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .gemini/skills/batch-extraction && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .gemini/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extractionInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .github/skills/batch-extraction && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .github/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins batch-extraction --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction .opencode/skills/batch-extraction && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "batch-extraction" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction into .opencode/skills/batch-extraction/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "batch-extraction", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
batch-extractionA skill your agent uses when extracting from many files at once with shared config, bounded parallelism, per-file overrides, and error recovery.
Batch Extraction is an agent skill from hashgraph-online/awesome-codex-plugins. Use when extracting from many files at once with shared config, bounded parallelism, per-file overrides, and error recovery. Covers the batch command, --file-configs, --max-concurrent, and output layout.
Its SKILL.md is about 1.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
The repository describes itself as: A curated list of awesome OpenAI Codex / ChatGPT plugins, skills, and resources. The 1 Codex Marketplace. See live plugins at: https://hol.org/plugins/best-codex-plugins. The licence is Apache-2.0.
Read from SKILL.md and the folder at commit 16b4156. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
jqFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Batch Extraction loads about 1.3k tokens when it runs. Until then it costs about 57 tokens; SKILL.md has 441 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from hashgraph-online/awesome-codex-plugins at commit 16b4156, republished under its Apache-2.0 licence (© hashgraph-online). 441 words, ~1,250 tokens.
.claude/skills/batch-extraction/SKILL.md (or your agent's skills folder).Use this when processing a directory or glob of documents in one pass.
kreuzberg batch shares one extraction config across every file, runs
extractions concurrently, and returns one structured array — failures on
individual files do not abort the run.
# Glob expands to many paths; results come back as a JSON array (default)
kreuzberg batch *.pdf
# Mixed formats, markdown content for LLM ingestion
kreuzberg batch docs/*.docx --content-format markdown
# Recurse with the shell, then extract
kreuzberg batch $(find ./corpus -name '*.pdf')batch defaults to --format json (vs --format text for single
extract). Each array entry is a full extraction result, so downstream
code can index by position into the input path list.
kreuzberg batch reports/*.pdf \
| jq '.[] | {chars: (.content | length), mime: .mime_type}'--max-concurrent caps how many files extract at once (default: CPU
count). Lower it on memory-constrained hosts or when OCR/ML models are
active, since each in-flight extraction holds its own buffers:
# Cap at 4 concurrent extractions
kreuzberg batch scans/*.pdf --ocr true --max-concurrent 4--max-threads additionally caps total internal threads (Rayon, ONNX
intra-op, the batch semaphore) for tightly constrained environments:
kreuzberg batch *.pdf --max-concurrent 2 --max-threads 4A single shared config does not always fit. --file-configs points at a
JSON file mapping each path to its own override object, merged on top of
the shared config for that file only:
{
"scan.pdf": { "force_ocr": true },
"report.pdf": { "output_format": "markdown" },
"data.xlsx": { "output_format": "json" }
}kreuzberg batch scan.pdf report.pdf data.xlsx --file-configs overrides.jsonKeys are file paths (matching the paths passed on the command line); values are per-file extraction config objects in snake_case, the same shape as a config file.
For text/toon output with image extraction, --output-dir controls where
referenced image files (e.g. image_0.png) are written; the directory
must already exist. JSON output embeds image bytes inline and ignores
--output-dir.
mkdir -p out/images
kreuzberg batch slides/*.pptx --extract-images true --output-dir out/images --format textBatch extraction is fault-tolerant per file: one unreadable or corrupt
document does not stop the rest. Inspect results for partial content and
surfaced errors rather than relying on the process exit code alone. Pair
with --max-concurrent to avoid exhausting memory when a few large files
sit in a big batch.
Every extract flag also applies to batch (OCR, chunking, layout,
content format, etc.) and is shared across all files unless a
--file-configs entry overrides it:
kreuzberg batch invoices/*.pdf \
--layout --layout-table-model slanet_wireless \
--content-format markdown --max-concurrent 8A config file works too and auto-discovers from the cwd upward:
output_format = "markdown"
[ocr]
backend = "tesseract"
language = "eng"kreuzberg batch corpus/*.pdf --config kreuzberg.tomlFrom Python, use the batch helpers (async and sync):
from kreuzberg import batch_extract_files, batch_extract_files_sync, ExtractionConfig
config = ExtractionConfig(output_format="markdown")
# Async
results = await batch_extract_files(["a.pdf", "b.docx", "c.xlsx"], config=config)
# Sync
results = batch_extract_files_sync(["a.pdf", "b.docx"], config=config)
for result in results:
print(len(result.content))Node.js mirrors this with batchExtractFiles; Rust uses
batch_extract_file (requires the tokio-runtime feature). See
references/python-api.md, references/nodejs-api.md, and
references/rust-api.md in the sibling kreuzberg skill.
When the kreuzberg MCP server is registered, prefer the
batch_extract_files tool over shelling out — it takes the file list and a
config object and returns structured results directly.
batch defaults to --format json,
extract to --format text. Set --format explicitly if a script
depends on one shape.--output-dir must exist — the CLI does not create it.--max-concurrent; the default is CPU count.--file-configs path keys — must match the paths as passed on the
command line, not absolute-resolved variants.See references/cli-reference.md for the full batch flag set.
© hashgraph-online, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction of hashgraph-online/awesome-codex-plugins.
Open the folder on GitHubat commit 16b4156
Batch Extraction next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Batch Extraction this skillhashgraph-online/awesome-codex-plugins | 1.2k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | |
| Batchasgeirtj/system_prompts_leaks | 69k | — | ~1.3k | Automated safety check: Pass | CC0-1.0 | |
| Batchcodewhale-hq/Codewhale | 41k | — | ~157 | Automated safety check: Pass | MIT | |
| Extractalirezarezvani/claude-skills | 28k | — | ~1.4k | Automated safety check: Pass | MIT | |
| Batch API PlannerQwenLM/qwen-code | 28k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | |
| Brand Extractnexu-io/open-design | 100k | — | ~3.1k | Automated safety check: Pass | Apache-2.0 |
asgeirtj/system_prompts_leaks
Research and plan a large-scale change, then execute it in parallel across 5–30 isolated worktree agents that each open a PR.
codewhale-hq/Codewhale
Break a large, parallelizable goal into bounded work units, coordinate existing agent/worktree machinery, integrate, and verify.
alirezarezvani/claude-skills
Turn a proven pattern or debugging solution into a standalone reusable skill with SKILL.md, reference docs, and examples.
QwenLM/qwen-code
Prepares many-file, single-turn transforms such as translating or rewriting as a plan, then submits it to the asynchronous, half-price DashScope Batch API through the qwen batch CLI.
nexu-io/open-design
Extract a complete Brand Kit from a live website by driving the in-app browser.
QwenLM/qwen-code
Runs one operation across many files with parallel worker agents: finds files by glob pattern, splits them into chunks, launches workers and summarizes the results.
hashgraph-online/awesome-codex-plugins
Create original anime-style reaction stickers as looping GIFs and MP4 previews, using generated character pose sheets and timed key poses.
hashgraph-online/awesome-codex-plugins
Manage and query Calibre libraries with the calibredb CLI (local paths or Calibre Content server URLs).
hashgraph-online/awesome-codex-plugins
A skill your agent uses when adding, changing, testing, or debugging Rust HTTP APIs and services, especially when Codex needs black-box integration tests, random-port app startup, real database test…
hashgraph-online/awesome-codex-plugins
Make a studio's game look like something at build time — a cover from a real frame of the game (free), painted covers, backdrops, textures and character plates from image models through the…
hashgraph-online/awesome-codex-plugins
Use CALL-E from Codex through the calle CLI. An agent skill from hashgraph-online/awesome-codex-plugins.
hashgraph-online/awesome-codex-plugins
Balance game difficulty, resources, rewards, probability, progression, economies, and dominant strategies.
A skill your agent uses when extracting from many files at once with shared config, bounded parallelism, per-file overrides, and error recovery. Batch Extraction is an agent skill from hashgraph-online/awesome-codex-plugins. Use when extracting from many files at once with shared config, bounded parallelism, per-file overrides, and error recovery.
Batch Extraction fits situations like: extracting from many files at once with shared config; bounded parallelism; per-file overrides.
Run `npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a claude-code`. Or copy the skill folder (plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction in hashgraph-online/awesome-codex-plugins) into .claude/skills/batch-extraction in your project. Claude Code loads it when a task matches its description.
Run `npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a codex`. Or copy the skill folder (plugins/kreuzberg-dev/plugins/plugins/kreuzberg/skills/batch-extraction in hashgraph-online/awesome-codex-plugins) into .agents/skills/batch-extraction in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add hashgraph-online/awesome-codex-plugins --skill batch-extraction -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/batch-extraction, .gemini/skills/batch-extraction, .github/skills/batch-extraction and .opencode/skills/batch-extraction in your project.
Going by SKILL.md and its folder, Batch Extraction needs the command-line tools its instructions call (jq). Our summary lists: Python 3; Node.js.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Batch Extraction is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.3k tokens (SKILL.md is roughly 5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Batch Extraction: Batch (asgeirtj/system_prompts_leaks, 69k stars), Batch (codewhale-hq/Codewhale, 41k stars), Extract (alirezarezvani/claude-skills, 28k stars) and Batch API Planner (QwenLM/qwen-code, 28k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
hashgraph-online (a GitHub organization) maintains it in hashgraph-online/awesome-codex-plugins, which has 1,232 GitHub stars. The repository holds 736 skills in this directory. The repository was last updated on October 6, 2026.
Source: hashgraph-online/awesome-codex-plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.