Tmux
trpc-group/trpc-agent-go
Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.
Move hardcoded lookup/config constants (gender maps, header dicts, value translations, column-label maps, date formats) out of a crawler and into the dataset .yml — as datapatch lookups wherever…
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install opensanctions/opensanctions crawler-constants-to-yml --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .claude/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .claude/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-ymlType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install opensanctions/opensanctions crawler-constants-to-yml --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .agents/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .agents/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install opensanctions/opensanctions crawler-constants-to-yml --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .cursor/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .cursor/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/opensanctions/opensanctions.git --path .claude/skills/crawler-constants-to-yml--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install opensanctions/opensanctions crawler-constants-to-yml --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .gemini/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .gemini/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install opensanctions/opensanctions crawler-constants-to-ymlInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .github/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .github/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install opensanctions/opensanctions crawler-constants-to-yml --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/opensanctions/opensanctions.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/crawler-constants-to-yml .opencode/skills/crawler-constants-to-yml && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "crawler-constants-to-yml" agent skill from https://github.com/opensanctions/opensanctions/tree/main/.claude/skills/crawler-constants-to-yml into .opencode/skills/crawler-constants-to-yml/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "crawler-constants-to-yml", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
crawler-constants-to-ymlMove hardcoded lookup/config constants (gender maps, header dicts, value translations, column-label maps, date formats) out of a crawler and into the dataset .yml — as datapatch lookups wherever…
Crawler Constants To Yml is an agent skill from opensanctions/opensanctions. Move hardcoded lookup/config constants (gender maps, header dicts, value translations, column-label maps, date formats) out of a crawler and into the dataset .yml — as datapatch lookups wherever possible, otherwise http / config / dates metadata. Use when asked to move, hoist, or "put in the yml" a crawler constant such as GENDERS or HEADERS.
Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Data & Analytics, covering Web scraping and Translation. The repository describes itself as: An open database of international sanctions data, persons of interest and politically exposed persons. The licence is MIT.
3 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 123b662. It shows what the files ask for, not the result of running them.
Pre-approves these tools, so the agent can use them without asking each time:
ReadEditBashGrepGlobFrom allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
ruffmypypython3From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Crawler Constants To Yml loads about 1.4k tokens when it runs. Until then it costs about 92 tokens; SKILL.md has 563 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
allowed-tools: Read, Edit, Bash, Grep, GlobAutomated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from opensanctions/opensanctions at commit 123b662, republished under its MIT licence (© opensanctions). 563 words, ~1,391 tokens.
.claude/skills/crawler-constants-to-yml/SKILL.md (or your agent's skills folder).Hoist a module-level constant out of a crawler and into its .yml. The guiding rule:
whatever the framework can handle from the YAML should live in the YAML — a lookup,
http, config, or dates block — leaving the crawler code free of one-off maps and
conditionals. Target: $ARGUMENTS (the .yml, or the sibling crawler.py).
Read the relevant zavod/docs before editing — do not re-derive the mechanics here:
zavod/docs/best_practices/datapatch_lookups.mdhttp.user_agent: zavod/docs/best_practices/http_operations.mdzavod/docs/best_practices/dates_meta.md| Constant in code | Belongs in | Mechanism |
|---|---|---|
| Map of source value → clean value for a typed property (gender, country, …) | lookups: type.<type> | Auto-applied on entity.add(prop, raw). Pass the raw value; delete the dict. |
| Map/translation of a value for a non-type property, or a categorical label | lookups: <name> | context.lookup_value("<name>", raw, warn_unmatched=True) |
Request User-Agent | http.user_agent | Applied to the session; drop headers= if it held only the UA. |
| Static structural/config dict (column labels, field maps used for validation) | config: | Read via context.dataset.config["key"] |
| Date formats / non-English month names | dates.formats, dates.months | Consumed by h.apply_date |
Type-property names map to their type.<type> per the table in datapatch_lookups.md
(e.g. gender → type.gender, citizenship/country → type.country, email → type.email).
If a constant genuinely cannot move (see gotchas), leave it in code — do not force it.
Type lookups (the common GENDERS case). Add the lookup and feed the raw value in:
lookups:
type.gender:
options:
- match: Masculino # exact source values, one per real value
value: male
- match: Femenino
value: femaleperson.add("gender", record["sexo"]) # was: GENDERS.get(record["sexo"])Delete the dict, its .get(...), and any or ""/None-guard it needed. Match the source
value exactly — vocabularies differ per source (Femenino vs Feminino, Hombre/Mujer
vs Masculino/Femenino); never copy a sibling dataset's lookup blind.
Named lookups. For a value that isn't a typed property (e.g. a seat/membership type), use an explicit lookup and store the result — if the old code fetched the value only to validate it, recording the mapped value is a strict improvement:
seat_type = context.lookup_value("membership_type", raw, warn_unmatched=True)
occupancy.add("summary", seat_type)Headers. Move only the User-Agent to http.user_agent. For any other header, first
prove whether the server actually needs it (§3) — drop inert ones, keep genuinely-required
ones inline (the http block supports only user_agent). A format-selecting Accept header
can sometimes be replaced by a URL suffix/param in data.url (e.g. …/atual.json),
removing the header entirely.
ruff check <crawler.py> && mypy --strict <crawler.py>
zavod crawl <dataset.yml>
# issues.log must be clean (or only carry warnings you expect):
python3 -c "import json;[print(json.loads(l)['level'],json.loads(l).get('message')) for l in open('data/datasets/<name>/issues.log')]" | sort | uniq -c
# confirm the value still lands, at the expected coverage:
grep -a ',<Schema>:<prop>,' data/datasets/<name>/statements.pack | awk -F, '{print $3}' | sort | uniq -cWhen probing which headers a server requires, drop them one at a time against the live endpoint and keep only those whose removal changes the status/response.
DICT.get(x)
returned None and silently dropped unmapped values; the lookup instead lets unmapped
values reach the type cleaner, which warns. After the change, confirm coverage in the
statement count and that no new Rejected property value warnings appear.config:. Making correctness depend on YAML key order is a
trap for the next editor.config: is for static config, not value cleaning. Reach for a lookups: block when
you're normalizing property values; use config: for parsing/structure knobs.Accept, Content-Type, Referer, etc. have no
http field — required ones stay as an inline headers= dict on the fetch call.© opensanctions, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .claude/skills/crawler-constants-to-yml of opensanctions/opensanctions.
Open the folder on GitHubat commit 123b662
Crawler Constants To Yml next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Crawler Constants To Yml this skillopensanctions/opensanctions | 834 | — | ~1.4k | Automated safety check: Notes | MIT | |
| Tmuxtrpc-group/trpc-agent-go | 1.9k | 23 repos | ~868 | Automated safety check: Pass | Apache-2.0 | |
| Ketch1broseidon/ketch | 702 | 1 repos | ~3.9k | Automated safety check: Pass | MIT | |
| Crawl4AI Web Scrapingsmallnest/goclaw | 599 | 1 repos | ~2.5k | Automated safety check: Pass | MIT | |
| Boss Zhipin Scrapereatmoreduck/boss-zhipin-scraper | 1.5k | — | ~2.6k | Automated safety check: Pass | MIT | |
| Axyusukebe/ax | 719 | 1 repos | ~918 | Automated safety check: Pass | MIT |
trpc-group/trpc-agent-go
Remote-control tmux sessions for interactive CLIs by sending keystrokes and scraping pane output.
1broseidon/ketch
Research skill for ketch — a fast stateless CLI for web search, OSS code search, curated library docs, page scraping, and site crawling; an optional MCP server exists for operators who want it, but…
smallnest/goclaw
Scrapes sites, handles JavaScript-heavy pages and extracts structured data with Crawl4AI, through its crwl CLI or Python SDK, including schema-based extraction without an LLM.
eatmoreduck/boss-zhipin-scraper
Scrape BOSS直聘 (job listing site) via Chrome CDP. An agent skill from eatmoreduck/boss-zhipin-scraper.
yusukebe/ax
Use the ax CLI instead of curl + throwaway parsing scripts whenever you fetch a URL, explore an unknown web page, or extract structured data from HTML.
Anakin-Inc/anakin
Scrape any website into clean markdown or structured JSON. An agent skill from Anakin-Inc/anakin.
opensanctions/opensanctions
Scaffold a new PEP (Politically Exposed Persons) crawler — members of a parliament, legislature, senate, chamber of deputies, cabinet, judiciary, or an asset-declaration register — from a source URL…
opensanctions/opensanctions
Bring a dataset .yml's metadata in line with house conventions (title, summary, description, coverage, publisher, maintainer comments).
opensanctions/opensanctions
Refactor the title, description and coverage frequency of a legislature/parliament PEP dataset .yml into the house style.
opensanctions/opensanctions
Migrate ad-hoc name cleaning in a crawler to h.reviewnames (Step 1 of the name framework migration).
opensanctions/opensanctions
Rewrite messy or AI-generated crawler code into clean, production-ready style that follows the zavod best practices.
opensanctions/opensanctions
Release one or more datasets by adding them to a topical collection, bumping coverage.start, and verifying.
Categories
Move hardcoded lookup/config constants (gender maps, header dicts, value translations, column-label maps, date formats) out of a crawler and into the dataset .yml — as datapatch lookups wherever…. Crawler Constants To Yml is an agent skill from opensanctions/opensanctions.yml — as datapatch lookups wherever possible, otherwise http / config / dates metadata.
Crawler Constants To Yml fits situations like: put in the yml a crawler constant such as GENDERS; tasks that involve Web scraping; tasks that involve Translation.
Run `npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a claude-code`. Or copy the skill folder (.claude/skills/crawler-constants-to-yml in opensanctions/opensanctions) into .claude/skills/crawler-constants-to-yml in your project. Claude Code loads it when a task matches its description.
Run `npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a codex`. Or copy the skill folder (.claude/skills/crawler-constants-to-yml in opensanctions/opensanctions) into .agents/skills/crawler-constants-to-yml in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add opensanctions/opensanctions --skill crawler-constants-to-yml -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/crawler-constants-to-yml, .gemini/skills/crawler-constants-to-yml, .github/skills/crawler-constants-to-yml and .opencode/skills/crawler-constants-to-yml in your project.
Going by SKILL.md and its folder, Crawler Constants To Yml needs the command-line tools its instructions call (ruff, mypy and python3). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Read, Edit, Bash, Grep, Glob.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.
Crawler Constants To Yml is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.4k tokens (SKILL.md is roughly 5.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Crawler Constants To Yml: Tmux (trpc-group/trpc-agent-go, 1.9k stars), Ketch (1broseidon/ketch, 702 stars), Crawl4AI Web Scraping (smallnest/goclaw, 599 stars) and Boss Zhipin Scraper (eatmoreduck/boss-zhipin-scraper, 1.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
opensanctions (a GitHub organization) maintains it in opensanctions/opensanctions, which has 834 GitHub stars. The repository holds 11 skills in this directory. The repository was last updated on October 9, 2026.
Source: opensanctions/opensanctions on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.