Agent skill

Backfill

by scaccogatto in scaccogatto/okf-skills

Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts).

MITAuto-check: notesDevelopment

Install Backfill

skills CLI
$ npx skills add scaccogatto/okf-skills --skill backfill -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install scaccogatto/okf-skills backfill --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/scaccogatto/okf-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/backfill .claude/skills/backfill && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
backfill
GitHub stars
408
Token cost
~4.3k tokens
SKILL.md length
1,951 words
Files
2 (incl. scripts)
Skills in repo
4
Repo updated
First seen
Licence
MIT

At a glance

Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts).

  • Works in 6 steps: Event schema → Protocol: Preflight → Extract →… → Skip rules → …
  • Creating an .okf/ bundle for an existing repository that predates this skill
  • SKILL.md covers 1. Event schema, 2. Protocol: Preflight →…, 3. Skip rules and 4. Anti-degeneration rules, plus 2 more sections
  • Runs Python scripts from its folder; calls uv, jq and git

What it does

Backfill is an agent skill from scaccogatto/okf-skills. Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts). Use when creating an .okf/ bundle for an existing repository that predates this skill, or when resuming an interrupted backfill session. Triggers on: "reconstruct the OKF bundle", "backfill the knowledge bundle", "event-source the history".

Its SKILL.md is about 4.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/okf_backfill_events.py`).

It sits in Development, covering Event-driven systems. It works with Git. The repository describes itself as: The OKF toolkit for Claude Code — author, maintain, validate & visualize Open Knowledge Format bundles. Plugin, agent skills, and a GitHub Action. The licence is MIT.

When your agent uses it

  • Creating an .okf/ bundle for an existing repository that predates this skill
  • Resuming an interrupted backfill session
  • : reconstruct the OKF bundle
  • Backfill the knowledge bundle

Example prompts

  • “reconstruct the OKF bundle”
  • “backfill the knowledge bundle”
  • “event-source the history”
  • “/backfill”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Bash, Read, Write, Edit

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Event schema
  2. Protocol: Preflight → Extract → Bootstrap → Replay (Map+Reduce) → Finalize
  3. Skip rules
  4. Anti-degeneration rules
  5. Implementation notes
  6. Capped diff emitter

What it can do on your machine

Read from SKILL.md and the folder at commit 8e31878. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash
    • Read
    • Write
    • Edit

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv
    • jq
    • git

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv and git, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Backfill loads about 4.3k tokens when it runs. Until then it costs about 90 tokens; SKILL.md has 1,951 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~90
When it runs · the whole SKILL.md, loaded when a task matches
~4.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash, Read, Write, Edit

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from scaccogatto/okf-skills at commit 8e31878, republished under its MIT licence (© scaccogatto). 1,951 words, ~4,254 tokens.

Download SKILL.mdSave it as .claude/skills/backfill/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
backfill
description
Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts). Use when creating an `.okf/` bundle for an existing repository that predates this skill, or when resuming an interrupted backfill session. Triggers on: "reconstruct the OKF bundle", "backfill the knowledge bundle", "event-source the history".
allowed-tools
Bash, Read, Write, Edit
user-invocable
true
argument-hint
[repo-dir] [--no-sessions] [--sessions-dir DIR]

Reconstruct an OKF bundle from history

This skill replays a repository's decision-making history (git commits and Claude session transcripts) to rebuild its OKF knowledge bundle as if the okf skill's Stop hook had been active from the start.

The extraction layer is deterministic (same repo → byte-identical events); the replay layer is an LLM loop, replayable and auditable but not byte-identical (timestamps, summaries change per run). See the event schema (§1) and skip rules (§3).

1. Event schema

Git commits and session turns become events (JSONL, one per line, sorted by timestamp then source):

json
{"id":"git:<sha>","source":"git","ts":"<ISO8601 Z>","sha":"...","author":"...","subject":"...","body":"...","files":[{"path":"...","add":N,"del":N}]}
{"id":"session:<file>:<lineno>","source":"session","ts":"<ISO8601 Z>","user":"...","outcome":"...","title":"...","branch":"...","skip":"..."}
  • Git events come from git log --first-parent --numstat.
  • Session events pair each user message with the last assistant text block of that turn (the wrap-up), extracted from ~/.claude/projects/<repo-slug>/*.jsonl and worktree subdirs.
  • Timestamps are normalized to UTC ISO 8601 strings ending in Z.
  • Skip field (optional, added by extraction): marks low-signal events — see §3, never overridden by replay.

2. Protocol: Preflight → Extract → Bootstrap → Replay (Map+Reduce) → Finalize

Preflight: check if bundle can be rebuilt
bash
if [ -d <repo>/.okf ] && ! [ -f <repo>/.okf/.backfill-state.json ]; then
  echo "ERROR: .okf/ exists but is incomplete. Delete it to rebuild from scratch, or pass --resume to continue from the last checkpoint."
  exit 1
fi

If .backfill-state.json exists, resume from the cursor; otherwise, fresh bootstrap.

Extract: generate event stream
bash
uv run "${CLAUDE_SKILL_DIR}/scripts/okf_backfill_events.py" <repo-dir> \
  --out events.jsonl \
  [--no-sessions] [--sessions-dir ~/.claude/projects] \
  [--max-text 2000] [--skip-globs "vendor/**"]

Write events.jsonl to the scratchpad (never committed). Report:

  • Total events extracted, per-source counts, and per-rule skip counts.
Bootstrap: initialize bundle (fresh only)

Create .okf/ with:

bash
mkdir -p <repo>/.okf
echo 'okf_version: "0.2"' > <repo>/.okf/index.md
echo '# Update Log' > <repo>/.okf/log.md

Cursor (.okf/.backfill-state.json):

json
{"last_id": null, "done": 0}
Replay: two-phase map/reduce protocol

The replay is now structured in two phases to enforce semantic depth and anti-degeneration rules:

Phase 1: Map (parallel analysis, per-event)

Launch okf:event-analyzer agents for the live events (those without skip field), in waves of 4 to 10 (see the cost note in §5), via Claude Code's Workflow agentType parameter. Each analyzer receives its event ids, the events.jsonl path, the repo path, the output directory and the emitter path, and:

  • Fetches its own event JSON with jq (the orchestrator dispatches ids, never content)
  • Reads git evidence only through the capped diff emitter (§6); session events need no other call
  • Writes analyses/<event-id>.md (event id with : sanitized to -) to scratchpad, with a truncated flag copied from the emitter's last line
  • Replies with one line of counts per event; the analysis never travels back in the reply
  • Never touches .okf/
  • Is resumable: skip events already in analyses/

Dispatch rule (deterministic, from events.jsonl, no content read): a live event is small when it is a session turn or a commit whose numstat totals at most 60 changed lines. Small events are grouped chronologically, eight per analyzer call; large commits go one per call. Measured on this repository (benchmark/map-tier/RESULTS.md): 57 calls instead of 147, a third off the map phase at the same tier, no loss on any per-event metric and no cross-event contamination.

bash
jq -c 'select(.skip==null) | {id, small: (.source=="session" or (([.files[]?|.add+.del]|add)//0) <= 60)}' events.jsonl

Host-agnostic fallback (if no Workflow support): spawn generic subagents with the same system prompt as agents/event-analyzer.md, one per event, collecting analyses to scratchpad.

Phase 2: Reduce (sequential folding, one agent)

Run a single okf:bundle-weaver agent that:

  • Reads all analyses in chronological order
  • Folds them into .okf/, updating or creating concepts
  • Enforces anti-degeneration rules (§4)
  • Adds a dated log.md bullet for each lifecycle event only (a concept created, deprecated or superseded); the "why" of a routine update goes into the concept body
  • Manages cursor (.okf/.backfill-state.json) for resume capability
  • Is the only actor that writes .okf/
  • Replies with one line of counts (folded, created, updated, bullets, conflicts, superseded, truncated inputs); the bundle never travels back in the reply

Resume behavior: both phases support resumption. Phase 1 skips already-analyzed event ids; Phase 2 restarts from last_id in the cursor.

Domain priming (advisory)

Before starting the map phase, read the repository's README and directory structure to sketch a candidate taxonomy of concepts (skills, integrations, decisions, etc.). This priming is advisory only — it helps the analyzer emit better candidate names. The guarantee that no events are lost comes from the deterministic coverage check in Finalize (§5), not from priming; analyses without candidate names still flow to the weaver.

Context hygiene (orchestrator)

The orchestrator handles ids and counts, never content. Events, analyses and diffs are read by the agents; the orchestrator reads:

bash
wc -l events.jsonl                                   # total events
jq -r 'select(.skip==null) | .id' events.jsonl       # live ids to dispatch
ls analyses | wc -l                                  # analyses written

Never cat events.jsonl or open an analysis from the orchestrator: whatever enters its context is billed at the frontier rate on every following turn, and the routing exists to prevent that.

Finalize: validate and clean up
  1. Write index.md per directory by hand (do NOT run okf_init.py — it would scaffold placeholder files): one # Section per directory, one bullet per concept, * [Title](file.md) - description taken from each concept's frontmatter (SPEC §8). The root index.md keeps its okf_version: "0.2" frontmatter and links every subdirectory.

  2. Add final log entry (today's date, appended to existing dated section if present):

    ## 2026-09-01
    
    - **Backfill**: reconstructed from N events by okf-backfill/0.9.5
  3. Delete cursor:

    bash
    rm -f .okf/.backfill-state.json
  4. Run self-checks for anti-degeneration rules:

    bash
    # No concept filenames derived from change type
    ! find .okf -name "*.md" | grep -E "merge-pull-request|(^|/)(feat|fix|chore|docs)[:-]"
    
    # No non-kebab-case characters in filenames
    ! find .okf -name "*.md" | grep -E "[^a-z0-9/.-]"
    
    # No identical consecutive log bullets
    awk 'p==$0 && /^- / {exit 1} {p=$0}' .okf/log.md
    
    # Every concept declares its lifecycle (absent status reads as `stable`, SPEC §5.4)
    [ -z "$(grep -rL '^status:' --include='*.md' .okf | grep -v '/index.md$\|/log.md$')" ]

    These are lexical guards; they do not catch a bundle that asserts a superseded state. That check is semantic and belongs to the weaver, at fold time, when it has both the analysis and the earlier concepts on disk (§4 rule 5).

  5. Run coverage check to guarantee every event is mapped:

    bash
    uv run "${CLAUDE_SKILL_DIR}/scripts/okf_backfill_events.py" \
      --check-coverage events.jsonl .okf

    Exit code 0 means all live events (non-skipped) appear in the bundle's sources or log. Exit code 1 with a list of unmapped ids means the weaver skipped some events; fix and re-run.

  6. Validate bundle schema:

    bash
    uv run "${CLAUDE_SKILL_DIR}/../validate/scripts/okf_validate.py" .okf --strict

    Fix every error before finishing.

  7. Report the resulting bundle:

    • Concept count per directory
    • Log entry samples (first and last)
    • Coverage check result (all events mapped)
    • Validation result (pass/fail, warnings)
    • Superseded resolutions (the weaver's superseded count) and deprecated concepts: grep -rl '^status: deprecated' --include='*.md' .okf | wc -l
    • Agents spawned per phase (analyzers, weaver invocations) and, when the host reports it, tokens per phase
    • Truncated analyses: grep -l '^truncated: true' analyses/*.md | wc -l, next to the deterministic estimate of git events whose full diff exceeds the default cap: jq -r 'select(.skip==null and .source=="git") | [.files[]|.add+.del] | add' events.jsonl | awk '$1>300' | wc -l

3. Skip rules

Events matching a rule below are marked skip: <rule-id> and skipped during replay (cursor advances, no log entry):

RuleConditionRationale
paths-only-generatedGit commit touches ONLY lockfiles or generated code (vendor/, node_modules/, dist/, .min., *.lock, etc.)Noise: build byproducts, no knowledge content
merge-no-filesGit merge commit with no file changes (first-parent only)Housekeeping
session-command-noiseSession user message is a slash command or <20 chars and has no useful outcomeEphemeral UI/chatter

Extend the first rule with --skip-globs "extra/**" for repo-specific patterns.

Show full SKILL.md (929 more words)Show less

4. Anti-degeneration rules

The analyzer and weaver enforce these rules to prevent the bundle from devolving into a mechanical listing of commits or a taxonomy-by-accident:

  1. Concept names describe domain entities, not changes. Forbidden:

    • Filenames derived from git subject: merge-pull-request-#2-....md, feat:-add-docs-skill.md, fix:-handle-edge-case.md
    • Bare action words: update.md, fix.md, add-feature.md
    • Non-kebab-case: spaces, underscores, capitals, special characters (allowed: [a-z0-9-] only)

    Example: a commit with subject "feat: add presales pipeline" touches a domain concept. The concept is named presales-pipeline.md (the entity), not feat:-add-presales-pipeline.md (the change). The change history lives in frontmatter sources and the body's ## History; log.md holds only lifecycle events.

  2. Prefer update over create. Every event is analyzed for which concepts it touches; if a concept already exists and the analysis fits, update its sources and body. Only create a new concept if the analysis reveals a distinct domain entity not yet captured.

  3. Log bullets are lifecycle events, and explain intent. Only a concept created, deprecated or superseded gets a bullet; a routine update does not. A bullet answers "why?" from the commit body or session outcome, never just re-reads the subject:

    • Bad: - Added presales-pipeline.md feature
    • Good: - **Creation**: [Presales pipeline](...), formalized to clarify handoff points
  4. No identical consecutive bullets. A run of same-concept bullets means updates leaked into the log:

    - Feature X update
    - Feature X update
    - Feature X update

    Instead: no bullet at all; the updates live in the concept's body and sources.

  5. A reversal is not growth. A replay walks the history forward, so a later event routinely overturns an earlier one (a phase dropped, a limit changed, a target downgraded). "Prefer update over create" is a merge rule and says nothing about this: applied alone it leaves the bundle asserting both states in the present tense, in one concept or in two. The weaver greps the whole bundle before writing and, on contradiction, replaces rather than appends: the body states what holds today, the previous position drops to a dated line under ## History, and an outdated concept that a newer one supersedes gets status: deprecated plus a link to its successor. A deprecated concept keeps its sources, so coverage stays green. None of the guards in Finalize can catch this — they are lexical, and a superseded claim is well-formed.

  6. Reconstructed concepts are draft, never stable. SPEC §5.4 reads an absent status as stable ("ready for consumption"); a bundle nobody has reread is draft ("not yet reviewed; possibly incomplete"). The weaver writes status: on every concept.

5. Implementation notes

  • Determinism: extraction is byte-identical; replay is not (time, LLM variance). Both are auditable: events.jsonl is deterministic, map analyses are stored and resumable, reduce loop is human-readable and cursor-backed.
  • Privacy: events.jsonl goes to scratchpad, never committed. Session turn text is truncated (head+tail) before extraction.
  • Cost note: deep replay reads one capped diff per git commit (§6) to extract rationale. For histories >500 events, consider splitting into sub-ranges and replaying sequentially, or use --skip-globs to exclude low-signal paths (e.g., vendored dependencies, generated code). Map phase is parallel in waves of 4 to 10: the gate benchmark lost 260 of 299 runs to a high-concurrency mass failure and finished at concurrency 4 (benchmark/gate/RESULTS.md). Reduce is sequential but much cheaper (concepts already analyzed).
  • Future sources (not implemented): GitHub PR/issue text, release notes, CI/deploy logs — all optional post-MVP. Lore protocol (arXiv 2603.15566): on repos that adopt git trailers (structured decision metadata), the extracted event's body already contains trailers in a machine-readable format. The analyzer and weaver should use trailers as the primary source of "why" (overriding generic inference from diff content).
  • Trust metadata: generated.by is the backfill agent (okf-backfill/0.9.5), not claimed as human-reviewed (human:...); concepts are correctly unverified (SPEC §5.3) and status: draft (§5.4) — the two axes are independent: trust is who says it, lifecycle is whether it has been reread. Writing status is not optional here: absent, §5.4 reads it as stable, so a bundle nobody has reread would declare itself ready for consumption. Map-phase analyses are working artifacts (stored for auditability during reduce); the weaver's output is the canonical bundle.

6. Capped diff emitter

Analyzers never read a raw git show: the harness cuts long tool output blindly (mid-hunk, character-based, host-dependent) and a cheap worker may not notice the cut. The extractor reads the whole diff instead and emits a deterministic sample:

bash
uv run "${CLAUDE_SKILL_DIR}/scripts/okf_backfill_events.py" <repo-dir> --show <sha> \
  [--only <path>] [--max-diff-lines 300] [--per-file-lines 120] [--max-line-chars 400] \
  [--skip-globs "vendor/**"]
  • Diff against the first parent, so merge commits agree with the extracted numstat.
  • The stat is always complete: entities survive any cap.
  • Patches are capped per file (breadth over depth) and cut at hunk or file boundaries, with a bracketed marker at every cut; lines longer than --max-line-chars are shortened with a marker (a generated one-line file is "1 line" but can be hundreds of KB).
  • Patches of generated files (lockfiles, vendor/**, --skip-globs) are omitted even inside mixed commits; their stat line stays.
  • The last line is fixed-form: [diff: shown=X total=Y files_shown=A files_total=B truncated=true|false]. The analyzer copies truncated into its frontmatter; finalize reports the count.
  • --only <path> is the one permitted follow-up when a truncated diff hides the rationale: same cap logic, one file.

Defaults keep one call under the harness limits with margin; tune per repo with the flags.

The analyzer's tier is configuration, not infrastructure: agents/event-analyzer.md carries model and effort in its frontmatter. Fork the file to retier the map phase; the skill resolves the agent by name and nothing else changes. The default is haiku since the map-tier benchmark (benchmark/map-tier/RESULTS.md): parity with sonnet on commits at under half the map cost, on the condition that the two explicit instructions in the agent file stay (which summary line the flag copies; claims report intent as intent).

© scaccogatto, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/backfill of scaccogatto/okf-skills.

  • SKILL.md
  • scripts/okf_backfill_events.py

Open the folder on GitHubat commit 8e31878

Compare with similar skills

Backfill next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Backfill compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Backfill this skillscaccogatto/okf-skills408—~4.3kAutomated safety check: NotesMIT
Evolutionary Modular Architecturetech-leads-club/agent-skills7k—~3.7kAutomated safety check: PassCC-BY-4.0
Deskcomm Contribuirmelgarafael/DeskcommCRM4.5k—~3.6kAutomated safety check: NotesMIT
Designing ArchitectureCloudAI-X/claude-workflow-v21.4k1 repos~1.4kAutomated safety check: PassMIT
Migration PathAxonIQ/AxonFramework3.6k—~594Automated safety check: PassApache-2.0
Bktavivsinai/bitbucket-cli2321 repos~1kAutomated safety check: PassMIT

Similar skills

  • Evolutionary Modular Architecture

    tech-leads-club/agent-skills

    Guides design of modular-monolith platforms with DDD, flat-by-aggregate modules, anti-corruption layers, outbox events and resilience, plus an architecture document with SVG diagrams.

    7k GitHub stars~3.7k tokensUpdated 18 days ago
    DevelopmentAuto-check passed
  • Deskcomm Contribuir

    melgarafael/DeskcommCRM

    Guia de contribuição ao DeskcommCRM para quem vai mexer no código e abrir um pull request, sobretudo de um fork.

    4.5k GitHub stars~3.6k tokensUpdated today
    DevelopmentAuto-check: notes
  • Designing Architecture

    CloudAI-X/claude-workflow-v2

    Designs software architecture and selects appropriate patterns for projects.

    1.4k GitHub starsUsed in 1 repo~1.4k tokens
    DevelopmentAuto-check passed
  • Migration Path

    AxonIQ/AxonFramework

    Create a new Axon Framework 4→5 migration path documentation page.

    3.6k GitHub stars~594 tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Bkt

    avivsinai/bitbucket-cli

    Operate Bitbucket Cloud or Data Center repositories, pull requests, branches, issues, pipelines, permissions, and webhooks with bkt.

    232 GitHub starsUsed in 1 repo~1k tokens
    DevelopmentAuto-check passed
  • Update Dependencies

    alorence/django-modern-rpc

    Routine update of all project dependencies — uv itself, uv.lock (all groups), tool versions pinned in GitHub workflows and .pre-commit-config.yaml (uv, ruff, mypy...), and SHA-pinned GitHub Actions.

    111 GitHub stars~1.3k tokensUpdated yesterday
    DevelopmentAuto-check passed

More from scaccogatto/okf-skills

  • Okf

    scaccogatto/okf-skills

    Author, maintain, and consume Open Knowledge Format (OKF) knowledge bundles — portable markdown + YAML frontmatter that both humans and agents read.

    408 GitHub stars~2.1k tokensUpdated 10 days ago
    Auto-check: notes
  • Validate

    scaccogatto/okf-skills

    Check that an Open Knowledge Format (OKF) bundle is conformant with the v0.2 spec (§11).

    408 GitHub stars~941 tokensUpdated 10 days ago
    Auto-check: notes
  • Visualize

    scaccogatto/okf-skills

    Render an Open Knowledge Format (OKF) bundle as a single self-contained, interactive HTML graph (viz.html) — concepts as nodes coloured by type and sized by body length, markdown links and…

    408 GitHub stars~714 tokensUpdated 10 days ago
    Auto-check: notes

Works with

Questions about Backfill

What does Backfill do?

Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts). Backfill is an agent skill from scaccogatto/okf-skills. Reconstruct an OKF bundle by event-sourcing a repository's history (git log and Claude session transcripts).

When should I use Backfill?

Backfill fits situations like: creating an .okf/ bundle for an existing repository that predates this skill; resuming an interrupted backfill session; : reconstruct the OKF bundle; backfill the knowledge bundle.

How do I install Backfill in Claude Code?

Run `npx skills add scaccogatto/okf-skills --skill backfill -a claude-code`. Or copy the skill folder (skills/backfill in scaccogatto/okf-skills) into .claude/skills/backfill in your project. Claude Code loads it when a task matches its description.

How do I install Backfill in Codex?

Run `npx skills add scaccogatto/okf-skills --skill backfill -a codex`. Or copy the skill folder (skills/backfill in scaccogatto/okf-skills) into .agents/skills/backfill in your project. Codex loads it when a task matches its description.

Can I use Backfill in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add scaccogatto/okf-skills --skill backfill -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/backfill, .gemini/skills/backfill, .github/skills/backfill and .opencode/skills/backfill in your project.

What does Backfill need to run?

Going by SKILL.md and its folder, Backfill needs Python for the scripts in its folder and the command-line tools its instructions call (uv, jq and git). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Bash, Read, Write, Edit.

Does Backfill access the network?

SKILL.md contains no URLs. Its commands use uv and git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Backfill safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Backfill use?

Backfill is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Backfill use?

About 4.3k tokens (SKILL.md is roughly 17k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Backfill?

Skills that share tags, products or a category with Backfill: Evolutionary Modular Architecture (tech-leads-club/agent-skills, 7k stars), Deskcomm Contribuir (melgarafael/DeskcommCRM, 4.5k stars), Designing Architecture (CloudAI-X/claude-workflow-v2, 1.4k stars) and Migration Path (AxonIQ/AxonFramework, 3.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Backfill?

scaccogatto (a GitHub user) maintains it in scaccogatto/okf-skills, which has 408 GitHub stars. The repository holds 4 skills in this directory. The repository was last updated on September 28, 2026.

Source: scaccogatto/okf-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.