Cluster keywords into pillar and spoke pages by SERP overlap via script.

MITAuto-check passedMarketing & SEO

Install Keyword Cluster

skills CLI
$ npx skills add indranilbanerjee/digital-marketing-pro --skill keyword-cluster -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install indranilbanerjee/digital-marketing-pro keyword-cluster --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/indranilbanerjee/digital-marketing-pro.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/keyword-cluster .claude/skills/keyword-cluster && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
keyword-cluster
GitHub stars
862
Used in
1 other repo
Token cost
~2.5k tokens
SKILL.md length
1,070 words
Files
1
Skills in repo
162
Repo updated
First seen
Licence
MIT

At a glance

Cluster keywords into pillar and spoke pages by SERP overlap via script.

  • Works in 4 steps: Read… → If no brand exists: ask "Set up a brand… → Apply industry-specific guidance from… → …
  • Tasks that involve Keyword research
  • SKILL.md covers Purpose, Context efficiency, When to Use and Brand context (auto-applied), plus 9 more sections
  • Calls python

What it does

Keyword Cluster is an agent skill from indranilbanerjee/digital-marketing-pro. Cluster keywords into pillar and spoke pages by SERP overlap via script. "cluster these keywords"

Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Marketing & SEO, covering Keyword research. The repository describes itself as: An open-source AI marketing operating system for strategy, SEO, AEO/GEO, paid media, content, CRM, and analytics - grounded in brand context, human approval, and verifiable… The licence is MIT.

When your agent uses it

  • Tasks that involve Keyword research

Example prompts

  • “cluster these keywords”
  • “/keyword-cluster”

Requirements

  • Python 3

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Read ~/.claude-marketing/brands/_active-brand.json for the active slug, then load ~/.claude-marketing/brands/{slug}/profile.json
  2. If no brand exists: ask "Set up a brand first (/digital-marketing-pro:brand-setup)?" — or proceed with defaults
  3. Apply industry-specific guidance from skills/context-engine/industry-profiles.md
  4. Apply skills/context-engine/compliance-rules.md to filter out banned terminology before clustering

What it can do on your machine

Read from SKILL.md and the folder at commit 9e949f3. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Keyword Cluster loads about 2.5k tokens when it runs. Until then it costs about 28 tokens; SKILL.md has 1,070 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~28
When it runs · the whole SKILL.md, loaded when a task matches
~2.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from indranilbanerjee/digital-marketing-pro at commit 9e949f3, republished under its MIT licence (© indranilbanerjee). 1,070 words, ~2,502 tokens.

Download SKILL.mdSave it as .claude/skills/keyword-cluster/SKILL.md (or your agent's skills folder).
name
keyword-cluster
description
Cluster keywords into pillar and spoke pages by SERP overlap via script. "cluster these keywords"
argument-hint
<seed-keywords or path/to/seeds.csv> [target-country]
user-invocable
true

/digital-marketing-pro:keyword-cluster

Script location. If your host does not set ${CLAUDE_PLUGIN_ROOT}, the scripts are in this plugin's scripts/ folder, next to skills/.

Purpose

Take a set of seed keywords and produce a publication-ready cluster plan: pillar pages with their spokes, intent-grouped, prioritised by an opinionated scoring formula, with an internal-link map and a four-gate quality scorecard. Output is structured for direct hand-off to /digital-marketing-pro:content-brief or /digital-marketing-pro:content-engine.

Context efficiency

Heavy skill. Grep before Read any referenced file, then Read only matched ranges with offset + limit. List ${CLAUDE_PLUGIN_DATA}/<brand>/ before opening files. On re-invocation mid-session, skip files already in context.

When to Use

  • Onboarding a new content programme — turn a 20-keyword brief into a structured topical hub
  • Auditing an existing content library for cannibalisation (two pages competing for the same intent)
  • Designing a pillar+spokes architecture before any writing begins
  • Staging programmatic SEO across hundreds of variants (use this once per topic family)
  • Reorganising an existing site's internal-link graph

Don't use when you just need keyword expansion (use /digital-marketing-pro:keyword-research) or when you need ranking / SERP-feature analysis (use /digital-marketing-pro:rank-monitor, with --features for SERP features).

Brand context (auto-applied)

  1. Read ~/.claude-marketing/brands/_active-brand.json for the active slug, then load ~/.claude-marketing/brands/{slug}/profile.json
  2. If no brand exists: ask "Set up a brand first (/digital-marketing-pro:brand-setup)?" — or proceed with defaults
  3. Apply industry-specific guidance from skills/context-engine/industry-profiles.md
  4. Apply skills/context-engine/compliance-rules.md to filter out banned terminology before clustering

Inputs

InputSourceRequired?
Seed keywords (3–500)CSV with keyword column (optional: volume, kd, intent)yes
SERP results per keywordJSON: {keyword: [top result URLs]} from any rank-tracker / Ahrefs / Semrush exportstrongly recommended — without this the script falls back to lexical clustering, which is lower-confidence
Target country / languageFrom brand profileoptional override
Min volume / max KD filtersCLI flagsoptional
Overlap thresholdCLI flag --overlap (default 0.4 for SERP mode, 0.3 for lexical)optional

If SERPs JSON is unavailable, you can build one quickly by running the brand's connected rank-tracker MCP (Ahrefs / SE Ranking / Semrush) for each seed and saving the top 10 URLs. Skip this step only if the seeds are too numerous to justify the API spend — but flag the lower-confidence mode in the final deliverable.

Process (10 steps, numbered-file output)

All outputs go to ${CLAUDE_PLUGIN_DATA}/{brand}/seo/keyword-cluster/{YYYY-MM-DD}/.

  1. 00-input.md — capture seeds, source, filters, brand context, run timestamp
  2. 01-seed-expansion.md — if seeds < 20, expand via brand's keyword-research MCP (Ahrefs getRelatedKeywords, etc.) to ~50–200; otherwise skip. Document expansion source.
  3. 02-filtered.csv — apply min-volume / max-KD / banned-word filters. Save the filtered set as CSV (this is what the script consumes).
  4. 03-serps.json — fetch top-10 SERP URLs per keyword via the connected rank-tracker (skip if SERPs already provided). Budget guard: if estimated cost > 500 credits, surface the cost and ask "Continue? (y/N — default N)" before fetching.
  5. 04-cluster-run.json — run the script:
    bash
    python "${CLAUDE_PLUGIN_ROOT}/scripts/keyword_cluster.py" \
        --keywords "${CLAUDE_PLUGIN_DATA}/{brand}/seo/keyword-cluster/{date}/02-filtered.csv" \
        --serps "${CLAUDE_PLUGIN_DATA}/{brand}/seo/keyword-cluster/{date}/03-serps.json" \
        --overlap 0.4 \
        --min-volume {profile.min_volume or 0} \
        --max-kd {profile.max_kd or 100} \
        --out "${CLAUDE_PLUGIN_DATA}/{brand}/seo/keyword-cluster/{date}/04-cluster-run.json"
  6. 05-quality-scorecard.md — read the quality_scorecard block from 04-cluster-run.json. If status: needs_review, diagnose:
    • cannibalisation: fail → two clusters share pillar+intent. Merge them or reassign the lower-priority cluster's pillar.
    • orphan: fail → a multi-keyword cluster has 0 spokes. Re-tokenise its members or lower --overlap.
    • coverage: fail → < 80% of seeds clustered. Lower --overlap to 0.3 or expand seeds.
    • anchor_diversity: fail → pillar names too similar. Rewrite cluster names with synonym variation.
    • fragmentation_warning: true (pillar-only > 50%) → overlap threshold too strict. Try --overlap 0.3 first.
  7. 06-pillar-pages.md — for each cluster with priority_score >= 0.5, draft a one-paragraph pillar page brief (intent, audience, length target, key questions to answer). These feed /digital-marketing-pro:content-brief.
  8. 07-internal-link-map.md — table view of internal_link_targets from the script output. Per cluster: which other clusters to link out to + suggested anchor text. This is the file your dev team or CMS template should consume.
  9. 08-build-order.md — sorted by priority_score descending. Recommended build cadence: top 10% in Q1, next 30% in Q2, remainder backlog.
  10. PLAN.md — single-page summary: stats + scorecard + top 5 priority clusters + handoff to next skill in chain.

Output format

${CLAUDE_PLUGIN_DATA}/{brand}/seo/keyword-cluster/2026-06-04/
├── 00-input.md
├── 01-seed-expansion.md      (only if seeds expanded)
├── 02-filtered.csv
├── 03-serps.json             (if SERP mode)
├── 04-cluster-run.json       (raw script output)
├── 05-quality-scorecard.md
├── 06-pillar-pages.md
├── 07-internal-link-map.md
├── 08-build-order.md
└── PLAN.md                   (the deliverable)

PLAN.md is what you hand to the brand / client / next skill. Everything else is auditable intermediate state.

Show full SKILL.md (429 more words)Show less

Quality scorecard (the four gates)

Every run produces a scorecard from scripts/keyword_cluster.py. All four must pass for status: ready:

GateWhat it checksWhy it matters
cannibalisationNo two clusters share the same (pillar, primary_intent) pairPrevents you from writing two pages competing for the same SERP
orphanEvery multi-keyword cluster has ≥1 spoke (pillar-only clusters are exempt and tagged)Catches clustering bugs where a cluster head has no supporting topics
coverage≥ 80% of input seeds are assigned to at least one clusterCatches "junk" seeds and overly strict thresholds
anchor_diversityEach multi-keyword cluster has ≥ 2 anchor-text variants suggestedStops anchor-text over-optimisation across the internal-link graph

A fragmentation_warning: true (pillar-only > 50%) is a soft signal — the run is valid but you should consider lowering --overlap and re-running.

After the cluster

Ask: "Would you like me to:

  • Brief the top pillar pages? (/digital-marketing-pro:content-brief)
  • Start writing the highest-priority pillar? (/digital-marketing-pro:content-engine)
  • Apply the internal-link map to your CMS? (/digital-marketing-pro:seo-implement)
  • Schedule a quarterly re-run via /digital-marketing-pro:seo-drift?"

Chain handoffs

This skill is a producer in the chain:

  1. /digital-marketing-pro:keyword-research — generate seeds
  2. /digital-marketing-pro:keyword-cluster — this skill
  3. /digital-marketing-pro:content-brief — consumes PLAN.md + 06-pillar-pages.md to brief each pillar
  4. /digital-marketing-pro:content-engine — drafts the content
  5. /digital-marketing-pro:seo-implement — applies the internal-link map to the CMS

Tips & caveats

  • SERP mode is strictly better than lexical mode. Lexical clustering can't see that "shopify seo" and "ecommerce platform seo" target overlapping SERPs while "shopify themes" doesn't.
  • Overlap threshold defaults are conservative. If you get fragmentation_warning: true, lower to 0.3 first. If you get cannibalisation: fail with too few clusters, raise to 0.5.
  • The priority score isn't a ranking — it's a starting build order. A cluster with priority_score: 0.3 may still be your highest-conversion opportunity if it maps to a high-margin product line. Use the brand profile's business_goals to override mechanically.
  • Don't run this on raw GSC query exports without filtering first. GSC dumps thousands of long-tail variants of the same query — they'll all cluster together and produce a single mega-cluster.
  • Pillar-only clusters are valid — they represent distinct intents that simply lack spoke candidates in your seed set. Add seeds via Step 2 expansion if you want spokes.
  • The internal-link map is suggestions, not commands. Final anchor text should be reviewed for brand voice (apply the brand profile's voice fields + skills/context-engine/guidelines-framework.md).

Agents used

  • seo-specialist (primary) — interpretation + final pillar-page recommendations
  • competitive-intel — for SERP-overlap reasoning when results look surprising
  • brand-guardian — anchor-text review against banned-term lists

See also

  • /digital-marketing-pro:keyword-research — generates seeds (use first)
  • /digital-marketing-pro:content-brief — consumes the cluster plan (use next)
  • /digital-marketing-pro:seo-implement — applies internal-link map to CMS
  • /digital-marketing-pro:seo-drift — re-run quarterly to detect cluster drift
  • scripts/keyword_cluster.py — the underlying clustering engine

© indranilbanerjee, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/keyword-cluster of indranilbanerjee/digital-marketing-pro.

Open the folder on GitHubat commit 9e949f3

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in indranilbanerjee/digital-marketing-pro, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Keyword Cluster next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Keyword Cluster compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Keyword Cluster this skillindranilbanerjee/digital-marketing-pro8621 repos~2.5kAutomated safety check: PassMIT
SEO Keyword ClusteringAgriciDaniel/claude-seo19k2 repos~3.3kAutomated safety check: PassMIT
Evaluate Skillevery-app/open-seo23k—~1.8kAutomated safety check: NotesMIT
SEO Content Brief GeneratorAgriciDaniel/claude-seo19k2 repos~2.6kAutomated safety check: PassMIT
Blog GoogleAgriciDaniel/claude-blog2.3k1 repos~3.3kAutomated safety check: NotesMIT
FLOW SEO FrameworkAgriciDaniel/claude-seo19k2 repos~1.4kAutomated safety check: PassMIT

Similar skills

  • SEO Keyword Clustering

    AgriciDaniel/claude-seo

    Clusters keywords by how much their search results overlap and designs a hub-and-spoke content plan with an internal link matrix and an interactive cluster map.

    19k GitHub starsUsed in 2 repos~3.3k tokens
    Marketing & SEOAuto-check passed
  • Evaluate Skill

    every-app/open-seo

    Test a candidate OpenSEO skill end to end by running fresh, isolated Codex sessions against the local backend and scoring the reports they save.

    23k GitHub stars~1.8k tokensUpdated 2 days ago
    Marketing & SEOAuto-check: notes
  • SEO Content Brief Generator

    AgriciDaniel/claude-seo

    Builds research-backed SEO content briefs with competitor scoring, per-section word counts and page-type templates, for new pages or improving existing ones.

    19k GitHub starsUsed in 2 repos~2.6k tokens
    Marketing & SEOAuto-check passed
  • Blog Google

    AgriciDaniel/claude-blog

    Google API integration for blog performance: PageSpeed Insights, CrUX Core Web Vitals with 25-week history, Search Console performance, URL Inspection, Indexing API, GA4 organic traffic, NLP entity…

    2.3k GitHub starsUsed in 1 repo~3.3k tokens
    Marketing & SEOAuto-check: notes
  • FLOW SEO Framework

    AgriciDaniel/claude-seo

    Brings the FLOW framework's stage-specific SEO prompts into the agent, from keyword discovery through backlinks, on-page work and conversion to local SEO, loaded on demand.

    19k GitHub starsUsed in 2 repos~1.4k tokens
    Marketing & SEOAuto-check passed
  • SEO Dataforseo

    AgriciDaniel/codex-seo

    Live SEO data via DataForSEO MCP server. An agent skill from AgriciDaniel/codex-seo.

    799 GitHub starsUsed in 2 repos~4.6k tokens
    Marketing & SEOAuto-check passed

More from indranilbanerjee/digital-marketing-pro

All 162 skills in this repo
  • Import Template

    indranilbanerjee/digital-marketing-pro

    Import a deliverable template as a reusable placeholder template per brand.

    862 GitHub starsUsed in 1 repo~1.6k tokens
    Auto-check passed
  • Ab Test Plan

    indranilbanerjee/digital-marketing-pro

    Plan an A/B test by script: sample size per variant, days to run, stopping rules.

    862 GitHub starsUsed in 1 repo~1.9k tokens
    Auto-check passed
  • Aeo Audit

    indranilbanerjee/digital-marketing-pro

    Run a one-time AEO audit of six AI answer engines, scored per surface.

    862 GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • Agent Readiness Audit

    indranilbanerjee/digital-marketing-pro

    Audit agent readiness by script: AI-crawler rules, product schema, no-JS HTML, feeds.

    862 GitHub starsUsed in 1 repo~3.7k tokens
    Auto-check passed
  • Backlink Gap

    indranilbanerjee/digital-marketing-pro

    Find backlink gap domains linking to competitors, not you, scored by script.

    862 GitHub starsUsed in 1 repo~2.6k tokens
    Auto-check passed
  • C2pa Metadata

    indranilbanerjee/digital-marketing-pro

    Embed C2PA provenance in AI-generated images, video or PDF by script.

    862 GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed

Categories

Questions about Keyword Cluster

What does Keyword Cluster do?

Cluster keywords into pillar and spoke pages by SERP overlap via script. Keyword Cluster is an agent skill from indranilbanerjee/digital-marketing-pro. Cluster keywords into pillar and spoke pages by SERP overlap via script.

When should I use Keyword Cluster?

Keyword Cluster fits situations like: tasks that involve Keyword research.

How do I install Keyword Cluster in Claude Code?

Run `npx skills add indranilbanerjee/digital-marketing-pro --skill keyword-cluster -a claude-code`. Or copy the skill folder (skills/keyword-cluster in indranilbanerjee/digital-marketing-pro) into .claude/skills/keyword-cluster in your project. Claude Code loads it when a task matches its description.

How do I install Keyword Cluster in Codex?

Run `npx skills add indranilbanerjee/digital-marketing-pro --skill keyword-cluster -a codex`. Or copy the skill folder (skills/keyword-cluster in indranilbanerjee/digital-marketing-pro) into .agents/skills/keyword-cluster in your project. Codex loads it when a task matches its description.

Can I use Keyword Cluster in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add indranilbanerjee/digital-marketing-pro --skill keyword-cluster -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/keyword-cluster, .gemini/skills/keyword-cluster, .github/skills/keyword-cluster and .opencode/skills/keyword-cluster in your project.

What does Keyword Cluster need to run?

Going by SKILL.md and its folder, Keyword Cluster needs the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Keyword Cluster access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Keyword Cluster safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Keyword Cluster use?

Keyword Cluster is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Keyword Cluster use?

About 2.5k tokens (SKILL.md is roughly 10k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Keyword Cluster?

Skills that share tags, products or a category with Keyword Cluster: SEO Keyword Clustering (AgriciDaniel/claude-seo, 19k stars), Evaluate Skill (every-app/open-seo, 23k stars), SEO Content Brief Generator (AgriciDaniel/claude-seo, 19k stars) and Blog Google (AgriciDaniel/claude-blog, 2.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Keyword Cluster?

indranilbanerjee (a GitHub user) maintains it in indranilbanerjee/digital-marketing-pro, which has 862 GitHub stars. The repository holds 162 skills in this directory. The repository was last updated on October 9, 2026.

Source: indranilbanerjee/digital-marketing-pro on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.