Agent skill

Purview Data Catalog

by Kilo-Org in Kilo-Org/kilo-marketplace

Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map.

MITAuto-check passedData & Analytics

Install Purview Data Catalog

skills CLI
$ npx skills add Kilo-Org/kilo-marketplace --skill purview-data-catalog -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Kilo-Org/kilo-marketplace purview-data-catalog --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Kilo-Org/kilo-marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/purview-data-catalog .claude/skills/purview-data-catalog && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
purview-data-catalog
GitHub stars
190
Used in
1 other repo
Token cost
~2.1k tokens
SKILL.md length
1,049 words
Files
2
Skills in repo
85
Repo updated
First seen
Licence
MIT

At a glance

Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map.

  • Works in 8 steps: Pre-requisite: Data Map is scanning and… → Pick the right governance domains —… → Publish 3-5 data products per domain — A… → …
  • Tasks that involve Data governance
  • SKILL.md covers When to use, Pick the right object for the…, Approach and Guardrails, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Purview Data Catalog is an agent skill from Kilo-Org/kilo-marketplace. Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map. Covers governance domains, data products, and curation. WHEN: Purview data catalog, unified catalog, data products, governance domain, business glossary, data quality, data discovery for analysts, curate data assets, data stewardship.

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file.

It sits in Data & Analytics, covering Data governance and Data cleaning. The repository describes itself as: Kilo Marketplace - A curated collection of Skills, MCP Servers, and Modes for enhancing AI agent capabilities across the Kilo ecosystem—including Kilo Code (VS Code extension)… The licence is MIT.

When your agent uses it

  • Tasks that involve Data governance
  • Tasks that involve Data cleaning

Example prompts

  • “/purview-data-catalog”

Workflow steps

8 steps, taken from the first numbered list in SKILL.md.

  1. Pre-requisite: Data Map is scanning and producing technical metadata — The Unified
  2. Pick the right governance domains — Start with 2-3 high-pain business areas, not the
  3. Publish 3-5 data products per domain — A data product groups related assets for a
  4. Build the glossary in parallel, not first — Building a 500-term glossary up front
  5. Assign stewardship with named individuals, not teams — "Owned by the data team" =
  6. Configure data quality on critical data elements only — Do not scan everything. Pick
  7. Drive adoption with discoverability — Wire the catalog into Power BI (catalog endorsement
  8. Operate as a continuous programme — Monthly: review unowned assets, stale products,

What it can do on your machine

Read from SKILL.md and the folder at commit ff51758. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • learn.microsoft.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Purview Data Catalog loads about 2.1k tokens when it runs. Until then it costs about 111 tokens; SKILL.md has 1,049 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~111
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Kilo-Org/kilo-marketplace at commit ff51758, republished under its MIT licence (© Kilo-Org). 1,049 words, ~2,134 tokens.

Download SKILL.mdSave it as .claude/skills/purview-data-catalog/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
purview-data-catalog
description
Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map. Covers governance domains, data products, and curation. WHEN: Purview data catalog, unified catalog, data products, governance domain, business glossary, data quality, data discovery for analysts, curate data assets, data stewardship.
metadata.author
Microsoft
metadata.version
0.1.0
metadata.category
data

Microsoft Purview Unified Catalog

The Unified Catalog provides business-friendly data discovery and governance on top of the Data Map: organising assets into governance domains and data products with glossary terms, ownership, and data quality. This skill is the catalog/business layer; the Data Map (purview-data-map) is the technical scanning layer underneath.

When to use

Use this skill when the user wants analysts and data consumers to find, trust, and request data - not when they want to scan sources (that is the Data Map's job).

Do not use this skill for:

  • Scanning lakehouses, databases, SaaS sources (use purview-data-map)
  • Sensitivity labels for security/compliance (use purview-data-classification)
  • Lifecycle/retention rules (use purview-data-lifecycle)
  • AI data discovery (use purview-dspm-ai)

Pick the right object for the job

If you want to...ObjectOwned by
Define a business area with accountability (Finance, HR, Supply Chain)Governance domainDomain owner (business)
Bundle related assets into something an analyst can request and consumeData productData product owner (business)
Give a business term a shared definition tied to data assetsGlossary termDomain steward
Measure trustworthiness of a critical data elementData quality rule + scoreData steward
Show how a column flows from source to reportLineage (auto, from Data Map scans)System-generated
Grant a consumer access to a data productAccess policy / request workflowData product owner

Rule of thumb: start with 1-2 governance domains and 5-10 data products. A catalog with 200 thinly-curated products is less useful than one with 10 great ones. Curation depth beats catalog breadth, every time.

Approach

A catalog rollout dies when IT "loads everything" without business ownership. Follow the order; each step gates the next.

  1. Pre-requisite: Data Map is scanning and producing technical metadata — The Unified Catalog sits on top of the Data Map. Confirm scans are running, lineage is populating, and classifications are firing before designing domains. Verify: Data Map shows scanned assets with classifications (e.g. "Credit Card Number") and at least one lineage link source → sink.
  2. Pick the right governance domains — Start with 2-3 high-pain business areas, not the org chart. Good first picks: Finance (regulatory pressure), Customer (consent and CRM sprawl), HR (privacy). Each domain needs a named business owner who has time, not a delegated IT proxy. Verify: each domain has a named accountable owner from the business side with a recurring 30-min weekly slot for curation.
  3. Publish 3-5 data products per domain — A data product groups related assets for a specific consumer use case (e.g. "Customer 360 for marketing analysts"). Each product has: description, glossary terms linked, owner, access guidance, and at least one data quality rule on a key column. Do not publish empty shells. Verify: a representative consumer can find the product, understand what is in it, and request access via the in-portal workflow without external help.
  4. Build the glossary in parallel, not first — Building a 500-term glossary up front produces a graveyard of terms with no asset links. Define terms as you build products; each term must link to at least one data product or asset at publication. Verify: every published glossary term is linked to ≥1 data asset.
  5. Assign stewardship with named individuals, not teams — "Owned by the data team" = owned by nobody. Each governance domain and data product has a single named steward. Stewardship workload should be < 4 hours/week per steward, or it stops happening.
  6. Configure data quality on critical data elements only — Do not scan everything. Pick 5-10 critical data elements per data product (e.g. customer_id, order_amount, date_of_birth) and define rules: completeness, uniqueness, validity, freshness. Publish the score on the product page. Verify: each high-priority data product shows a quality score and the rules behind it.
  7. Drive adoption with discoverability — Wire the catalog into Power BI (catalog endorsement surfaces in the BI service), Microsoft Search, and Teams. If consumers cannot find products where they already work, they will not use the catalog. Verify: a Power BI dataset surfaces its catalog endorsement; a Teams search returns a linked data product.
  8. Operate as a continuous programme — Monthly: review unowned assets, stale products, broken lineage. Quarterly: domain owner check-in on curation completeness. Annually: sunset retired products.
Show full SKILL.md (362 more words)Show less

Guardrails

  • Align catalog access with Data Map collection permissions. A consumer with no scan-level access cannot see the underlying asset even if the catalog product is published to them - permissions are AND, not OR. Plan this together.
  • Curation is a continuous programme, not a one-time load. Without weekly steward effort, the catalog goes stale within a quarter and trust collapses.
  • Do not auto-bulk-publish. Importing 10,000 assets via scan ≠ a catalog. Assets without curation are noise that hides the good products.
  • Glossary terms without asset links are graveyards. Enforce "must link to publish".
  • Data quality rules cost compute. Each rule runs on a schedule and pulls data. Limit to critical data elements; do not put quality on every column.
  • Domain count creep kills the programme. 3 domains with depth beats 30 with names only. Add a domain only when an existing one has > 10 well-curated data products.
  • Personal data needs governance, not just publishing. If a product contains personal data, coordinate with privacy (Priva) before exposing it to consumer search.

Common anti-patterns

  • "Build the glossary first." 500 terms, 0 asset links, programme dies in 6 months. Build terms with the products that need them.
  • "Auto-publish every scanned asset." Catalog floods, consumers cannot find anything, trust drops, never recovers.
  • "One steward owns 50 products." Not stewardship - a queue. Cap at ~10 products per steward.
  • "Catalog as compliance theatre." Published to satisfy an audit, no consumer ever uses it, no curation budget allocated. Sunset it or invest properly.
  • "Skip discoverability integration." Catalog exists but consumers search Teams/SharePoint for data instead. Wire it into existing workflows or accept zero adoption.
  • "Data quality on everything." Compute cost explodes; meaningful signal drowns in noise. Critical data elements only.

Example prompts

  • Set up the Purview unified catalog with governance domains and data products.
  • Which governance domains should I start with?
  • Build a business glossary and data quality rules.
  • How do analysts discover and curate data assets?
  • Establish data stewardship in the data catalog.
  • How many data products should one steward own?

Microsoft Learn

© Kilo-Org, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/purview-data-catalog of Kilo-Org/kilo-marketplace.

  • SKILL.md
  • LICENSE

Open the folder on GitHubat commit ff51758

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Kilo-Org/kilo-marketplace, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Purview Data Catalog next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Purview Data Catalog compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Purview Data Catalog this skillKilo-Org/kilo-marketplace1901 repos~2.1kAutomated safety check: PassMIT
Data Quality Frameworkswshobson/agents40k11 repos~1.1kAutomated safety check: PassMIT
Research Data Feasibility and Leakage ChecksLight0305/Light-skills641—~4.9kAutomated safety check: PassMIT
Datalineage Summarygoogle/skills21k—~1.7kAutomated safety check: PassApache-2.0
Monte Carlo Context Detectionsickn33/agentic-awesome-skills47k1 repos~2.6kAutomated safety check: WarnMIT
Querying AWS Sagemaker Catalogaws/agent-toolkit-for-aws2.8k—~2.6kAutomated safety check: PassApache-2.0

Similar skills

  • Sets up data quality checks with Great Expectations, dbt tests and data contracts, with checkpoints and pass-fail reports for pipelines.

    40k GitHub starsUsed in 11 repos~1.1k tokens
    Data & AnalyticsAuto-check passed
  • Finds usable public datasets, judges whether the data can support a research idea, and checks train and test splits for leakage before results are trusted.

    641 GitHub stars~4.9k tokensUpdated 3 mo ago
    Data & AnalyticsAuto-check passed
  • Datalineage Summary

    google/skills

    Official

    Summarizes Google Cloud Data Lineage graphs to help users debug data quality issues and understand data provenance for BQ/GCS.

    21k GitHub stars~1.7k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Monte Carlo Context Detection

    sickn33/agentic-awesome-skills

    Route data-related requests to the right Monte Carlo skill or workflow.

    47k GitHub starsUsed in 1 repo~2.6k tokens
    Data & AnalyticsAuto-check: warnings
  • Querying AWS Sagemaker Catalog

    aws/agent-toolkit-for-aws

    Official

    Runs SQL analytics on SageMaker Catalog asset metadata tables exported as Apache Iceberg in S3 Tables.

    2.8k GitHub stars~2.6k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Data Governance

    cbrock84/headcount

    Establishes ownership, definitions, quality, access, and lineage for the organization's data.

    2k GitHub stars~864 tokensUpdated 20 days ago
    Data & AnalyticsAuto-check passed

More from Kilo-Org/kilo-marketplace

All 85 skills in this repo
  • AzureML Project Scaffolding

    Kilo-Org/kilo-marketplace

    Sets up and maintains AzureML-ready Python projects as uv workspaces with devcontainers, a Makefile and job YAML, so local runs match cloud jobs and experiments stay reproducible.

    190 GitHub stars~3.1k tokensUpdated 9 days ago
    Auto-check: notes
  • Jupyter Notebook Builder

    Kilo-Org/kilo-marketplace

    Creates, inspects, edits and runs Jupyter notebooks, scaffolding experiment or tutorial notebooks from templates and preferring a Jupyter MCP server over raw JSON edits.

    190 GitHub stars~1.3k tokensUpdated 9 days ago
    Auto-check passed
  • Tableau Dashboard Creator

    Kilo-Org/kilo-marketplace

    Takes a plain-language dashboard request through brand setup, data exploration, planning, an interactive HTML mock and a Tableau implementation spec.

    190 GitHub stars~3.8k tokensUpdated 9 days ago
    Auto-check: notes
  • Elasticsearch File Ingest

    Kilo-Org/kilo-marketplace

    Ingest and transform data files (CSV/JSON/Parquet/Arrow IPC) into Elasticsearch with stream processing and custom transforms.

    190 GitHub stars~2.8k tokensUpdated 9 days ago
    Auto-check passed
  • Nifi Flow Layout

    Kilo-Org/kilo-marketplace

    A skill your agent uses when arranging Apache NiFi processors, process groups, ports, comments, numbering, crossing connections, dense fan-in/fan-out, or reusable readable canvas layouts.

    190 GitHub stars~1.5k tokensUpdated 9 days ago
    Auto-check passed
  • Splunk Ingest Processor Setup

    Kilo-Org/kilo-marketplace

    Render Cisco Data Fabric ingest-time routing workflows and Splunk Cloud Platform Ingest Processor setup plans with SPL2 pipelines, source types, destinations, lifecycle handoffs, queue and…

    190 GitHub stars~1.2k tokensUpdated 9 days ago
    Auto-check passed

Questions about Purview Data Catalog

What does Purview Data Catalog do?

Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map. Purview Data Catalog is an agent skill from Kilo-Org/kilo-marketplace. Guidance for the Microsoft Purview Unified Catalog (data catalog) — business-friendly discovery, governance domains, data products, glossary terms, and data quality on top of the Data Map.

When should I use Purview Data Catalog?

Purview Data Catalog fits situations like: tasks that involve Data governance; tasks that involve Data cleaning.

How do I install Purview Data Catalog in Claude Code?

Run `npx skills add Kilo-Org/kilo-marketplace --skill purview-data-catalog -a claude-code`. Or copy the skill folder (skills/purview-data-catalog in Kilo-Org/kilo-marketplace) into .claude/skills/purview-data-catalog in your project. Claude Code loads it when a task matches its description.

How do I install Purview Data Catalog in Codex?

Run `npx skills add Kilo-Org/kilo-marketplace --skill purview-data-catalog -a codex`. Or copy the skill folder (skills/purview-data-catalog in Kilo-Org/kilo-marketplace) into .agents/skills/purview-data-catalog in your project. Codex loads it when a task matches its description.

Can I use Purview Data Catalog in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Kilo-Org/kilo-marketplace --skill purview-data-catalog -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/purview-data-catalog, .gemini/skills/purview-data-catalog, .github/skills/purview-data-catalog and .opencode/skills/purview-data-catalog in your project.

What does Purview Data Catalog need to run?

SKILL.md names no scripts, command-line tools or credentials: Purview Data Catalog is instructions for the agent only.

Does Purview Data Catalog access the network?

SKILL.md names 1 domain. As links in the text: learn.microsoft.com. This is read from the text; nothing was executed.

Is Purview Data Catalog safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Purview Data Catalog use?

Purview Data Catalog is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Purview Data Catalog use?

About 2.1k tokens (SKILL.md is roughly 8.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Purview Data Catalog?

Skills that share tags, products or a category with Purview Data Catalog: Data Quality Frameworks (wshobson/agents, 40k stars), Research Data Feasibility and Leakage Checks (Light0305/Light-skills, 641 stars), Datalineage Summary (google/skills, 21k stars) and Monte Carlo Context Detection (sickn33/agentic-awesome-skills, 47k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Purview Data Catalog?

Kilo-Org (a GitHub organization) maintains it in Kilo-Org/kilo-marketplace, which has 190 GitHub stars. The repository holds 85 skills in this directory. The repository was last updated on September 28, 2026.

Source: Kilo-Org/kilo-marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.