Datalineage Summary
google/skills
Summarizes Google Cloud Data Lineage graphs to help users debug data quality issues and understand data provenance for BQ/GCS.
Query Fabric lakehouse and warehouse data using DuckDB, either locally or inside a Fabric notebook.
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdb --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .claude/skills/using-duckdb && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .claude/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdbType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdb --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .agents/skills/using-duckdb && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .agents/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdb --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .cursor/skills/using-duckdb && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .cursor/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/data-goblin/power-bi-agentic-development.git --path plugins/etl/skills/using-duckdb--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdb --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .gemini/skills/using-duckdb && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .gemini/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdbInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .github/skills/using-duckdb && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .github/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install data-goblin/power-bi-agentic-development using-duckdb --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/data-goblin/power-bi-agentic-development.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/etl/skills/using-duckdb .opencode/skills/using-duckdb && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "using-duckdb" agent skill from https://github.com/data-goblin/power-bi-agentic-development/tree/main/plugins/etl/skills/using-duckdb into .opencode/skills/using-duckdb/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "using-duckdb", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
using-duckdbQuery Fabric lakehouse and warehouse data using DuckDB, either locally or inside a Fabric notebook.
Using Duckdb is an agent skill from data-goblin/power-bi-agentic-development. Query Fabric lakehouse and warehouse data using DuckDB, either locally or inside a Fabric notebook. Automatically invoke when the user mentions "DuckDB", "query Delta tables locally", or asks to "attach DuckDB to a lakehouse", "query OneLake data", "explore lakehouse data", "data freshness check", "validate data quality", "use DuckDB in Fabric".
Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/common-patterns.md` and `references/in-notebook-setup.md`).
It sits in Data & Analytics, covering Data warehousing and Data cleaning. It works with DuckDB and Microsoft Azure. The repository describes itself as: Power BI AI skills and Power BI agents for Claude Code and GitHub Copilot: a plugin marketplace of Power BI skills, subagents, and hooks for semantic models, DAX, TMDL, reports… The licence is GPL-3.0.
Read from SKILL.md and the folder at commit 41886f2. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
duckdbazbrewFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
duckdb.orggithub.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
ACCESS_TOKENFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Using Duckdb loads about 1.2k tokens when it runs, and up to ~2.1k if it reads all its reference files. Until then it costs about 90 tokens; SKILL.md has 243 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from data-goblin/power-bi-agentic-development at commit 41886f2, republished under its GPL-3.0 licence (© data-goblin). 243 words, ~1,173 tokens.
.claude/skills/using-duckdb/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.Query Delta Lake tables and raw files in OneLake using DuckDB. Works both locally (CLI/Python) and inside Fabric notebooks. Read-only; for writes, use the executing-spark skill.
| Mode | Where it runs | Auth | Best for |
|---|---|---|---|
| Local | Developer machine | Azure CLI (az login) | Exploration, validation, ad-hoc analysis |
| In-notebook | Fabric Spark container | notebookutils.credentials.getToken('storage') | Combining DuckDB speed with Spark write-back |
brew install duckdb on macOS)az login)INSTALL delta; INSTALL azure; (one-time)WS_ID=$(fab get "Workspace.Workspace" -q "id" | tr -d '"')
LH_ID=$(fab get "Workspace.Workspace/LH.Lakehouse" -q "id" | tr -d '"')
duckdb -c "
LOAD delta; LOAD azure;
CREATE SECRET (TYPE azure, PROVIDER credential_chain, CHAIN 'cli');
SELECT * FROM delta_scan(
'abfss://${WS_ID}@onelake.dfs.fabric.microsoft.com/${LH_ID}/Tables/schema/table'
) LIMIT 10;
"The CHAIN 'cli' parameter uses Azure CLI credentials. Without it, DuckDB tries managed identity first (fails on local machines).
BASE="abfss://${WS_ID}@onelake.dfs.fabric.microsoft.com/${LH_ID}/Files"
duckdb -c "
LOAD azure;
CREATE SECRET (TYPE azure, PROVIDER credential_chain, CHAIN 'cli');
SELECT * FROM read_csv('${BASE}/data.csv') LIMIT 10;
SELECT * FROM read_parquet('${BASE}/facts.parquet') LIMIT 10;
SELECT * FROM read_json('${BASE}/events/*.json');
"Glob patterns (*, **) work for reading multiple files.
Inside a Fabric notebook, DuckDB can query lakehouse Delta tables directly using a storage token. This approach is faster than Spark SQL for analytical queries on single-node data.
import duckdb
import time
# Get storage token from notebook context
token = notebookutils.credentials.getToken('storage')
# Create DuckDB connection
con = duckdb.connect(f'temp_{time.time_ns()}.duckdb')
con.sql('SET enable_object_cache=true')
# Register OneLake secret
con.sql(f"""
CREATE OR REPLACE SECRET onelake (
TYPE AZURE,
PROVIDER ACCESS_TOKEN,
ACCESS_TOKEN '{token}'
)
""")
# Query Delta tables
workspace = "<workspace-id>"
lakehouse = "<lakehouse-name>"
path = f"abfss://{workspace}@onelake.dfs.fabric.microsoft.com/{lakehouse}.Lakehouse/Tables"
df = con.sql(f"""
SELECT * FROM delta_scan('{path}/schema/table_name') LIMIT 100
""").df()
print(df)Dynamically find all Delta tables in a lakehouse:
tables = con.sql(f"""
SELECT DISTINCT split_part(file, '_delta_log', 1) as table_path
FROM glob('{path}/*/*/*_delta_log/*.json')
""").df()['table_path'].tolist()
for t in tables:
view_name = t.split('/')[-1]
con.sql(f"CREATE OR REPLACE VIEW {view_name} AS SELECT * FROM delta_scan('{t}')")
print(f"Created view: {view_name}")abfss://<workspace-id>@onelake.dfs.fabric.microsoft.com/<item-id>/Tables/<schema>/<table>
abfss://<workspace-id>@onelake.dfs.fabric.microsoft.com/<item-id>/Files/<path>| Item type | ID source |
|---|---|
| Lakehouse | fab get "ws/LH.Lakehouse" -q "id" |
| Warehouse | fab get "ws/WH.Warehouse" -q "id" |
| SQL Database | fab get "ws/DB.SQLDatabase" -q "id" |
Cross-item joins work in a single DuckDB query; use different abfss:// paths.
For data freshness checks, quality validation, schema discovery, cross-table joins, and row count audits, see references/common-patterns.md.
references/common-patterns.md -- Data freshness, quality, schema discovery, cross-joinsreferences/in-notebook-setup.md -- Full notebook setup with auto-discovery and write-back patterns© data-goblin, GPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files (references) in plugins/etl/skills/using-duckdb of data-goblin/power-bi-agentic-development.
Open the folder on GitHubat commit 41886f2
Using Duckdb next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Using Duckdb this skilldata-goblin/power-bi-agentic-development | 1k | — | ~1.2k | Automated safety check: Pass | GPL-3.0 | |
| Datalineage Summarygoogle/skills | 21k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | |
| Google Cloud Solution Agentic Analytics Spark Knowledge Cataloggoogle/skills | 21k | — | ~4.4k | Automated safety check: Pass | Apache-2.0 | |
| Bloodhound AnalysisSpecterOps/skills | 702 | — | ~1.4k | Automated safety check: Pass | MIT | |
| Ga4 Auditcognyai/claude-code-marketing-skills | 104 | — | ~1.6k | Automated safety check: Pass | None | |
| Duckdb Experttheneoai/awesome-skills | 183 | — | ~4.2k | Automated safety check: Pass | MIT |
google/skills
Summarizes Google Cloud Data Lineage graphs to help users debug data quality issues and understand data provenance for BQ/GCS.
Discovers requirements and designs an end-to-end governed agentic analytics solution using Knowledge Catalog and Managed Service for Apache Spark (Lightning Engine).
SpecterOps/skills
Use as the default router for generic BloodHound asks: check the BloodHound connection, verify MCP health, analyze BloodHound data, find or explain a path, inspect shortest paths, find a path to…
cognyai/claude-code-marketing-skills
Google Analytics 4 configuration and data-quality audit — key events, data streams, custom dimensions, attribution, retention, PII, Ads link, BigQuery export
theneoai/awesome-skills
DuckDB expert for embedded OLAP analytics, Parquet/CSV querying, and high-performance analytical SQL on local data.
walkthru-earth/geocoding-playground
Investigates geocoder data quality issues by querying live S3 parquet files via MotherDuck MCP.
data-goblin/power-bi-agentic-development
Author, validate, publish, and test Power BI paginated reports in the RDL format.
data-goblin/power-bi-agentic-development
Automatically invoke this skill whenever the user asks about Fabric tenant settings or Power BI tenant settings or auditing tenant settings.
data-goblin/power-bi-agentic-development
Interactive BPA rule generation for Power BI semantic models; guided discovery, model investigation, and expert rule authoring.
data-goblin/power-bi-agentic-development
Guidance for Power BI Project (PBIP) structure, thick and thin reports, project renames, forks, and validation.
data-goblin/power-bi-agentic-development
Actionable feedback on the quality, usage, and effectiveness of Power BI reports.
data-goblin/power-bi-agentic-development
This skill should be used whenever the user mentions a "semantic model", "data model", or "dataset", or asks to "build", "model", "design", "optimize", "review", or "audit" one, or to "add a…
Works with
Categories
Query Fabric lakehouse and warehouse data using DuckDB, either locally or inside a Fabric notebook. Using Duckdb is an agent skill from data-goblin/power-bi-agentic-development. Query Fabric lakehouse and warehouse data using DuckDB, either locally or inside a Fabric notebook.
Using Duckdb fits situations like: mentions DuckDB; query Delta tables locally; asks to attach DuckDB to a lakehouse; query OneLake data.
Run `npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a claude-code`. Or copy the skill folder (plugins/etl/skills/using-duckdb in data-goblin/power-bi-agentic-development) into .claude/skills/using-duckdb in your project. Claude Code loads it when a task matches its description.
Run `npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a codex`. Or copy the skill folder (plugins/etl/skills/using-duckdb in data-goblin/power-bi-agentic-development) into .agents/skills/using-duckdb in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add data-goblin/power-bi-agentic-development --skill using-duckdb -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/using-duckdb, .gemini/skills/using-duckdb, .github/skills/using-duckdb and .opencode/skills/using-duckdb in your project.
Going by SKILL.md and its folder, Using Duckdb needs the command-line tools its instructions call (duckdb, az and brew) and credentials named ACCESS_TOKEN. Our summary lists: Python 3; A credential in ACCESS_TOKEN.
SKILL.md names 2 domains. As links in the text: duckdb.org and github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Using Duckdb is published under the GPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.2k tokens (SKILL.md is roughly 4.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 889 tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Using Duckdb: Datalineage Summary (google/skills, 21k stars), Google Cloud Solution Agentic Analytics Spark Knowledge Catalog (google/skills, 21k stars), Bloodhound Analysis (SpecterOps/skills, 702 stars) and Ga4 Audit (cognyai/claude-code-marketing-skills, 104 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
data-goblin (a GitHub user) maintains it in data-goblin/power-bi-agentic-development, which has 1,026 GitHub stars. The repository holds 33 skills in this directory. The repository was last updated on October 5, 2026.
Source: data-goblin/power-bi-agentic-development on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.