Data Quality Frameworks
wshobson/agents
Sets up data quality checks with Great Expectations, dbt tests and data contracts, with checkpoints and pass-fail reports for pipelines.
Audit a dbt project against the Entropy Data reference layout and add anything missing — Open Data Product Specification (ODPS), Open Data Contract Standard (ODCS), OpenLineage transport config…
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-sync --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .claude/skills/entropy-data-sync && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .claude/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-syncType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-sync --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .agents/skills/entropy-data-sync && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .agents/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-sync --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .cursor/skills/entropy-data-sync && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .cursor/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/hashgraph-online/awesome-codex-plugins.git --path plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-sync --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .gemini/skills/entropy-data-sync && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .gemini/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-syncInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .github/skills/entropy-data-sync && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .github/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install hashgraph-online/awesome-codex-plugins entropy-data-sync --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync .opencode/skills/entropy-data-sync && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "entropy-data-sync" agent skill from https://github.com/hashgraph-online/awesome-codex-plugins/tree/main/plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync into .opencode/skills/entropy-data-sync/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "entropy-data-sync", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
entropy-data-syncAudit a dbt project against the Entropy Data reference layout and add anything missing — Open Data Product Specification (ODPS), Open Data Contract Standard (ODCS), OpenLineage transport config…
Entropy Data Sync is an agent skill from hashgraph-online/awesome-codex-plugins. Audit a dbt project against the Entropy Data reference layout and add anything missing — Open Data Product Specification (ODPS), Open Data Contract Standard (ODCS), OpenLineage transport config, output-port model layout, and the GitHub Actions publish workflow. Trigger when the user asks to integrate a dbt project with Entropy Data, set up Entropy Data publishing, or check whether a dbt project follows the Entropy Data conventions.
Its SKILL.md is about 4.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 10 other files (for example `templates/.github/workflows/data-product.yml`, `templates/data-product.odps.yaml` and `templates/models/output_ports/v1/contract.odcs.yaml`).
It sits in Data & Analytics, covering Data pipelines and ETL and Data governance. It works with dbt and GitHub Actions. The repository describes itself as: A curated list of awesome OpenAI Codex / ChatGPT plugins, skills, and resources. The 1 Codex Marketplace. See live plugins at: https://hol.org/plugins/best-codex-plugins. The licence is Apache-2.0.
6 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 16b4156. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
gituvFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
github.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
ENTROPY_DATA_API_KEYDBT_DATABRICKS_TOKENFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Entropy Data Sync loads about 4.6k tokens when it runs. Until then it costs about 113 tokens; SKILL.md has 2,163 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from hashgraph-online/awesome-codex-plugins at commit 16b4156, republished under its Apache-2.0 licence (© hashgraph-online). 2,163 words, ~4,588 tokens.
.claude/skills/entropy-data-sync/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.Make sure a dbt project is well-integrated with Entropy Data.
A dbt project is well-integrated with Entropy Data when it has all of:
| # | Artifact | Path | Purpose |
|---|---|---|---|
| 1 | Open Data Product Specification | <data-product-id>.odps.yaml at repo root | Declares the data product, team, output ports |
| 2 | Output-port data contracts | models/output_ports/v<N>/<contract-id>.odcs.yaml (one per output port — what this data product commits to produce) | Schema + server config the contract test runs against; colocated with the SQL that implements it |
| 3 | Input-port data contracts | models/input_ports/<provider-output-port-id>.odcs.yaml (one per active access agreement — what this data product trusts upstream to produce) | Cached snapshot of the upstream provider's ODCS; refreshed via entropy-data datacontracts get, never hand-edited |
| 4 | OpenLineage transport | openlineage.yml at repo root | Makes dbt-ol run send lineage to api.entropy-data.com |
| 5 | Model layout | models/{input_ports,staging,intermediate,output_ports/v1}/ | Convention that mirrors the data product's lifecycle |
| 6 | Publish workflow | .github/workflows/data-product.yml | CI: dbt run/test → publish ODPS + output ODCS → run contract test |
| 7 | Git connections | One per ODPS + one per output-port ODCS, registered via entropy-data dataproducts gitconnection put and entropy-data datacontracts gitconnection put | Lets Entropy Data link the published spec back to the YAML in the repo, and enables pull / push / push-pr from the CLI. Input-port ODCS files are not registered — they belong to the upstream data product, which owns its own git connection |
Work in this exact order. Do not skip the audit.
${PLUGIN_ROOT}below refers to the root of this plugin — the directory that containsskills/. On Claude Code it is set automatically as${CLAUDE_PLUGIN_ROOT}— use that. On any other agent (Codex, Copilot CLI, etc.) it is unset; resolve it as../..relative to thisSKILL.mdfile's directory (i.e. the grandparent ofskills/<this-skill>/).
Before running Step 0, print the following plan to the user verbatim so they know what's about to happen:
Running entropy-data-sync. I'll:
- Verify the
entropy-dataCLI is installed and connected.- Confirm this is a dbt project and pick up its name.
- Audit existing Entropy Data artifacts (ODPS, ODCS, OpenLineage, model layout, publish workflow, git connections).
- Gather any missing parameters from you (one batched question).
- Apply fixes — create missing files, patch incomplete ones, register git connections.
- Summarize what changed and what's deferred.
Then proceed.
Confirm uv run --quiet entropy-data --version succeeds from the project root. If it fails, run uv sync (the bootstrap template seeds entropy-data as a dev dep) and retry. If still missing, stop and tell the user to verify entropy-data is in pyproject.toml's [dependency-groups].dev. Use uv run entropy-data … for every CLI invocation in this skill.
Run entropy-data connection test. If it fails (no connection, expired key, etc.), stop and tell the user to run entropy-data connection add <name> --host <host> --api-key <key> first. Do not prompt for the key yourself.
Check that dbt_project.yml exists at the working directory root. If not, stop and tell the user this skill only works inside a dbt project.
Read dbt_project.yml and remember the name: value — call it DBT_PROJECT_NAME. By convention it is also the dbt profile and the data product id.
For each row in the table above, check whether the artifact is present. For row 7 (git connections), call:
entropy-data dataproducts gitconnection get <DATA_PRODUCT_ID> -o jsonentropy-data datacontracts gitconnection get <CONTRACT_ID> -o json for each output-port contract under models/output_ports/**/For rows 2 and 3, glob the file system:
models/output_ports/**/*.odcs.yamlmodels/input_ports/*.odcs.yamlLegacy projects may still have contracts at the old datacontracts/ path. If datacontracts/ exists and contains *.odcs.yaml files, flag those rows as migration needed and surface the move target (output contracts → models/output_ports/v1/; input contracts, if any, → models/input_ports/). Do not move them silently — Step 4 asks the user.
If a get returns a 404 (or "not found"), mark that connection as missing. If it returns a connection whose repository-url / repository-path / repository-branch does not match the local repo, mark it as drifted and call it out separately — do not silently overwrite. If the underlying data product or contract doesn't exist on the platform yet (the workflow hasn't run for the first time), or if the working directory is not a git repository (git rev-parse --is-inside-work-tree errors or returns false), mark git connections as deferred with a one-line explanation.
For row 1 (ODPS file), also check that the top-level customProperties list contains an entry with property: "dataProductBuilder" and value: "https://github.com/entropy-data/dataproduct-builder-dbt". If the file exists but the property is missing, mark the ODPS as incomplete with a one-line note ("missing dataProductBuilder customProperty"); Step 4 will add it without touching other fields. Forks of this plugin should substitute their own builder URL in the template before publishing.
Produce a short audit report like:
Entropy Data integration audit for <DBT_PROJECT_NAME>:
[✓] ODPS file
[✗] Output-port contracts (no *.odcs.yaml under models/output_ports/)
[⏸] Input-port contracts (none — populated by dataproduct-implement from access agreements)
[✓] openlineage.yml
[✗] Model layout (no models/output_ports)
[✗] GitHub Actions publish workflow
[⏸] Git connections (deferred: data product not yet published — run the workflow first)Show the report. Then list what you intend to create. Wait for the user to confirm before writing any files.
Before generating files, fill in these placeholders. Infer from the project where you can; ask the user for the rest in one batched question.
| Placeholder | Default / inference | Notes |
|---|---|---|
DATA_PRODUCT_ID | DBT_PROJECT_NAME | Used as id in ODPS and as the dbt profile name |
DATA_PRODUCT_NAME | Title-cased DBT_PROJECT_NAME | Human-friendly name |
OUTPUT_PORT_NAME | DBT_PROJECT_NAME | One output port per ODCS file |
CONTRACT_ID | <DATA_PRODUCT_ID>-v1 | Stable id used by entropy-data datacontracts put |
CONTRACT_FILE | <contract_id>.odcs.yaml | File under models/output_ports/v1/ |
CONTRACT_PATH | models/output_ports/v1/<CONTRACT_FILE> | Full repo-relative path; used by --repository-path, the CI workflow, and datacontract test |
TABLE | last segment of DBT_PROJECT_NAME | Output table name |
PURPOSE | — | Ask the user (one sentence) |
TEAM_NAME | — | If <DATA_PRODUCT_ID>.odps.yaml already exists with a team.name, use that. Otherwise, prefer a team id registered in Entropy Data — invoke the entropy-data-teams skill (in this same plugin) so the user can pick from the existing teams, and use the returned id. Fall back to a free-text answer only if entropy-data-teams cannot run (CLI unavailable / not authenticated) |
TAG | — | Ask the user (e.g. a usecases/... slug) |
PLATFORM | — | Ask the user: databricks, snowflake, bigquery, s3, postgres |
CATALOG / SCHEMA | — | Ask the user (Databricks: catalog + schema; Snowflake: database + schema; BigQuery: project + dataset) |
DBT_PROFILE | DBT_PROJECT_NAME | Used in the workflow's profiles.yml block |
ODPS_FILE | <DATA_PRODUCT_ID>.odps.yaml | Path passed to entropy-data dataproducts put |
API_HOST | entropy-data connection get -o json → host | Resolve in Step 4, only when writing openlineage.yml or the workflow. Uses the same host the CLI is authenticated against, so lineage and CI publish hit the same deployment |
GIT_REPOSITORY_URL | git remote get-url origin | Used by gitconnection put. If no origin, ask the user; if the remote is git@… SSH form, convert to the equivalent HTTPS URL the platform expects |
GIT_REPOSITORY_BRANCH | git rev-parse --abbrev-ref HEAD, falling back to main | Used by gitconnection put; if HEAD is detached, ask the user |
GIT_CONNECTION_TYPE | inferred from GIT_REPOSITORY_URL: github.com → github, gitlab.com → gitlab, bitbucket.org → bitbucket, dev.azure.com / *.visualstudio.com → azuredevops | Ask the user only if the host doesn't match any of these |
GIT_HOST | the URL host, only when self-hosted (i.e. not one of the SaaS hosts above); otherwise omit | Passed as --host to gitconnection put |
GIT_CREDENTIAL_EXTERNAL_ID | — | Optional. Ask the user; if they don't have one yet, leave the connection unauthenticated (it can still be used for read-only metadata in the UI) |
For each missing artifact, copy the corresponding template from ${PLUGIN_ROOT}/skills/entropy-data-sync/templates/ into the user's project, substituting placeholders. Do not overwrite existing files; if a file is present but incomplete, surface the diff and ask before changing.
When (and only when) you're about to write openlineage.yml or .github/workflows/data-product.yml, resolve API_HOST from the active CLI connection:
entropy-data connection get -o jsonUse the host field to substitute {{API_HOST}} in those templates. Self-hosted deployments are handled via entropy-data connection add --host <host>, not a plugin-level setting.
If the ODPS file exists but was flagged as incomplete — missing dataProductBuilder customProperty in Step 2, append the entry to the top-level customProperties list (do not reorder or touch other entries):
customProperties:
- property: "dataProductBuilder"
value: "https://github.com/entropy-data/dataproduct-builder-dbt"Surface the diff and ask before saving.
The templates live at:
templates/data-product.odps.yaml → write to <DATA_PRODUCT_ID>.odps.yamltemplates/models/output_ports/v1/contract.odcs.yaml → write to <CONTRACT_PATH> (i.e. models/output_ports/v1/<CONTRACT_FILE>)templates/openlineage.yml → write to openlineage.ymltemplates/.github/workflows/data-product.yml → write to .github/workflows/data-product.ymlIf the audit reported a legacy datacontracts/ directory, ask the user before moving its contents. The default migration is:
datacontracts/<contract>.odcs.yaml → models/output_ports/v1/<contract>.odcs.yaml (if it matches an output port). If multiple output port versions exist, ask which one.--repository-path for each moved contract (Step 4b).datacontracts/ directory only after the user confirms the move.This skill does not create input-port ODCS files. They appear only when dataproduct-implement resolves an access agreement and writes the cached upstream contract to models/input_ports/. If the audit found stale input-port ODCS files (no matching .source.yaml), surface them as orphans and let the user decide whether to delete them.
For the model layout, create the directories models/input_ports/, models/staging/, models/intermediate/, models/output_ports/v1/ if absent, plus _models.yml placeholders so dbt does not warn about empty directories. Do not move existing models — only add the empty subfolders the user is missing, and note it in the report.
Also update dbt_project.yml's models: block so the materializations match the reference (output port = table, staging/intermediate = view):
models:
<DBT_PROJECT_NAME>:
+materialized: table
staging:
+materialized: view
intermediate:
+materialized: viewIf the models: block already exists, only add missing keys; do not clobber the user's customizations.
Only run this sub-step if the audit (Step 2) flagged at least one git connection as missing or the user confirmed re-creating a drifted one. Skip entirely if every connection is already correct, if the audit marked them as deferred, or if the working directory is not a git repository (check with git rev-parse --is-inside-work-tree — if it errors or returns false, there's no remote to register; tell the user to run git init and add a remote first, then re-run this skill).
For the data product:
entropy-data dataproducts gitconnection put <DATA_PRODUCT_ID> \
--repository-url <GIT_REPOSITORY_URL> \
--repository-path <ODPS_FILE> \
--repository-branch <GIT_REPOSITORY_BRANCH> \
--git-connection-type <GIT_CONNECTION_TYPE> \
[--host <GIT_HOST>] \
[--git-credential-external-id <GIT_CREDENTIAL_EXTERNAL_ID>]For each output-port ODCS file (models/output_ports/**/*.odcs.yaml):
entropy-data datacontracts gitconnection put <CONTRACT_ID> \
--repository-url <GIT_REPOSITORY_URL> \
--repository-path <CONTRACT_PATH> \
--repository-branch <GIT_REPOSITORY_BRANCH> \
--git-connection-type <GIT_CONNECTION_TYPE> \
[--host <GIT_HOST>] \
[--git-credential-external-id <GIT_CREDENTIAL_EXTERNAL_ID>]Do not register git connections for input-port ODCS files. They are cached copies of upstream contracts; the upstream data product owns the canonical record.
Notes:
--repository-path is relative to the repo root, not the working directory. The ODPS path is just <DATA_PRODUCT_ID>.odps.yaml; output-port contract paths look like models/output_ports/v<N>/<CONTRACT_FILE>.--host for SaaS providers (github.com, gitlab.com, bitbucket.org, dev.azure.com); set it only for self-hosted instances.put is upsert.Always end with this exact two-part format so the user gets a consistent recap.
Part 1 — outcome table. One row per artifact from the audit. Use the Status enum below; Details is a short, plain-text note (file path, or "—" if nothing to add).
| Artifact | Status | Details |
|---|---|---|
| ODPS file | … | … |
| Output-port contracts | … | <N> file(s) at models/output_ports/v<N>/<CONTRACT_FILE> |
| Input-port contracts | … | <N> file(s) at models/input_ports/<provider-output-port-id>.odcs.yaml (or "—" if no access agreements yet) |
| OpenLineage transport | … | … |
| Model layout | … | … |
| Publish workflow | … | … |
| Git connections | … | … |
Status enum (use exactly these words):
created — the skill wrote a new file or registered a new connection.updated — the skill patched an existing file or fixed a drifted connection.already present — no change needed.deferred — skipped intentionally (data product/contract not yet published, or no git repo). The deferred command(s) appear in Part 2.skipped — the user declined when asked to confirm.Part 2 — next steps. Bullet list, only include the items that apply:
deferred git connection, the exact entropy-data dataproducts gitconnection put … or entropy-data datacontracts gitconnection put … command to run after the first CI publish.ENTROPY_DATA_API_KEY, plus platform creds (DBT_DATABRICKS_HOST, DBT_DATABRICKS_HTTP_PATH, DBT_DATABRICKS_TOKEN for Databricks; equivalents for other platforms)."<CONTRACT_PATH> — the template only seeds id and updated_at."dbt-ol run locally once to verify lineage flows to Entropy Data (requires OPENLINEAGE__TRANSPORT__AUTH__APIKEY)."If there is nothing in Part 2, write a single line: No further action required.
dbt-databricks install line, the Create profiles.yml block, and the DATACONTRACT_* env vars to the matching dialect. Do not generate a Databricks workflow for a non-Databricks project.id + updated_at and tell the user to fill in the rest, or — if dbt models already exist for the output port — derive columns from _models.yml if available.gitconnection get returns a record matching the local repo URL / branch / path, do not call put.© hashgraph-online, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 4 other files in plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync of hashgraph-online/awesome-codex-plugins.
Open the folder on GitHubat commit 16b4156
Entropy Data Sync next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Entropy Data Sync this skillhashgraph-online/awesome-codex-plugins | 1.2k | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | |
| Data Quality Frameworkswshobson/agents | 40k | 10 repos | ~1.1k | Automated safety check: Pass | MIT | |
| Monte Carlo Preventsickn33/agentic-awesome-skills | 47k | 1 repos | ~3.3k | Automated safety check: Pass | MIT | |
| Phy Pipeline Contract EnforcerLeoYeAI/openclaw-master-skills | 2.2k | — | ~4.7k | Automated safety check: Pass | Apache-2.0 | |
| Modeling Warehouse FoundationsPostHog/posthog-foss | 721 | — | ~2.1k | Automated safety check: Pass | MIT | |
| Dbt Databricks PR Readydatabricks/dbt-databricks | 379 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 |
wshobson/agents
Sets up data quality checks with Great Expectations, dbt tests and data contracts, with checkpoints and pass-fail reports for pipelines.
sickn33/agentic-awesome-skills
Surfaces Monte Carlo data observability context (table health, alerts, lineage, blast radius) before SQL/dbt edits.
LeoYeAI/openclaw-master-skills
Data pipeline contract enforcer. An agent skill from LeoYeAI/openclaw-master-skills.
PostHog/posthog-foss
Shared foundations for building reusable data models in PostHog, on either of two stacks: PostHog-native data-warehouse views / materialized views (HogQL, via the view- MCP tools), or an external…
databricks/dbt-databricks
A skill your agent uses for an open dbt-databricks pull request, including your own PR or a fork PR, to assess merge readiness and optionally repair selected gaps on the PR head branch.
MaterializeInc/materialize
Cut a dbt-materialize PyPI release: bump the version in version.py and setup.py, date the Unreleased CHANGELOG entry, and open the release PR with a Ship: <url body.
hashgraph-online/awesome-codex-plugins
Create original anime-style reaction stickers as looping GIFs and MP4 previews, using generated character pose sheets and timed key poses.
hashgraph-online/awesome-codex-plugins
Manage and query Calibre libraries with the calibredb CLI (local paths or Calibre Content server URLs).
hashgraph-online/awesome-codex-plugins
A skill your agent uses when adding, changing, testing, or debugging Rust HTTP APIs and services, especially when Codex needs black-box integration tests, random-port app startup, real database test…
hashgraph-online/awesome-codex-plugins
Make a studio's game look like something at build time — a cover from a real frame of the game (free), painted covers, backdrops, textures and character plates from image models through the…
hashgraph-online/awesome-codex-plugins
Use CALL-E from Codex through the calle CLI. An agent skill from hashgraph-online/awesome-codex-plugins.
hashgraph-online/awesome-codex-plugins
Balance game difficulty, resources, rewards, probability, progression, economies, and dominant strategies.
Works with
Categories
Audit a dbt project against the Entropy Data reference layout and add anything missing — Open Data Product Specification (ODPS), Open Data Contract Standard (ODCS), OpenLineage transport config…. Entropy Data Sync is an agent skill from hashgraph-online/awesome-codex-plugins. Audit a dbt project against the Entropy Data reference layout and add anything missing — Open Data Product Specification (ODPS), Open Data Contract Standard (ODCS), OpenLineage transport config, output-port model layout, and the GitHub Actions publish workflow.
Entropy Data Sync fits situations like: the user asks to integrate a dbt project with Entropy Data; set up Entropy Data publishing; check whether a dbt project follows the Entropy Data conventions.
Run `npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a claude-code`. Or copy the skill folder (plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync in hashgraph-online/awesome-codex-plugins) into .claude/skills/entropy-data-sync in your project. Claude Code loads it when a task matches its description.
Run `npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a codex`. Or copy the skill folder (plugins/entropy-data/dataproduct-builder-dbt/skills/entropy-data-sync in hashgraph-online/awesome-codex-plugins) into .agents/skills/entropy-data-sync in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add hashgraph-online/awesome-codex-plugins --skill entropy-data-sync -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/entropy-data-sync, .gemini/skills/entropy-data-sync, .github/skills/entropy-data-sync and .opencode/skills/entropy-data-sync in your project.
Going by SKILL.md and its folder, Entropy Data Sync needs the command-line tools its instructions call (git and uv) and credentials named ENTROPY_DATA_API_KEY and DBT_DATABRICKS_TOKEN.
SKILL.md names 1 domain. In commands or code: github.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Entropy Data Sync is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 4.6k tokens (SKILL.md is roughly 18k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Entropy Data Sync: Data Quality Frameworks (wshobson/agents, 40k stars), Monte Carlo Prevent (sickn33/agentic-awesome-skills, 47k stars), Phy Pipeline Contract Enforcer (LeoYeAI/openclaw-master-skills, 2.2k stars) and Modeling Warehouse Foundations (PostHog/posthog-foss, 721 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
hashgraph-online (a GitHub organization) maintains it in hashgraph-online/awesome-codex-plugins, which has 1,232 GitHub stars. The repository holds 736 skills in this directory. The repository was last updated on October 6, 2026.
Source: hashgraph-online/awesome-codex-plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.