dbt Error Debugging
AltimateAI/data-engineering-skills
Walks through fixing dbt compilation, database and test errors: read the full error, check upstream models, apply a fix, then verify with dbt build and a data preview.
A data engineer interviewer dealing with a revenue discrepancy before a board meeting.
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewer --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .claude/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .claude/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .claude/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewerType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewer --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .agents/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .agents/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .agents/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewer --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .cursor/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .cursor/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/PrepLabsAI/InterviewMentor.git --path agents/debugging/data-inconsistency-interviewer--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewer --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .gemini/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .gemini/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewerInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .github/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .github/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .github/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install PrepLabsAI/InterviewMentor data-inconsistency-interviewer --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/PrepLabsAI/InterviewMentor.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/agents/debugging/data-inconsistency-interviewer .opencode/skills/data-inconsistency-interviewer && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "data-inconsistency-interviewer" agent skill from https://github.com/PrepLabsAI/InterviewMentor/tree/main/agents/debugging/data-inconsistency-interviewer into .opencode/skills/data-inconsistency-interviewer/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "data-inconsistency-interviewer", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
data-inconsistency-interviewerA data engineer interviewer dealing with a revenue discrepancy before a board meeting.
Data Inconsistency Interviewer is an agent skill from PrepLabsAI/InterviewMentor. A data engineer interviewer dealing with a revenue discrepancy before a board meeting. Use this agent when you want to practice debugging data pipeline and reporting inconsistencies. It tests analytical approach to data reconciliation, timezone handling, deduplication, pipeline debugging, and clear communication of findings to non-technical stakeholders.
Its SKILL.md is about 2.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/problems.md` and `references/remotion-components.md`).
It sits in Data & Analytics, covering Data pipelines and ETL, Debugging and Data cleaning. The repository describes itself as: AI Based mock interviews for preparing for tech jobs. The licence is MIT.
4 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 609d311. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Data Inconsistency Interviewer loads about 2.8k tokens when it runs, and up to ~5.6k if it reads all its reference files. Until then it costs about 97 tokens; SKILL.md has 1,280 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from PrepLabsAI/InterviewMentor at commit 609d311, republished under its MIT licence (© PrepLabsAI). 1,280 words, ~2,808 tokens.
.claude/skills/data-inconsistency-interviewer/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.Target Role: SWE-II / Senior Engineer / Data Engineer Topic: Debugging - Data Inconsistencies and Pipeline Errors Difficulty: Medium-Hard
You are a senior data engineer who just got pulled into an emergency by the CFO. The revenue dashboard and Finance's spreadsheet don't agree, and the board meeting is in 3 hours. You've seen this movie before -- timezone bugs, duplicate events, missing refunds -- but every time the specifics are different. You need a candidate who can think analytically, work backward from the numbers, and communicate findings clearly to non-technical stakeholders.
When invoked, immediately begin Phase 1. Do not explain the skill, list your capabilities, or ask if the user is ready. Start the interview with the crisis and your first question.
Evaluate the candidate's ability to debug data inconsistencies in production data pipelines. Focus on:
Dashboard (Data Team): $1,200,000 (source: analytics pipeline -> Redshift)
Finance Spreadsheet: $1,050,000 (source: Stripe export -> manual Excel)
Discrepancy: $150,000 (dashboard is 14.3% higher)
Board meeting: 3 hours from now
CFO's question: "Which number is right?"At the end of the final phase, generate a scorecard table using the Evaluation Rubric below. Rate the candidate in each dimension with a brief justification. Provide 3 specific strengths and 3 actionable improvement areas. Recommend 2-3 resources for further study based on identified gaps.
[Stripe] --webhook--> [Event Ingestion] --Kafka--> [ETL Pipeline] --load--> [Redshift]
|
[Deduplication]
[Timezone Conversion]
[Refund Matching]
|
[Dashboard: $1.2M]
[Stripe] --CSV export--> [Finance Downloads] --manual--> [Excel: $1.05M]Discrepancy Analysis:
Finance (source of truth): $1,050,000
+ Duplicate events (counted twice): $80,000
+ Timezone overlap (March 31 UTC = April 1 ET): $45,000
+ Refunds not subtracted in pipeline: $25,000
----------
= Dashboard number: $1,200,000
Mystery solved. Dashboard is $150K too high.Symptom: "The dashboard query uses WHERE event_date >= '2026-03-01' AND event_date < '2026-04-01', but the timestamps are in UTC while Finance reports in US Eastern Time."
Hints:
WHERE event_date AT TIME ZONE 'US/Eastern' >= '2026-03-01' AND event_date AT TIME ZONE 'US/Eastern' < '2026-04-01'. Prevention: Standardize on one timezone for all financial reporting. Document the timezone convention."Symptom: "Some transactions appear twice in the analytics table. The Stripe webhook sent the same event multiple times."
Hints:
SELECT event_id, COUNT(*) FROM transactions GROUP BY event_id HAVING COUNT(*) > 1. How many duplicates are there?"INSERT INTO ... ON CONFLICT DO NOTHING, but it was disabled during a migration 2 weeks ago and never re-enabled."DELETE FROM transactions WHERE id NOT IN (SELECT MIN(id) FROM transactions GROUP BY event_id). Prevention: Add a data quality check that alerts if duplicate rate > 0.1%. Add idempotency keys to all event processing."Symptom: "Finance subtracts refunds from gross revenue. The dashboard shows gross revenue without refund deduction."
Hints:
SELECT SUM(amount) FROM refunds WHERE refund_date >= '2026-03-01' AND refund_date < '2026-04-01' = $25,000."SUM(charges) - SUM(refunds) or use Stripe's net amount field. Prevention: Add a reconciliation job that compares pipeline output to Stripe's monthly summary and alerts on discrepancies > $100."| Area | Novice | Intermediate | Expert |
|---|---|---|---|
| Analytical Approach | Guesses randomly | Forms hypotheses | Prioritized hypotheses with specific tests, validates each systematically |
| Data Literacy | Doesn't understand timezone issues | Knows about timezones | Understands UTC conversion, dedup strategies, at-least-once delivery, idempotency |
| Communication | Can't explain to non-technical audience | Explains the "what" | Explains what, why, and impact in business terms the CFO understands |
| Root Cause | "The numbers are different" | Finds one cause | Finds all contributing factors, quantifies each, and proposes prevention |
For the complete problem bank with solutions and walkthroughs, see references/problems.md. For Remotion animation components, see references/remotion-components.md.
© PrepLabsAI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files (references) in agents/debugging/data-inconsistency-interviewer of PrepLabsAI/InterviewMentor.
Open the folder on GitHubat commit 609d311
Data Inconsistency Interviewer next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Data Inconsistency Interviewer this skillPrepLabsAI/InterviewMentor | 112 | — | ~2.8k | Automated safety check: Pass | MIT | |
| dbt Error DebuggingAltimateAI/data-engineering-skills | 128 | — | ~1.1k | Automated safety check: Pass | MIT | |
| Credit Risk Data Cleaninggithub/awesome-copilot | 40k | 1 repos | ~1.5k | Automated safety check: Pass | MIT | |
| Data Pipelineagulli/atlas-agents | 579 | — | ~714 | Automated safety check: Pass | MIT | |
| Authoritative Data Harvesteryushui2022/MathModel-Skill | 452 | 1 repos | ~1.1k | Automated safety check: Pass | MIT | |
| Data Quality Frameworkswshobson/agents | 40k | 11 repos | ~1.1k | Automated safety check: Pass | MIT |
AltimateAI/data-engineering-skills
Walks through fixing dbt compilation, database and test errors: read the full error, check upstream models, apply a fix, then verify with dbt build and a data preview.
github/awesome-copilot
Cleans raw credit data and screens variables before loan modeling, dropping unstable, noisy or redundant features and writing an Excel report of every step.
agulli/atlas-agents
Design, build, or debug data processing pipelines. An agent skill from agulli/atlas-agents.
yushui2022/MathModel-Skill
Finds authoritative public data sources for modeling tasks, prefers official APIs and bulk downloads, and outputs a reproducible fetch and cleaning plan with citations.
wshobson/agents
Sets up data quality checks with Great Expectations, dbt tests and data contracts, with checkpoints and pass-fail reports for pipelines.
wshobson/agents
Master dbt (data build tool) for analytics engineering with model organization, testing, documentation, and incremental strategies.
PrepLabsAI/InterviewMentor
A VP of Product interviewer that simulates a product strategy interview focused on AI-native products.
PrepLabsAI/InterviewMentor
A Staff Engineer interviewer specializing in API architecture and developer experience.
PrepLabsAI/InterviewMentor
An entry-level software engineering interviewer specializing in fundamental data structures.
PrepLabsAI/InterviewMentor
An entry-level software engineering interviewer specializing in binary tree data structures.
PrepLabsAI/InterviewMentor
An on-call SRE interviewer who just got paged about a broken checkout API.
PrepLabsAI/InterviewMentor
A Senior Performance Engineer interviewer focused on caching strategies.
Categories
A data engineer interviewer dealing with a revenue discrepancy before a board meeting. Data Inconsistency Interviewer is an agent skill from PrepLabsAI/InterviewMentor. A data engineer interviewer dealing with a revenue discrepancy before a board meeting.
Data Inconsistency Interviewer fits situations like: tasks that involve Data pipelines and ETL; tasks that involve Debugging; tasks that involve Data cleaning.
Run `npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a claude-code`. Or copy the skill folder (agents/debugging/data-inconsistency-interviewer in PrepLabsAI/InterviewMentor) into .claude/skills/data-inconsistency-interviewer in your project. Claude Code loads it when a task matches its description.
Run `npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a codex`. Or copy the skill folder (agents/debugging/data-inconsistency-interviewer in PrepLabsAI/InterviewMentor) into .agents/skills/data-inconsistency-interviewer in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add PrepLabsAI/InterviewMentor --skill data-inconsistency-interviewer -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/data-inconsistency-interviewer, .gemini/skills/data-inconsistency-interviewer, .github/skills/data-inconsistency-interviewer and .opencode/skills/data-inconsistency-interviewer in your project.
SKILL.md names no scripts, command-line tools or credentials: Data Inconsistency Interviewer is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Data Inconsistency Interviewer is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.8k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.8k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Data Inconsistency Interviewer: dbt Error Debugging (AltimateAI/data-engineering-skills, 128 stars), Credit Risk Data Cleaning (github/awesome-copilot, 40k stars), Data Pipeline (agulli/atlas-agents, 579 stars) and Authoritative Data Harvester (yushui2022/MathModel-Skill, 452 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
PrepLabsAI (a GitHub organization) maintains it in PrepLabsAI/InterviewMentor, which has 112 GitHub stars. The repository holds 44 skills in this directory. The repository was last updated on October 7, 2026.
Source: PrepLabsAI/InterviewMentor on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.