Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human…
MITAuto-check passed
Install Assessing Comprehension Debt
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a claude-code
Project install by default; add -g for ~/.claude/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .claude/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Type this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a codex
Project install goes to .agents/skills/; add -g for ~/.codex/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .agents/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a cursor
Project install goes to .agents/skills/; add -g for ~/.cursor/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .cursor/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a gemini-cli
Project install goes to .agents/skills/; add -g for ~/.gemini/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .gemini/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Installs for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a github-copilot
Project install goes to .agents/skills/; add -g for ~/.copilot/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .github/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
skills CLI
$ npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a opencode
OpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
Install the "assessing-comprehension-debt" agent skill from https://github.com/jaktestowac/awesome-copilot-for-testers/tree/main/skills/assessing-comprehension-debt into .opencode/skills/assessing-comprehension-debt/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "assessing-comprehension-debt", then confirm the skill loads.
OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
Facts
Skill name
assessing-comprehension-debt
GitHub stars
116
Token cost
~2.3k tokens
SKILL.md length
1,234 words
Files
3
Skills in repo
13
Repo updated
First seen
Licence
MIT
At a glance
Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human…
Works in 4 steps: Establish scope and whether provenance… → Compute the risk band per change or module → Check teach-back attestation on… → …
An AI-assisted codebase grows faster than the team reads it
SKILL.md covers When to Use, Operating Principles, Workflow and What This Skill Cannot Do, plus 4 more sections
Calls git
What it does
Assessing Comprehension Debt is an agent skill from jaktestowac/awesome-copilot-for-testers. Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human explanation accompanied it, and optional AI-authorship provenance. Findings stay advisory by design. Use when an AI-assisted codebase grows faster than the team reads it, when reviews are rubber-stamped, when nobody can explain a module that ships weekly, or when leadership asks how much of the code the team can actually maintain.
Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `resources/risk-band-rubric.md` and `resources/teach-back-protocol.md`).
The repository describes itself as: 👨💻 Instructions, prompts, and chat modes to help You with test automation for GitHub Copilot 🤖. The licence is MIT.
When your agent uses it
An AI-assisted codebase grows faster than the team reads it
Reviews are rubber-stamped
Nobody can explain a module that ships weekly
Leadership asks how much of the code the team can actually maintain
Example prompts
“Use the assessing-comprehension-debt skill to measure the risk that code shipped without anyone understanding it: a teach-back attestation on…”
“/assessing-comprehension-debt”
Workflow steps
4 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 8910672. It shows what the files ask for, not the result of running them.
Tool permissions
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Runs code
Shell commands in SKILL.md call:
git
From the folder's file list and the shell code blocks in SKILL.md.
Network
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Credentials
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Context cost
Assessing Comprehension Debt loads about 2.3k tokens when it runs. Until then it costs about 137 tokens; SKILL.md has 1,234 words of instructions outside code blocks.
Always· name and description, kept in context so the agent knows when to use it
~137
When it runs· the whole SKILL.md, loaded when a task matches
~2.3k
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
Safety
Auto-check passed
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
Download SKILL.mdSave it as .claude/skills/assessing-comprehension-debt/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
assessing-comprehension-debt
description
Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human explanation accompanied it, and optional AI-authorship provenance. Findings stay advisory by design. Use when an AI-assisted codebase grows faster than the team reads it, when reviews are rubber-stamped, when nobody can explain a module that ships weekly, or when leadership asks how much of the code the team can actually maintain.
argument-hint
Commit range or release scope, whether AI-provenance trailers are in use, and which modules matter most
user-invocable
true
Assessing Comprehension Debt
Use this skill when code is arriving faster than anyone is reading it, and every other quality signal still looks fine.
Comprehension debt is the gap between the code that exists and the code the team actually understands. Tests pass, coverage holds, lint is clean - and nobody can explain why the pricing module works. When an agent writes the implementation and the tests, both signals go green without a single human forming a mental model. That is the debt this skill makes visible.
It is advisory, permanently, on purpose. Understanding lives in people's heads and cannot be proven by any signal. What can be measured is the risk of comprehension debt and the absence of evidence of understanding. A gate that claims to measure understanding is lying, and once someone notices, every finding it ever produced loses credibility.
When to Use
an AI-assisted or agent-generated codebase is growing faster than the team reads it
reviews are fast, approvals are frequent, and nobody asks questions
a module ships weekly and no one volunteers to explain it
a bus-factor conversation needs evidence instead of anecdote
a quarterly quality review should cover more than test coverage
onboarding is slow in a codebase that looks well-tested
Operating Principles
Never blocks. Findings cap below the project's blocking threshold. State that ceiling in the report, every time.
Measure absence of evidence, not understanding. Say it in those words. The band is a proxy; treat it as a prompt for a conversation, not a score.
The risk band is directional, not precise. It says "this change is the shape of one nobody understands", not "nobody understands this change".
AI provenance is a signal, never a penalty. The useful figure is AI-authored and unattested. AI-authored and well-understood is a good outcome, and if trailers are not in use the honest answer is "unknown".
A human explanation lowers the band. An Intent: trailer or an ADR is evidence that somebody thought about it. That is exactly the behaviour to reward.
Teach-back beats approval. "I approve" is a click. "I can explain what happens when the provider retries this webhook" is comprehension.
Never blame individuals. A high band on a module is a system outcome: review load, delivery pressure, tooling. Naming people converts a useful signal into something nobody will run twice.
Trend over snapshot. One reading is noise. Direction over a quarter is the finding.
Workflow
Phase 1: Establish scope and whether provenance exists
Pick the range: a release, a quarter, a module's history. Then check what signals are available:
If Assisted-by: trailers are not in use, provenance is unknown - report it as unknown and move on. Do not infer AI authorship from commit size, style, or timing; those inferences are wrong often enough to poison the whole report.
Phase 2: Compute the risk band per change or module
Three inputs, from ./resources/risk-band-rubric.md:
Input
Signal
Complexity added
new branches, new conditionals, nesting depth, new cross-module calls
Size
changed lines, files touched, whether it lands as one commit or a reviewable sequence
Explanation present
an Intent: trailer, an ADR link, a module register entry, or a substantive review discussion
Bands: high (large, branch-heavy, no explanation), medium (one of the three), low (small, or well explained, or both).
A high band means: if nobody understands this, we would not be able to tell. That is all it means, and saying so plainly is what keeps the metric usable.
Phase 3: Check teach-back attestation on high-risk surface
For changes on high-risk surface (recording-change-intent has the rules), look for a record that a named human can explain it:
a Comprehension-Attested-by: <name> trailer, author-side or reviewer-side
an entry in the attestation register (attesting-manual-verification)
Status is FULL (attested), NONE (high-risk surface, no record), or N/A (not high-risk surface).
Where a record is missing and it matters, run the teach-back in ./resources/teach-back-protocol.md: four questions, ten minutes, and the answers tell you more than the band ever will. The protocol's value is not the record - it is that the conversation happens.
Phase 4: Report the picture
Risk bands by module, with the inputs that produced each one
Attestation status on high-risk surface, and the specific changes with none
Provenance, if known: AI-authored share, and the share that is AI-authored and unattested
Trend against the previous assessment, per module
Concentration - comprehension debt clusters. A module where one person authored everything and nobody reviewed it is the finding, regardless of any band
The ceiling, stated explicitly: these findings are advisory and do not fail the build
Then the only recommendations worth making: which specific modules deserve a teach-back session, which deserve a walkthrough written down, and which deserve a second pair of eyes on the next change.
Show full SKILL.md (453 more words)Show less
What This Skill Cannot Do
Worth stating in the report, because the temptation to over-read the number is strong:
It cannot tell you whether a person understands code. Only whether evidence exists.
It cannot distinguish elegant-and-obvious from complex-and-opaque. A high band on genuinely well-written code is a false positive; check before acting.
It cannot detect understanding that lives in a conversation, a diagram, or a head. Absence of a trailer is not absence of thought.
It cannot be used for performance assessment, and using it that way guarantees the trailers become theatre.
Common Failure Modes
Letting it block. The single failure that discredits the whole practice.
Treating the band as a verdict. "High band" is a prompt to have a conversation, not a defect.
Inferring AI authorship. From commit size, from phrasing, from the hour of the day. Wrong often, and corrosive when it is.
Penalising Assisted-by:. Teams stop using the trailer, and the one honest signal disappears.
Naming individuals. Converts a system signal into a personnel matter, and the next assessment never happens.
Confusing approval with comprehension. A merged PR with an approving click is not evidence of understanding.
Reporting one number for the repo. Comprehension debt concentrates; an average hides exactly the module you needed to see.
Skipping the teach-back. The band is the cheap proxy; the conversation is the actual value.
Resource Map
./resources/risk-band-rubric.md - the three inputs, how to compute each, band boundaries, worked examples, and known false positives
./resources/teach-back-protocol.md - the four questions, how to run a ten-minute session, what a weak answer looks like, and how to record the outcome
Related Skills
recording-change-intent - an Intent: record is an input to the band, and the sibling debt this one pairs with
attesting-manual-verification - where a teach-back record lives and how it expires
governing-quality-waivers - the third governance record: why a check is off
tech-debt-analysis - code and architecture health, as opposed to whether anyone understands it
documenting-test-suites - the usual remediation when a suite is unmaintainable by anyone who did not build it
code-review-advanced - where a teach-back conversation naturally belongs
analyzing-quality-metrics - for trending the band honestly, with its caveats attached
Definition of Done
This skill is complete when:
the scope is stated and the available signals are established, with provenance marked unknown where trailers are not in use
every change or module in scope has a band with the three inputs that produced it
teach-back status on high-risk surface is FULL / NONE / N/A, with the specific unattested changes named
concentration is reported, not just averages
the trend against a previous assessment is included where one exists
the advisory ceiling is stated explicitly in the report
recommendations are specific sessions and walkthroughs, not "improve documentation"
Assessing Comprehension Debt next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
Assessing Comprehension Debt compared with similar skills
Skill
Stars
Used in
Tokens
Auto-check
Licence
Repo updated
Assessing Comprehension Debt this skilljaktestowac/awesome-copilot-for-testers
Conduct a defensible cybersecurity risk assessment using the NIST SP 800-30 Rev 1 methodology: prepare scope and a risk model, identify threat sources and threat events, identify vulnerabilities and…
Measures portfolio and backtest risk with VaR, CVaR, maximum drawdown, Monte Carlo simulation, tail modeling and stress tests, using one tested risk module.
Assess an enterprise’s regulatory penalty risk across four dimensions: licensing/qualifications, compliance with regulatory rules, and historical penalty/credit records.
Packages repository skills as installable Copilot plugins: marketplace registration, plugin.json manifests, generated skill copies, and the sync check CI enforces.
Turns "we will skip this check for now" into a dated, attributed, expiring waiver with a stated reason and owner, inventories the silent skips already hiding in a repo - skipped tests, disabled lint…
Requires an externalised rationale for high-risk changes - new public exports, new endpoints, auth edits, migrations, removed guards - recorded as an Intent commit trailer, an ADR reference, or a…
Sets up and maintains visual regression testing: what to snapshot, baseline strategy, masking dynamic regions, threshold tuning, containerized baselines, and the review-and-update workflow.
116 GitHub stars~2.6k tokensUpdated 1 mo ago
Auto-check passed
Questions about Assessing Comprehension Debt
What does Assessing Comprehension Debt do?
Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human…. Assessing Comprehension Debt is an agent skill from jaktestowac/awesome-copilot-for-testers. Measures the risk that code shipped without anyone understanding it: a teach-back attestation on high-risk changes, a risk band from changed-code complexity, diff size and whether a human explanation accompanied it, and optional AI-authorship provenance.
When should I use Assessing Comprehension Debt?
Assessing Comprehension Debt fits situations like: an AI-assisted codebase grows faster than the team reads it; reviews are rubber-stamped; nobody can explain a module that ships weekly; leadership asks how much of the code the team can actually maintain.
How do I install Assessing Comprehension Debt in Claude Code?
Run `npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a claude-code`. Or copy the skill folder (skills/assessing-comprehension-debt in jaktestowac/awesome-copilot-for-testers) into .claude/skills/assessing-comprehension-debt in your project. Claude Code loads it when a task matches its description.
How do I install Assessing Comprehension Debt in Codex?
Run `npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a codex`. Or copy the skill folder (skills/assessing-comprehension-debt in jaktestowac/awesome-copilot-for-testers) into .agents/skills/assessing-comprehension-debt in your project. Codex loads it when a task matches its description.
Can I use Assessing Comprehension Debt in Cursor, Gemini CLI or GitHub Copilot?
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jaktestowac/awesome-copilot-for-testers --skill assessing-comprehension-debt -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/assessing-comprehension-debt, .gemini/skills/assessing-comprehension-debt, .github/skills/assessing-comprehension-debt and .opencode/skills/assessing-comprehension-debt in your project.
What does Assessing Comprehension Debt need to run?
Going by SKILL.md and its folder, Assessing Comprehension Debt needs the command-line tools its instructions call (git).
Does Assessing Comprehension Debt access the network?
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Is Assessing Comprehension Debt safe to install?
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
What licence does Assessing Comprehension Debt use?
Assessing Comprehension Debt is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
How many tokens does Assessing Comprehension Debt use?
About 2.3k tokens (SKILL.md is roughly 9.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
What are the alternatives to Assessing Comprehension Debt?
Skills that share tags, products or a category with Assessing Comprehension Debt: Conducting Cyber Risk Assessment With Nist 800 30 (mukul975/Anthropic-Cybersecurity-Skills, 34k stars), Feature Risk Assessment (anthropics/claude-for-legal, 9.6k stars), Risk Measurement and Stress Testing (HKUDS/Vibe-Trading, 35k stars) and Climate Risk Assessment (mohitagw15856/pm-claude-skills, 1.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Who maintains Assessing Comprehension Debt?
jaktestowac (a GitHub user) maintains it in jaktestowac/awesome-copilot-for-testers, which has 116 GitHub stars. The repository holds 13 skills in this directory. The repository was last updated on August 26, 2026.