Agent skill

Test Metrics Dashboard

by proffesor-for-testing in proffesor-for-testing/agentic-qe

A skill your agent uses when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data.

MITAuto-check passedTesting & QA

Install Test Metrics Dashboard

skills CLI
$ npx skills add proffesor-for-testing/agentic-qe --skill test-metrics-dashboard -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install proffesor-for-testing/agentic-qe test-metrics-dashboard --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/proffesor-for-testing/agentic-qe.git skills-src && mkdir -p .claude/skills && cp -r skills-src/assets/skills/test-metrics-dashboard .claude/skills/test-metrics-dashboard && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-metrics-dashboard
GitHub stars
495
Token cost
~718 tokens
SKILL.md length
167 words
Files
1
Skills in repo
93
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data.

  • Querying test history
  • SKILL.md covers Activation, Key Metrics, Run History and Composition, plus 1 more section
  • Calls jq and npx
  • Analyzing flakiness rates

What it does

Test Metrics Dashboard is an agent skill from proffesor-for-testing/agentic-qe. Use when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data.

Its SKILL.md is about 720 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Failing and flaky tests. The repository describes itself as: Agentic QE Fleet is an open-source AI-powered QA/QE platform designed for use with Coding Agents (works best with Claude Code) featuring specialized agents and skills to support… The licence is MIT.

When your agent uses it

  • Querying test history
  • Analyzing flakiness rates
  • Building quality trend dashboards from test execution data

Example prompts

  • “/test-metrics-dashboard”

Requirements

  • Node.js

What it can do on your machine

Read from SKILL.md and the folder at commit 829d030. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • jq
    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Metrics Dashboard loads about 718 tokens when it runs. Until then it costs about 40 tokens; SKILL.md has 167 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~40
When it runs · the whole SKILL.md, loaded when a task matches
~718

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from proffesor-for-testing/agentic-qe at commit 829d030, republished under its MIT licence (© proffesor-for-testing). 167 words, ~718 tokens.

Download SKILL.mdSave it as .claude/skills/test-metrics-dashboard/SKILL.md (or your agent's skills folder).
name
test-metrics-dashboard
description
Use when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data.
user-invocable
true

Test Metrics Dashboard

Data & Analysis skill for querying test execution history, identifying trends, and surfacing actionable quality metrics.

Activation

/test-metrics-dashboard

Key Metrics

Test Health Metrics
MetricFormulaTargetAlert
Pass RatePassed / Total> 95%< 90%
Flakiness RateFlaky / Total< 5%> 10%
MTTRAvg time from failure to fix< 4 hours> 24 hours
Execution TimeTotal suite duration< 10 min> 20 min
Coverage DeltaCurrent - Previous>= 0%< -2%
Data Collection
bash
# Export Jest results to JSON
npx jest --json --outputFile=test-results/$(date +%Y-%m-%d).json

# Parse results for dashboard
jq '{
  date: .startTime,
  total: .numTotalTests,
  passed: .numPassedTests,
  failed: .numFailedTests,
  duration_ms: (.testResults | map(.endTime - .startTime) | add),
  pass_rate: ((.numPassedTests / .numTotalTests) * 100),
  flaky: [.testResults[] | select(.numPendingTests > 0)] | length
}' test-results/$(date +%Y-%m-%d).json
Trend Analysis
bash
# Compare last 5 runs
for f in $(ls -t test-results/*.json | head -5); do
  jq --arg file "$f" '{
    file: $file,
    pass_rate: ((.numPassedTests / .numTotalTests) * 100 | floor),
    duration_s: ((.testResults | map(.endTime - .startTime) | add) / 1000 | floor)
  }' "$f"
done
Top Failing Tests
bash
# Find most frequently failing tests across runs
for f in test-results/*.json; do
  jq -r '.testResults[] | select(.numFailingTests > 0) | .testFilePath' "$f"
done | sort | uniq -c | sort -rn | head -10

Run History

Store dashboard data in ${CLAUDE_PLUGIN_DATA}/test-metrics.log:

2026-03-18|95.2|4.1|312|82.5|3

Format: date|pass_rate|flakiness_rate|duration_s|coverage_pct|failed_count

Read history for trend detection:

bash
# Coverage trending down?
tail -5 "${CLAUDE_PLUGIN_DATA}/test-metrics.log" | awk -F'|' '{print $5}' | sort -n | head -1

Composition

Feeds into:

  • /qe-quality-assessment — quality gate decisions based on metrics
  • /test-failure-investigator — investigate top failing tests
  • /coverage-drop-investigator — when coverage trends down

Gotchas

  • Metrics without baselines are meaningless — establish baselines before tracking trends
  • Flakiness rate is underreported — a test that fails 1/100 times still breaks CI weekly
  • Duration trends upward over time as test count grows — set alerts on rate of increase, not absolute value
  • Agent may report metrics from a single run as "trends" — need 5+ data points for meaningful trends

© proffesor-for-testing, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in assets/skills/test-metrics-dashboard of proffesor-for-testing/agentic-qe.

Open the folder on GitHubat commit 829d030

Compare with similar skills

Test Metrics Dashboard next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Metrics Dashboard compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Metrics Dashboard this skillproffesor-for-testing/agentic-qe495—~718Automated safety check: PassMIT
Swig Testswig/swig6.3k—~2.3kAutomated safety check: PassCustom licence
Triage CI FailureDataDog/datadog-agent3.8k—~2.3kAutomated safety check: PassApache-2.0
Dynamo Jira TicketDynamoDS/Dynamo2k—~1.1kAutomated safety check: PassApache-2.0
Fix Ready PRsfastrepl/anarlog9.5k—~1.4kAutomated safety check: PassMIT
Trx Analysismicrosoft/vstest969—~1.8kAutomated safety check: PassMIT

Similar skills

  • Swig Test

    swig/swig

    Run SWIG test suite for specific languages. An agent skill from swig/swig.

    6.3k GitHub stars~2.3k tokensUpdated 2 days ago
    Testing & QAAuto-check passed
  • Triage CI Failure

    DataDog/datadog-agent

    Official

    Classify a failed CI as either caused by an active incident, flakiness, or a true code regression.

    3.8k GitHub stars~2.3k tokensUpdated today
    Testing & QAAuto-check passed
  • Dynamo Jira Ticket

    DynamoDS/Dynamo

    Create structured Jira tickets for Dynamo from bug reports, failing tests, or feature requests.

    2k GitHub stars~1.1k tokensUpdated today
    Testing & QAAuto-check passed
  • Fix Ready PRs

    fastrepl/anarlog

    Inspect every open non-draft PR for CI failures and unresolved Cursor Bugbot findings, then fix them on the existing PR branches.

    9.5k GitHub stars~1.4k tokensUpdated today
    Testing & QAAuto-check passed
  • Trx Analysis

    microsoft/vstest

    Official

    Parse and analyze Visual Studio TRX test result files. An agent skill from microsoft/vstest.

    969 GitHub stars~1.8k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Wio

    workersio/skills

    Testing workflow skill for finding high-value test candidates, writing focused tests, generating realistic workloads, reviewing test value, and diagnosing test-suite health.

    200 GitHub stars~5.8k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed

More from proffesor-for-testing/agentic-qe

All 93 skills in this repo
  • Contract Testing

    proffesor-for-testing/agentic-qe

    Consumer-driven contract testing for microservices using Pact, schema validation, API versioning, and backward compatibility testing.

    495 GitHub stars~1.8k tokensUpdated 5 days ago
    Auto-check passed
  • Mutation Testing

    proffesor-for-testing/agentic-qe

    Test quality validation through mutation testing, assessing test suite effectiveness by introducing code mutations and measuring kill rate.

    495 GitHub stars~1.7k tokensUpdated 5 days ago
    Auto-check passed
  • Performance Testing

    proffesor-for-testing/agentic-qe

    Profiles application performance under load using k6, Artillery, or JMeter to measure latency, throughput, and error rates.

    495 GitHub stars~2.4k tokensUpdated 5 days ago
    Auto-check passed
  • Code Review Quality

    proffesor-for-testing/agentic-qe

    Conduct context-driven code reviews focusing on quality, testability, and maintainability.

    495 GitHub starsUsed in 1 repo~1.9k tokens
    Auto-check passed
  • Security Testing

    proffesor-for-testing/agentic-qe

    Scans for security vulnerabilities including XSS, SQL injection, CSRF, and auth flaws using OWASP Top 10 methodology.

    495 GitHub stars~2.7k tokensUpdated 5 days ago
    Auto-check: notes
  • Database Testing

    proffesor-for-testing/agentic-qe

    Database schema validation, data integrity testing, migration testing, transaction isolation, and query performance.

    495 GitHub starsUsed in 1 repo~1.7k tokens
    Auto-check passed

Categories

Questions about Test Metrics Dashboard

What does Test Metrics Dashboard do?

A skill your agent uses when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data. Test Metrics Dashboard is an agent skill from proffesor-for-testing/agentic-qe. Use when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data.

When should I use Test Metrics Dashboard?

Test Metrics Dashboard fits situations like: querying test history; analyzing flakiness rates; building quality trend dashboards from test execution data.

How do I install Test Metrics Dashboard in Claude Code?

Run `npx skills add proffesor-for-testing/agentic-qe --skill test-metrics-dashboard -a claude-code`. Or copy the skill folder (assets/skills/test-metrics-dashboard in proffesor-for-testing/agentic-qe) into .claude/skills/test-metrics-dashboard in your project. Claude Code loads it when a task matches its description.

How do I install Test Metrics Dashboard in Codex?

Run `npx skills add proffesor-for-testing/agentic-qe --skill test-metrics-dashboard -a codex`. Or copy the skill folder (assets/skills/test-metrics-dashboard in proffesor-for-testing/agentic-qe) into .agents/skills/test-metrics-dashboard in your project. Codex loads it when a task matches its description.

Can I use Test Metrics Dashboard in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add proffesor-for-testing/agentic-qe --skill test-metrics-dashboard -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-metrics-dashboard, .gemini/skills/test-metrics-dashboard, .github/skills/test-metrics-dashboard and .opencode/skills/test-metrics-dashboard in your project.

What does Test Metrics Dashboard need to run?

Going by SKILL.md and its folder, Test Metrics Dashboard needs the command-line tools its instructions call (jq and npx). Our summary lists: Node.js.

Does Test Metrics Dashboard access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Test Metrics Dashboard safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Test Metrics Dashboard use?

Test Metrics Dashboard is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Metrics Dashboard use?

About 718 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Test Metrics Dashboard?

Skills that share tags, products or a category with Test Metrics Dashboard: Swig Test (swig/swig, 6.3k stars), Triage CI Failure (DataDog/datadog-agent, 3.8k stars), Dynamo Jira Ticket (DynamoDS/Dynamo, 2k stars) and Fix Ready PRs (fastrepl/anarlog, 9.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Metrics Dashboard?

proffesor-for-testing (a GitHub user) maintains it in proffesor-for-testing/agentic-qe, which has 495 GitHub stars. The repository holds 93 skills in this directory. The repository was last updated on October 4, 2026.

Source: proffesor-for-testing/agentic-qe on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.