Official agent skill

Validation Diagnostics

by microsoft in microsoft/haste

Validation and diagnostic skill for HASTE. An agent skill from microsoft/haste.

OfficialMITAuto-check passed

Install Validation Diagnostics

skills CLI
$ npx skills add microsoft/haste --skill validation-diagnostics -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install microsoft/haste validation-diagnostics --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/microsoft/haste.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.github/skills/validation-diagnostics .claude/skills/validation-diagnostics && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
validation-diagnostics
GitHub stars
106
Token cost
~1.1k tokens
SKILL.md length
354 words
Files
1
Skills in repo
9
Repo updated
First seen
Licence
MIT

At a glance

Validation and diagnostic skill for HASTE. An agent skill from microsoft/haste.

  • Works in 4 steps: Document the miss in the diagnostic report → Identify the pattern (was it a… → Propose a skill update or instruction… → …
  • : validate implementation
  • SKILL.md covers Overview, Key Concepts, Patterns & Techniques and Decision Framework, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Validation Diagnostics is an agent skill from microsoft/haste, published by the product's own GitHub organization. Validation and diagnostic skill for HASTE. Compare planned vs implemented work, generate diagnostic reports, and feed misses back into skill refinement. Use when: 'validate implementation', 'compare to spec', 'diagnostic report', 'drift analysis', 'coverage check', 'implementation review'.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It works with Pydantic. The repository describes itself as: High Speed Assessment and Satellite Tracking for Emergencies. The licence is MIT.

When your agent uses it

  • : validate implementation
  • Compare to spec
  • Diagnostic report
  • Implementation review

Example prompts

  • “validate implementation”
  • “compare to spec”
  • “diagnostic report”
  • “/validation-diagnostics”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Document the miss in the diagnostic report
  2. Identify the pattern (was it a convention violation? missing test? spec ambiguity?)
  3. Propose a skill update or instruction addition to prevent recurrence
  4. Flag to the Orchestrator for tracking

What it can do on your machine

Read from SKILL.md and the folder at commit 079bd89. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Validation Diagnostics loads about 1.1k tokens when it runs. Until then it costs about 78 tokens; SKILL.md has 354 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from microsoft/haste at commit 079bd89, republished under its MIT licence (© microsoft). 354 words, ~1,093 tokens.

Download SKILL.mdSave it as .claude/skills/validation-diagnostics/SKILL.md (or your agent's skills folder).
name
validation-diagnostics
description
Validation and diagnostic skill for HASTE. Compare planned vs implemented work, generate diagnostic reports, and feed misses back into skill refinement. Use when: 'validate implementation', 'compare to spec', 'diagnostic report', 'drift analysis', 'coverage check', 'implementation review'.
source
HASTE validation practices
domain
quality
level
intermediate
agents
backend-validation, ui-validation, security-validation, orchestrator
created_date
2026-04-27
status
draft

Validation & Diagnostics

Overview

Structured process for comparing planned vs implemented work, generating diagnostic reports, and identifying gaps. Used by validation agents to provide concrete, evidence-based assessments.

Key Concepts

The Trust Problem

Agents will claim work is complete when it isn't. Validation must be:

  • Observable — Based on test output, not agent claims
  • Deterministic — Same input produces same verdict
  • Evidence-based — Every finding references concrete code or test results
  • Structured — Consistent format for easy human review

Patterns & Techniques

Planned vs Implemented Comparison

Step 1: Extract planned items From the spec, issue, or plan, extract a checklist of:

  • Acceptance criteria
  • Required endpoints/functions
  • Expected model fields
  • Required test coverage
  • UI components specified

Step 2: Verify each item For each planned item, check:

  • Does the code exist? (file search, grep)
  • Does it match the specification? (read and compare)
  • Is it tested? (find corresponding test)
  • Does the test pass? (run and capture output)

Step 3: Generate drift report

markdown
## Drift Analysis: [Feature]

| Planned Item | Status | Evidence |
|-------------|--------|----------|
| [spec item] | ✅ Implemented | [file:line] |
| [spec item] | ⚠️ Partial | [what's missing] |
| [spec item] | ❌ Not found | [searched in...] |
| [unplanned] | ⚡ Scope creep | [file:line] |
Diagnostic Report Template
markdown
## Diagnostic Report: [Component/Feature]

### Summary
[1-2 sentence verdict]

### Test Results

[Actual test output — copy/paste, not paraphrased]


### Code Quality
| Metric | Result |
|--------|--------|
| Type hints present | ✅ / ❌ |
| Pydantic models used | ✅ / ❌ |
| Config class used (no hardcoded secrets) | ✅ / ❌ |
| Error handling present | ✅ / ❌ |
| Logger used (not print) | ✅ / ❌ |

### Findings
| # | Severity | Finding | Location | Recommendation |
|---|----------|---------|----------|----------------|
| 1 | [High/Med/Low] | [what] | [file:line] | [fix] |

### Coverage Gaps
[Code paths without tests]

### Verdict
✅ PASS | ⚠️ CONDITIONAL PASS | ❌ FAIL
[Explanation with evidence]
HASTE-Specific Validation Checks
ComponentMust Verify
New API endpointAuth level, Pydantic validation, error codes, CORS
New processorConfig injection, logger usage, error handling
New data modelPydantic BaseModel, field types, validation rules
New data layerAbstract interface compliance, connection handling
New UI componentFluentUI usage, no alt frameworks, responsive
Geospatial codeCRS preservation, COG compliance, GDAL/rasterio usage
Show full SKILL.md (136 more words)Show less
Feedback Loop

When validation reveals a miss:

  1. Document the miss in the diagnostic report
  2. Identify the pattern (was it a convention violation? missing test? spec ambiguity?)
  3. Propose a skill update or instruction addition to prevent recurrence
  4. Flag to the Orchestrator for tracking

Decision Framework

Validation ResultAction
All checks pass, tests green✅ Approve
Minor issues, tests pass⚠️ Conditional — list issues for human review
Tests fail❌ Block — must fix before proceeding
Scope creep detected⚠️ Flag — human decides if extra work is acceptable
Spec ambiguity found⚠️ Flag — needs clarification before validation

Common Pitfalls

  • Accepting "tests passed" without seeing output — Always run and capture
  • Validating only happy path — Check error handling and edge cases
  • Skipping convention checks — HASTE has specific patterns that must be followed
  • Not checking for scope creep — Extra changes can introduce regressions

© microsoft, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .github/skills/validation-diagnostics of microsoft/haste.

Open the folder on GitHubat commit 079bd89

Compare with similar skills

Validation Diagnostics next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Validation Diagnostics compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Validation Diagnostics this skillmicrosoft/haste106—~1.1kAutomated safety check: PassMIT
Building Pydantic AI Agentsdocling-project/docling68k—~2.8kAutomated safety check: PassMIT
Datachain Knowledgedatachain-ai/datachain2.8k—~3kAutomated safety check: PassApache-2.0
Fastcrudbenavlabs/fastcrud1.6k—~5kAutomated safety check: PassMIT
Building Pydantic AI Agentspydantic/pydantic-ai20k—~8kAutomated safety check: PassMIT
Investor Panel Stock Reviewwbh604/UZI-Skill7.1k—~757Automated safety check: PassMIT

Similar skills

  • Building Pydantic AI Agents

    docling-project/docling

    Patterns and tested examples for building agents with Pydantic AI: tools, capabilities, structured output, dependency injection, hooks, YAML specs, streaming and testing.

    68k GitHub stars~2.8k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Datachain Knowledge

    datachain-ai/datachain

    A skill your agent uses whenever datasets, cloud storage buckets, or data pipelines are mentioned — creating, saving, querying, listing, exploring, deleting, or processing data in S3, GCS, Azure…

    2.8k GitHub stars~3k tokensUpdated today
    Knowledge ManagementAuto-check passed
  • Fastcrud

    benavlabs/fastcrud

    A skill your agent uses when building or modifying CRUD endpoints with FastCRUD (the fastcrud PyPI package) in a FastAPI project — covers FastCRUD, crudrouter, EndpointCreator, FilterConfig…

    1.6k GitHub stars~5k tokensUpdated 12 days ago
    Backend & APIsAuto-check passed
  • Building Pydantic AI Agents

    pydantic/pydantic-ai

    Official

    Build AI agents with Pydantic AI — tools, capabilities (including on-demand loading), workspaces, structured output, streaming, testing, and multi-agent patterns.

    20k GitHub stars~8k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Scores a stock through a panel of well-known investor personas, each applying their own method and returning a structured signal, then tallies the votes.

    7.1k GitHub stars~757 tokensUpdated 1 mo ago
    Business, Finance & HRAuto-check passed
  • Fastapi

    Open-TutorAi/open-tutor-ai-CE

    FastAPI best practices and conventions. An agent skill from Open-TutorAi/open-tutor-ai-CE.

    108 GitHub starsUsed in 2 repos~2.6k tokens
    Backend & APIsAuto-check passed

More from microsoft/haste

All 9 skills in this repo
  • API Design

    microsoft/haste

    Official

    Guide for designing and documenting RESTful APIs. An agent skill from microsoft/haste.

    106 GitHub stars~663 tokensUpdated 5 days ago
    Auto-check passed
  • Copilot CLI Modes

    microsoft/haste

    Official

    Comprehensive guide to all Copilot CLI running modes: Interactive, Plan, Autopilot, Fleet, Research, Chronicle, and Delegate.

    106 GitHub stars~4.2k tokensUpdated 5 days ago
    Auto-check passed
  • Debug Test Failures

    microsoft/haste

    Official

    Systematic debugging workflow for test failures. An agent skill from microsoft/haste.

    106 GitHub stars~735 tokensUpdated 5 days ago
    Auto-check passed
  • Dependency Update

    microsoft/haste

    Official

    Guide for safely updating project dependencies. An agent skill from microsoft/haste.

    106 GitHub stars~551 tokensUpdated 5 days ago
    Auto-check passed
  • Official

    Imagery provider adaptation skill for HASTE. An agent skill from microsoft/haste.

    106 GitHub stars~1.3k tokensUpdated 5 days ago
    Auto-check passed
  • Security Analysis

    microsoft/haste

    Official

    Dependabot and security analysis skill for HASTE. An agent skill from microsoft/haste.

    106 GitHub stars~1k tokensUpdated 5 days ago
    Auto-check passed

Works with

Questions about Validation Diagnostics

What does Validation Diagnostics do?

Validation and diagnostic skill for HASTE. An agent skill from microsoft/haste. Validation Diagnostics is an agent skill from microsoft/haste, published by the product's own GitHub organization. Validation and diagnostic skill for HASTE.

When should I use Validation Diagnostics?

Validation Diagnostics fits situations like: : validate implementation; compare to spec; diagnostic report; implementation review.

How do I install Validation Diagnostics in Claude Code?

Run `npx skills add microsoft/haste --skill validation-diagnostics -a claude-code`. Or copy the skill folder (.github/skills/validation-diagnostics in microsoft/haste) into .claude/skills/validation-diagnostics in your project. Claude Code loads it when a task matches its description.

How do I install Validation Diagnostics in Codex?

Run `npx skills add microsoft/haste --skill validation-diagnostics -a codex`. Or copy the skill folder (.github/skills/validation-diagnostics in microsoft/haste) into .agents/skills/validation-diagnostics in your project. Codex loads it when a task matches its description.

Can I use Validation Diagnostics in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add microsoft/haste --skill validation-diagnostics -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/validation-diagnostics, .gemini/skills/validation-diagnostics, .github/skills/validation-diagnostics and .opencode/skills/validation-diagnostics in your project.

What does Validation Diagnostics need to run?

SKILL.md names no scripts, command-line tools or credentials: Validation Diagnostics is instructions for the agent only.

Does Validation Diagnostics access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Validation Diagnostics safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Validation Diagnostics use?

Validation Diagnostics is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Validation Diagnostics use?

About 1.1k tokens (SKILL.md is roughly 4.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Validation Diagnostics?

Skills that share tags, products or a category with Validation Diagnostics: Building Pydantic AI Agents (docling-project/docling, 68k stars), Datachain Knowledge (datachain-ai/datachain, 2.8k stars), Fastcrud (benavlabs/fastcrud, 1.6k stars) and Building Pydantic AI Agents (pydantic/pydantic-ai, 20k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Validation Diagnostics?

microsoft (a GitHub organization, an official publisher) maintains it in microsoft/haste, which has 106 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on October 2, 2026.

Source: microsoft/haste on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.