Official agent skill

Analyze Comparison Tests

by microsoft in microsoft/GitHub-Copilot-for-Azure

Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run.

OfficialMITAuto-check passed

Install Analyze Comparison Tests

skills CLI
$ npx skills add microsoft/GitHub-Copilot-for-Azure --skill analyze-comparison-tests -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install microsoft/GitHub-Copilot-for-Azure analyze-comparison-tests --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/microsoft/GitHub-Copilot-for-Azure.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.github/skills/analyze-comparison-tests .claude/skills/analyze-comparison-tests && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
analyze-comparison-tests
GitHub stars
255
Token cost
~579 tokens
SKILL.md length
199 words
Files
2 (incl. references)
Skills in repo
56
Repo updated
First seen
Licence
MIT

At a glance

Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run.

  • Calls npm

What it does

Analyze Comparison Tests is an agent skill from microsoft/GitHub-Copilot-for-Azure, published by the product's own GitHub organization. Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run. WHEN TO USE: collect comparison test artifacts

Its SKILL.md is about 580 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/report-template.md`).

The repository describes itself as: GitHub Copilot for Azure. The licence is MIT.

Example prompts

  • “Use the analyze-comparison-tests skill to collect comparison test run artifacts and answers the user's questions based on the trajectories of each run”
  • “/analyze-comparison-tests”

What it can do on your machine

Read from SKILL.md and the folder at commit d8f4f4e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Analyze Comparison Tests loads about 579 tokens when it runs, and up to ~634 if it reads all its reference files. Until then it costs about 46 tokens; SKILL.md has 199 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~579
With references · SKILL.md plus every file in references/, read only if the agent opens them
~634

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from microsoft/GitHub-Copilot-for-Azure at commit d8f4f4e, republished under its MIT licence (© microsoft). 199 words, ~579 tokens.

Download SKILL.mdSave it as .claude/skills/analyze-comparison-tests/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
analyze-comparison-tests
description
Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run. WHEN TO USE: collect comparison test artifacts
license
MIT
metadata.author
Microsoft
metadata.version
1.0.0

Steps

  1. Collect run artifacts

Execute the collect-artifacts script to download the test run artifacts.

The user must provide an JSON file to correlate each comparison test run with the GitHub Actions run. The script expects one input argument as the path to this JSON file. The JSON input is supposed to be the JSON output when queuing the comparison test runs using the npm run compare:run command.

bash
cd tests/
npm run compare:collect -- input.json

The collect-artifacts script will download the test run artifacts to a directory named comparison-artifacts in the current working directory. Before executing the script, check if there is already such an directory. If so, skip executing the script and proceed to step 2.

  1. Extract insights

The downloaded artifacts will have the following folder structure:

text
comparison-artifacts/
├── <branch-name>/
│   ├── <stimulus-name-1>/
│   │   ├── <model>-with-skill/
│   │   │   ├── agent-metadata-<date-string-1>.md
│   │   │   ├── agent-metadata-<date-string-2>.md
│   │   │   └── ...
│   │   └── <model>-without-skill/
│   │       ├── agent-metadata-<date-string-1>.md
│   │       ├── agent-metadata-<date-string-2>.md
│   │       └── ...
│   └── <stimulus-name-2>/
│       ├── <model>-with-skill/
│       │   └── agent-metadata-*.md
│       └── <model>-without-skill/
│           └── agent-metadata-*.md
└── <branch-name-2>/
    └── ...

Each <branch-name>/<stimulus-name>/<model>-with-skill or <branch-name>/<stimulus-name>/<model>-without-skill directory contains the test run trajectories for that stimulus and model on that branch, with or without skills. Each trajectory is a markdown file that records user prompts, tool call requests, tool execution results, assistant responses that happened during the run. It also contains statistics such as token usage and turns. Based on the trajectories, answer the user's questions for each test run. Generate a report following the report-template to show your answers.

© microsoft, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in .github/skills/analyze-comparison-tests of microsoft/GitHub-Copilot-for-Azure.

  • SKILL.md
  • references/report-template.md

Open the folder on GitHubat commit d8f4f4e

Compare with similar skills

Analyze Comparison Tests next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Analyze Comparison Tests compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Analyze Comparison Tests this skillmicrosoft/GitHub-Copilot-for-Azure255—~579Automated safety check: PassMIT
Agent Collective Intelligence Coordinatorruvnet/ruflo74k3 repos~1kAutomated safety check: PassMIT
Gaia Architecture Comparisonruvnet/ruflo74k—~1.3kAutomated safety check: NotesMIT
Postman Collection Generatorsickn33/agentic-awesome-skills47k1 repos~1.4kAutomated safety check: PassMIT
Logseq Answer Machinelogseq/logseq45k—~1.2kAutomated safety check: WarnAGPL-3.0
Dossier Collectruvnet/ruflo74k—~1.1kAutomated safety check: NotesMIT

Similar skills

  • Agent skill for collective-intelligence-coordinator - invoke with $agent-collective-intelligence-coordinator

    74k GitHub starsUsed in 3 repos~1k tokens
    Auto-check passed
  • Side-by-side comparison of ruflo vs HAL vs other GAIA harnesses — capability gaps, design decisions, and improvement roadmap

    74k GitHub stars~1.3k tokensUpdated today
    DevelopmentAuto-check: notes
  • Postman Collection Generator

    sickn33/agentic-awesome-skills

    Generate complete, import-ready Postman Collection v2.1 JSON files from natural language API descriptions or cURL commands.

    47k GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check passed
  • Answer user questions about the Logseq repository by researching source code, docs, tests, runtime behavior, and local tools.

    45k GitHub stars~1.2k tokensUpdated today
    Knowledge ManagementAuto-check: warnings
  • Dossier Collect

    ruvnet/ruflo

    Build a graph-structured dossier on a seed entity via parallel fan-out + recursive expansion across web, memory, knowledge-graph, codebase, ADR index, and git intel

    74k GitHub stars~1.1k tokensUpdated today
    DevelopmentAuto-check: notes
  • Collecting Indicators Of Compromise

    mukul975/Anthropic-Cybersecurity-Skills

    Systematically collects, categorizes, and distributes indicators of compromise (IOCs) during and after security incidents to enable detection, blocking, and threat intelligence sharing.

    34k GitHub stars~2.7k tokensUpdated 1 mo ago
    SecurityAuto-check passed

More from microsoft/GitHub-Copilot-for-Azure

All 56 skills in this repo
  • Capacity

    microsoft/GitHub-Copilot-for-Azure

    Official

    Discovers available Azure OpenAI model capacity across regions and projects.

    255 GitHub starsUsed in 2 repos~1.7k tokens
    Auto-check passed
  • Deploy Model

    microsoft/GitHub-Copilot-for-Azure

    Official

    Unified Azure OpenAI model deployment skill with intelligent intent-based routing.

    255 GitHub starsUsed in 1 repo~1.8k tokens
    Auto-check passed
  • Entra Agent Id

    microsoft/GitHub-Copilot-for-Azure

    Official

    Provision Microsoft Entra Agent Identity Blueprints, BlueprintPrincipals, and per-instance Agent Identities via Microsoft Graph, and configure OAuth 2.0 token exchange (fmipath, OBO, cross-tenant)…

    255 GitHub starsUsed in 3 repos~4k tokens
    Auto-check passed
  • Microsoft Foundry

    microsoft/GitHub-Copilot-for-Azure

    Official

    Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end.

    255 GitHub starsUsed in 1 repo~6.7k tokens
    Auto-check passed
  • Azure Storage

    microsoft/GitHub-Copilot-for-Azure

    Official

    Azure Storage Services including Blob Storage, File Shares, Queue Storage, Table Storage, and Data Lake.

    255 GitHub starsUsed in 2 repos~1.3k tokens
    Auto-check passed
  • Azure Diagnostics

    microsoft/GitHub-Copilot-for-Azure

    Official

    Debug Azure production issues on Azure using AppLens, Azure Monitor, resource health, and safe triage.

    255 GitHub starsUsed in 1 repo~1.6k tokens
    Auto-check passed

Questions about Analyze Comparison Tests

What does Analyze Comparison Tests do?

Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run. Analyze Comparison Tests is an agent skill from microsoft/GitHub-Copilot-for-Azure, published by the product's own GitHub organization. Collects comparison test run artifacts and answers the user's questions based on the trajectories of each run.

How do I install Analyze Comparison Tests in Claude Code?

Run `npx skills add microsoft/GitHub-Copilot-for-Azure --skill analyze-comparison-tests -a claude-code`. Or copy the skill folder (.github/skills/analyze-comparison-tests in microsoft/GitHub-Copilot-for-Azure) into .claude/skills/analyze-comparison-tests in your project. Claude Code loads it when a task matches its description.

How do I install Analyze Comparison Tests in Codex?

Run `npx skills add microsoft/GitHub-Copilot-for-Azure --skill analyze-comparison-tests -a codex`. Or copy the skill folder (.github/skills/analyze-comparison-tests in microsoft/GitHub-Copilot-for-Azure) into .agents/skills/analyze-comparison-tests in your project. Codex loads it when a task matches its description.

Can I use Analyze Comparison Tests in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add microsoft/GitHub-Copilot-for-Azure --skill analyze-comparison-tests -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/analyze-comparison-tests, .gemini/skills/analyze-comparison-tests, .github/skills/analyze-comparison-tests and .opencode/skills/analyze-comparison-tests in your project.

What does Analyze Comparison Tests need to run?

Going by SKILL.md and its folder, Analyze Comparison Tests needs the command-line tools its instructions call (npm).

Does Analyze Comparison Tests access the network?

SKILL.md contains no URLs. Its commands use npm, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Analyze Comparison Tests safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Analyze Comparison Tests use?

Analyze Comparison Tests is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Analyze Comparison Tests use?

About 579 tokens (SKILL.md is roughly 2.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 55 tokens, read only when the agent opens those files.

What are the alternatives to Analyze Comparison Tests?

Skills that share tags, products or a category with Analyze Comparison Tests: Agent Collective Intelligence Coordinator (ruvnet/ruflo, 74k stars), Gaia Architecture Comparison (ruvnet/ruflo, 74k stars), Postman Collection Generator (sickn33/agentic-awesome-skills, 47k stars) and Logseq Answer Machine (logseq/logseq, 45k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Analyze Comparison Tests?

microsoft (a GitHub organization, an official publisher) maintains it in microsoft/GitHub-Copilot-for-Azure, which has 255 GitHub stars. The repository holds 56 skills in this directory. The repository was last updated on October 7, 2026.

Source: microsoft/GitHub-Copilot-for-Azure on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.