Search

Testing & QA · OpenAI · For developers

40 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

This skill should be used to "verify OpenHands features", "test the Canvas UI like a user", "drive Agent Canvas", "check a UI change in the real app", "create or update the feature map", "run the…

OpenHands/OpenHands91k—~3.4kAutomated safety check: PassMITtoday
2

Prepares the current DevSpace checkout or worktree for isolated local manual QA, covering QA state seeding, UI asset builds and snapshot resets.

Waishnav/devspace5.2k—~440Automated safety check: PassMIT2 days ago
3

This skill should be used when the user asks to "add an E2E test", "run live E2E", "run mock-LLM tests", "debug Playwright CI", "test the Docker image", or changes tests/e2e, Playwright configs, E2E…

OpenHands/OpenHands91k—~308Automated safety check: PassMITtoday
4

Local test harness for the Weave router: a docker compose stack plus codex exec runs that confirm how Codex requests are routed, translated and marked.

weave-os/router5.6k—~4.7kAutomated safety check: NotesApache-2.0today
5

Checks that ReasoningConfig fields are serialized into the right provider-specific JSON for OpenRouter, Anthropic, GitHub Copilot and Codex requests.

tailcallhq/forgecode7.6k—~1kAutomated safety check: PassApache-2.0today
6

Runs a real Codex CLI session through claude-tap and produces trace evidence and viewer screenshots for pull requests that touch capture, proxying or the viewer.

liaohch3/claude-tap3.3k—~3kAutomated safety check: PassMITyesterday
7

A skill your agent uses when creating, editing, validating, or running agent-qa tests, suites, or hooks.

vostride/agent-qa904—~569Automated safety check: PassUnknown2 mo ago
8

Make @uiverify/vitest (Vitest browser-mode) captures deterministic so component visual tests stop coming back "changed" without a real change (flaky diffs).

FranciscoMoretti/chat-js1.2k—~2.5kAutomated safety check: PassApache-2.0yesterday
9

Inspect a ChatGPT Apps MCP server codebase and generate chatgpt-app-submission.json with app info suggestions, tool hint justifications, test cases, and negative test cases, then report review-check…

nteract/semiotic2.7k—~2.8kAutomated safety check: PassApache-2.0today
10

browser-based page capture and text extraction for public-opinion research.

123321kk/opinion-agent-ultimate107—~631Automated safety check: PassNo licence6 mo ago
11

Record real OpenAI/Anthropic HTTP back-and-forth (requests + responses, including streaming text/event-stream) and print paste-ready Swift fixtures for SwiftAgent unit tests (ReplayHTTPClient) using…

SwiftedMind/SwiftAgent227—~906Automated safety check: PassMIT8 mo ago
12

Look at Kiln's UI in a real browser, and run its end-to-end tests.

Kiln-AI/Kiln5.2k—~1.3kAutomated safety check: PassUnknownyesterday
13

A skill your agent uses when investigating failed agent-qa runs, inspecting artifacts, classifying failures, or comparing recent runs.

vostride/agent-qa904—~394Automated safety check: PassUnknown2 mo ago
14

Triage Chatbook's auto-generated bug-report issues on GitHub by grouping them on their Failure Data signature (the failing Wolfram Language function plus the confirmed expression/pattern), finding…

WolframResearch/Chatbook124—~2.7kAutomated safety check: PassMITyesterday
15

GitHub issues, pull requests, bug reports, scope questions, and support threads.

roryeckel/wyoming_openai219—~2.9kAutomated safety check: PassApache-2.06 days ago
16

Build, launch, and drive shunt — the Claude Code LLM gateway (a Rust/axum Anthropic-Messages proxy).

pleaseai/shunt203—~2.6kAutomated safety check: PassApache-2.02 days ago
17

Adds a new LLM provider implementing LLMProvider interface with call() and stream() methods.

caliber-ai-org/ai-setup1.3k—~2.7kAutomated safety check: PassMIT17 days ago
18

Plan-first, one-time human plan approval; batched execution in either Gated (human confirms between batches) or Auto-loop (continuous run after plan approval); strict task-state updates; automatic…

AlephantAI/AIephant-AI-Agent-Gateway114—~2.7kAutomated safety check: PassGPL-3.04 mo ago
19

A skill your agent uses when the user wants to set up synthetic data generation for the first time, or when sdghub is not yet installed/configured in the current environment.

Red-Hat-AI-Innovation-Team/sdg_hub164—~1.1kAutomated safety check: PassApache-2.0yesterday
20

A skill your agent uses when the agent is building or iterating on a web game (HTML/JS) and needs a reliable development + testing loop: implement small changes, run a Playwright-based test script…

trailofbits/skills-curated513—~2.3kAutomated safety check: NotesCC-BY-SA-4.02 mo ago
21

Use after an agent-qa run has failed and you need to debug, patch, and verify the issue using MCP evidence, logs, artifacts, and local code changes instead of generated fix suggestions.

vostride/agent-qa904—~398Automated safety check: PassUnknown2 mo ago
22

Measure Python SDK coverage or address measured coverage gaps.

openai/openai-agents-python30k—~687Automated safety check: PassMIT2 days ago
23

Fixed workflow for developing MMSP itself — adding or updating model support, and changing its pages.

Prism-Shadow/model-message-stream-protocol113—~7.2kAutomated safety check: PassApache-2.0yesterday
24

Test a pre-built afm binary at any path — runs pre-flight safety checks, then any combination of unit tests, assertions, smart analysis, promptfoo evals, batch validation, OpenAI compat, GPU…

scouzi1966/maclocal-api346—~3.8kAutomated safety check: PassMITyesterday
25

Run the Kiln pre-release smoke test suite plus the standard CI checks (checks.sh), diagnose every prerelease test that broke (and why), and write a clean readable report with recommended actions.

Kiln-AI/Kiln5.2k—~7.8kAutomated safety check: NotesUnknownyesterday
26

A skill your agent uses when the task requires automating a real browser from the terminal (navigation, form filling, snapshots, screenshots, data extraction, UI-flow debugging) via playwright-cli…

trailofbits/skills-curated513—~945Automated safety check: NotesCC-BY-SA-4.02 mo ago
27

Measure JS SDK coverage or address measured coverage gaps. An agent skill from openai/openai-agents-js.

openai/openai-agents-js3.9k—~729Automated safety check: PassMITyesterday
28

Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

gooseworks-ai/goose-skills1.2k—~2.3kAutomated safety check: PassMITtoday
29

Deploys one assigned DynamoGraphDeployment and proves it with an OpenAI-compatible smoke test.

ai-dynamo/dynamo8.3k—~4.7kAutomated safety check: PassApache-2.0today
30

Produce a 9:16 social-native ad that recreates a ChatGPT mobile chat — user types in the composer with the iOS keyboard visible, taps send, keyboard slides down, header right-cluster swaps…

criptogus/agent-evolve-network288—~5kAutomated safety check: PassUnknown2 days ago
31

Create a Mastra project using create-mastra and smoke test the studio in Chrome using Chrome MCP server

mastra-ai/mastra29k—~3.3kAutomated safety check: NotesUnknowntoday
32

Authenticate a browser against a local ChatJS app without OAuth.

FranciscoMoretti/chat-js1.2k—~141Automated safety check: PassApache-2.0yesterday
33

Visual quality assurance: analyze game screenshots for defects, compare against reference, check motion in frame sequences.

RandallLiuXin/GodotMaker550—~1.8kAutomated safety check: PassUnknown23 days ago
34

When you want to integrate an external tool, API, MCP server, or service into a project — the wizard walks you through auth, config, env vars, client wrapper code, example usage, and an optional…

coreyhaines31/makerskills851—~2.9kAutomated safety check: NotesMIT2 days ago
35

Creates a new GAIK software component as an installable Python package.

GAIK-project/gaik-toolkit100—~3kAutomated safety check: PassMIT2 days ago
36

Jev browser automation with indexed actions. An agent skill from openqa-cn/codexqa.

openqa-cn/codexqa152—~1.4kAutomated safety check: NotesMIT8 days ago
37

Connect external agents and MCP hosts (Claude, Claude Desktop, Claude Code, ChatGPT custom MCP apps, Codex, Cursor, Claude Cowork, VS Code GitHub Copilot, Goose, Postman, MCPJam) to an agent-native…

BuilderIO/agent-native7.1k—~7.2kAutomated safety check: NotesNo licenceyesterday
38

Assemble an Apple Notes list video ad from a note + end-card JSON — a frame-accurate fake iPhone screen recording of a short list being typed into Apple Notes (character by character, key pops…

gooseworks-ai/goose-skills1.2k—~1.6kAutomated safety check: PassMIT2 days ago
39

A skill your agent uses when debugging, comparing, or regression-testing LLM / agent calls — when the user wants to capture LLM traffic, see a conversation as a branchable DAG, fork an alternative…

ccplugins/awesome-claude-code-plugins970—~868Automated safety check: PassApache-2.01 mo ago
40
40.Openai Gh Fix CIOfficial

A skill your agent uses when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use gh to inspect checks and logs, summarize failure context, draft a fix plan, and…

trailofbits/skills-curated513—~955Automated safety check: NotesCC-BY-SA-4.02 mo ago