Search
Testing & QA · OpenAI · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | This skill should be used to "verify OpenHands features", "test the Canvas UI like a user", "drive Agent Canvas", "check a UI change in the real app", "create or update the feature map", "run the… | OpenHands/ | 91k | — | ~3.4k | Automated safety check: Pass | MIT | today |
| 2 | Prepares the current DevSpace checkout or worktree for isolated local manual QA, covering QA state seeding, UI asset builds and snapshot resets. | Waishnav/ | 5.2k | — | ~440 | Automated safety check: Pass | MIT | 2 days ago |
| 3 | This skill should be used when the user asks to "add an E2E test", "run live E2E", "run mock-LLM tests", "debug Playwright CI", "test the Docker image", or changes tests/e2e, Playwright configs, E2E… | OpenHands/ | 91k | — | ~308 | Automated safety check: Pass | MIT | today |
| 4 | Local test harness for the Weave router: a docker compose stack plus codex exec runs that confirm how Codex requests are routed, translated and marked. | weave-os/ | 5.6k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | today |
| 5 | Checks that ReasoningConfig fields are serialized into the right provider-specific JSON for OpenRouter, Anthropic, GitHub Copilot and Codex requests. | tailcallhq/ | 7.6k | — | ~1k | Automated safety check: Pass | Apache-2.0 | today |
| 6 | Runs a real Codex CLI session through claude-tap and produces trace evidence and viewer screenshots for pull requests that touch capture, proxying or the viewer. | liaohch3/ | 3.3k | — | ~3k | Automated safety check: Pass | MIT | yesterday |
| 7 | A skill your agent uses when creating, editing, validating, or running agent-qa tests, suites, or hooks. | vostride/ | 904 | — | ~569 | Automated safety check: Pass | Unknown | 2 mo ago |
| 8 | Make @uiverify/vitest (Vitest browser-mode) captures deterministic so component visual tests stop coming back "changed" without a real change (flaky diffs). | FranciscoMoretti/ | 1.2k | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 9 | Inspect a ChatGPT Apps MCP server codebase and generate chatgpt-app-submission.json with app info suggestions, tool hint justifications, test cases, and negative test cases, then report review-check… | nteract/ | 2.7k | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 10 | browser-based page capture and text extraction for public-opinion research. | 123321kk/ | 107 | — | ~631 | Automated safety check: Pass | No licence | 6 mo ago |
| 11 | Record real OpenAI/Anthropic HTTP back-and-forth (requests + responses, including streaming text/event-stream) and print paste-ready Swift fixtures for SwiftAgent unit tests (ReplayHTTPClient) using… | SwiftedMind/ | 227 | — | ~906 | Automated safety check: Pass | MIT | 8 mo ago |
| 12 | 12.Playwright Look at Kiln's UI in a real browser, and run its end-to-end tests. | Kiln-AI/ | 5.2k | — | ~1.3k | Automated safety check: Pass | Unknown | yesterday |
| 13 | A skill your agent uses when investigating failed agent-qa runs, inspecting artifacts, classifying failures, or comparing recent runs. | vostride/ | 904 | — | ~394 | Automated safety check: Pass | Unknown | 2 mo ago |
| 14 | Triage Chatbook's auto-generated bug-report issues on GitHub by grouping them on their Failure Data signature (the failing Wolfram Language function plus the confirmed expression/pattern), finding… | WolframResearch/ | 124 | — | ~2.7k | Automated safety check: Pass | MIT | yesterday |
| 15 | GitHub issues, pull requests, bug reports, scope questions, and support threads. | roryeckel/ | 219 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 16 | 16.Run Shunt Build, launch, and drive shunt — the Claude Code LLM gateway (a Rust/axum Anthropic-Messages proxy). | pleaseai/ | 203 | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 17 | 17.LLM Provider Adds a new LLM provider implementing LLMProvider interface with call() and stream() methods. | caliber-ai-org/ | 1.3k | — | ~2.7k | Automated safety check: Pass | MIT | 17 days ago |
| 18 | Plan-first, one-time human plan approval; batched execution in either Gated (human confirms between batches) or Auto-loop (continuous run after plan approval); strict task-state updates; automatic… | AlephantAI/ | 114 | — | ~2.7k | Automated safety check: Pass | GPL-3.0 | 4 mo ago |
| 19 | 19.Setup Guide A skill your agent uses when the user wants to set up synthetic data generation for the first time, or when sdghub is not yet installed/configured in the current environment. | Red-Hat-AI-Innovation-Team/ | 164 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 20 | A skill your agent uses when the agent is building or iterating on a web game (HTML/JS) and needs a reliable development + testing loop: implement small changes, run a Playwright-based test script… | trailofbits/ | 513 | — | ~2.3k | Automated safety check: Notes | CC-BY-SA-4.0 | 2 mo ago |
| 21 | Use after an agent-qa run has failed and you need to debug, patch, and verify the issue using MCP evidence, logs, artifacts, and local code changes instead of generated fix suggestions. | vostride/ | 904 | — | ~398 | Automated safety check: Pass | Unknown | 2 mo ago |
| 22 | Measure Python SDK coverage or address measured coverage gaps. | openai/ | 30k | — | ~687 | Automated safety check: Pass | MIT | 2 days ago |
| 23 | 23.Mmsp Dev Fixed workflow for developing MMSP itself — adding or updating model support, and changing its pages. | Prism-Shadow/ | 113 | — | ~7.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 24 | Test a pre-built afm binary at any path — runs pre-flight safety checks, then any combination of unit tests, assertions, smart analysis, promptfoo evals, batch validation, OpenAI compat, GPU… | scouzi1966/ | 346 | — | ~3.8k | Automated safety check: Pass | MIT | yesterday |
| 25 | Run the Kiln pre-release smoke test suite plus the standard CI checks (checks.sh), diagnose every prerelease test that broke (and why), and write a clean readable report with recommended actions. | Kiln-AI/ | 5.2k | — | ~7.8k | Automated safety check: Notes | Unknown | yesterday |
| 26 | A skill your agent uses when the task requires automating a real browser from the terminal (navigation, form filling, snapshots, screenshots, data extraction, UI-flow debugging) via playwright-cli… | trailofbits/ | 513 | — | ~945 | Automated safety check: Notes | CC-BY-SA-4.0 | 2 mo ago |
| 27 | Measure JS SDK coverage or address measured coverage gaps. An agent skill from openai/openai-agents-js. | openai/ | 3.9k | — | ~729 | Automated safety check: Pass | MIT | yesterday |
| 28 | Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard… | gooseworks-ai/ | 1.2k | — | ~2.3k | Automated safety check: Pass | MIT | today |
| 29 | Deploys one assigned DynamoGraphDeployment and proves it with an OpenAI-compatible smoke test. | ai-dynamo/ | 8.3k | — | ~4.7k | Automated safety check: Pass | Apache-2.0 | today |
| 30 | Produce a 9:16 social-native ad that recreates a ChatGPT mobile chat — user types in the composer with the iOS keyboard visible, taps send, keyboard slides down, header right-cluster swaps… | criptogus/ | 288 | — | ~5k | Automated safety check: Pass | Unknown | 2 days ago |
| 31 | 31.Smoke Test Create a Mastra project using create-mastra and smoke test the studio in Chrome using Chrome MCP server | mastra-ai/ | 29k | — | ~3.3k | Automated safety check: Notes | Unknown | today |
| 32 | Authenticate a browser against a local ChatJS app without OAuth. | FranciscoMoretti/ | 1.2k | — | ~141 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 33 | 33.Visual QA Visual quality assurance: analyze game screenshots for defects, compare against reference, check motion in frame sequences. | RandallLiuXin/ | 550 | — | ~1.8k | Automated safety check: Pass | Unknown | 23 days ago |
| 34 | 34.Toolify When you want to integrate an external tool, API, MCP server, or service into a project — the wizard walks you through auth, config, env vars, client wrapper code, example usage, and an optional… | coreyhaines31/ | 851 | — | ~2.9k | Automated safety check: Notes | MIT | 2 days ago |
| 35 | Creates a new GAIK software component as an installable Python package. | GAIK-project/ | 100 | — | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 36 | Jev browser automation with indexed actions. An agent skill from openqa-cn/codexqa. | openqa-cn/ | 152 | — | ~1.4k | Automated safety check: Notes | MIT | 8 days ago |
| 37 | Connect external agents and MCP hosts (Claude, Claude Desktop, Claude Code, ChatGPT custom MCP apps, Codex, Cursor, Claude Cowork, VS Code GitHub Copilot, Goose, Postman, MCPJam) to an agent-native… | BuilderIO/ | 7.1k | — | ~7.2k | Automated safety check: Notes | No licence | yesterday |
| 38 | Assemble an Apple Notes list video ad from a note + end-card JSON — a frame-accurate fake iPhone screen recording of a short list being typed into Apple Notes (character by character, key pops… | gooseworks-ai/ | 1.2k | — | ~1.6k | Automated safety check: Pass | MIT | 2 days ago |
| 39 | 39.Forkmind A skill your agent uses when debugging, comparing, or regression-testing LLM / agent calls — when the user wants to capture LLM traffic, see a conversation as a branchable DAG, fork an alternative… | ccplugins/ | 970 | — | ~868 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 40 | A skill your agent uses when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use gh to inspect checks and logs, summarize failure context, draft a fix plan, and… | trailofbits/ | 513 | — | ~955 | Automated safety check: Notes | CC-BY-SA-4.0 | 2 mo ago |