Search

Testing & QA · OpenAI

41 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

This skill should be used to "verify OpenHands features", "test the Canvas UI like a user", "drive Agent Canvas", "check a UI change in the real app", "create or update the feature map", "run the…

OpenHands/OpenHands91k—~3.4kAutomated safety check: PassMITtoday
2

Prepares the current DevSpace checkout or worktree for isolated local manual QA, covering QA state seeding, UI asset builds and snapshot resets.

Waishnav/devspace5.2k—~440Automated safety check: PassMIT2 days ago
3

This skill should be used when the user asks to "add an E2E test", "run live E2E", "run mock-LLM tests", "debug Playwright CI", "test the Docker image", or changes tests/e2e, Playwright configs, E2E…

OpenHands/OpenHands91k—~308Automated safety check: PassMITtoday
4

Local test harness for the Weave router: a docker compose stack plus codex exec runs that confirm how Codex requests are routed, translated and marked.

weave-os/router5.6k—~4.7kAutomated safety check: NotesApache-2.0today
5

Checks that ReasoningConfig fields are serialized into the right provider-specific JSON for OpenRouter, Anthropic, GitHub Copilot and Codex requests.

tailcallhq/forgecode7.6k—~1kAutomated safety check: PassApache-2.0today
6

Runs a real Codex CLI session through claude-tap and produces trace evidence and viewer screenshots for pull requests that touch capture, proxying or the viewer.

liaohch3/claude-tap3.3k—~3kAutomated safety check: PassMITyesterday
7

A skill your agent uses when creating, editing, validating, or running agent-qa tests, suites, or hooks.

vostride/agent-qa904—~569Automated safety check: PassUnknown2 mo ago
8

Make @uiverify/vitest (Vitest browser-mode) captures deterministic so component visual tests stop coming back "changed" without a real change (flaky diffs).

FranciscoMoretti/chat-js1.2k—~2.5kAutomated safety check: PassApache-2.0yesterday
9

Inspect a ChatGPT Apps MCP server codebase and generate chatgpt-app-submission.json with app info suggestions, tool hint justifications, test cases, and negative test cases, then report review-check…

nteract/semiotic2.7k—~2.8kAutomated safety check: PassApache-2.0today
10

browser-based page capture and text extraction for public-opinion research.

123321kk/opinion-agent-ultimate107—~631Automated safety check: PassNo licence6 mo ago
11

Record real OpenAI/Anthropic HTTP back-and-forth (requests + responses, including streaming text/event-stream) and print paste-ready Swift fixtures for SwiftAgent unit tests (ReplayHTTPClient) using…

SwiftedMind/SwiftAgent227—~906Automated safety check: PassMIT8 mo ago
12

Look at Kiln's UI in a real browser, and run its end-to-end tests.

Kiln-AI/Kiln5.2k—~1.3kAutomated safety check: PassUnknownyesterday
13

A skill your agent uses when investigating failed agent-qa runs, inspecting artifacts, classifying failures, or comparing recent runs.

vostride/agent-qa904—~394Automated safety check: PassUnknown2 mo ago
14

Triage Chatbook's auto-generated bug-report issues on GitHub by grouping them on their Failure Data signature (the failing Wolfram Language function plus the confirmed expression/pattern), finding…

WolframResearch/Chatbook124—~2.7kAutomated safety check: PassMITyesterday
15

GitHub issues, pull requests, bug reports, scope questions, and support threads.

roryeckel/wyoming_openai219—~2.9kAutomated safety check: PassApache-2.06 days ago
16

Build, launch, and drive shunt — the Claude Code LLM gateway (a Rust/axum Anthropic-Messages proxy).

pleaseai/shunt203—~2.6kAutomated safety check: PassApache-2.02 days ago
17

Adds a new LLM provider implementing LLMProvider interface with call() and stream() methods.

caliber-ai-org/ai-setup1.3k—~2.7kAutomated safety check: PassMIT17 days ago
18

Plan-first, one-time human plan approval; batched execution in either Gated (human confirms between batches) or Auto-loop (continuous run after plan approval); strict task-state updates; automatic…

AlephantAI/AIephant-AI-Agent-Gateway114—~2.7kAutomated safety check: PassGPL-3.04 mo ago
19

A skill your agent uses when the user wants to set up synthetic data generation for the first time, or when sdghub is not yet installed/configured in the current environment.

Red-Hat-AI-Innovation-Team/sdg_hub164—~1.1kAutomated safety check: PassApache-2.0yesterday
20

A skill your agent uses when the agent is building or iterating on a web game (HTML/JS) and needs a reliable development + testing loop: implement small changes, run a Playwright-based test script…

trailofbits/skills-curated513—~2.3kAutomated safety check: NotesCC-BY-SA-4.02 mo ago
21

Use after an agent-qa run has failed and you need to debug, patch, and verify the issue using MCP evidence, logs, artifacts, and local code changes instead of generated fix suggestions.

vostride/agent-qa904—~398Automated safety check: PassUnknown2 mo ago
22

Measure Python SDK coverage or address measured coverage gaps.

openai/openai-agents-python30k—~687Automated safety check: PassMIT2 days ago
23

Fixed workflow for developing MMSP itself — adding or updating model support, and changing its pages.

Prism-Shadow/model-message-stream-protocol113—~7.2kAutomated safety check: PassApache-2.0yesterday
24

Test a pre-built afm binary at any path — runs pre-flight safety checks, then any combination of unit tests, assertions, smart analysis, promptfoo evals, batch validation, OpenAI compat, GPU…

scouzi1966/maclocal-api346—~3.8kAutomated safety check: PassMITyesterday
25

Run the Kiln pre-release smoke test suite plus the standard CI checks (checks.sh), diagnose every prerelease test that broke (and why), and write a clean readable report with recommended actions.

Kiln-AI/Kiln5.2k—~7.8kAutomated safety check: NotesUnknownyesterday
26

A skill your agent uses when the task requires automating a real browser from the terminal (navigation, form filling, snapshots, screenshots, data extraction, UI-flow debugging) via playwright-cli…

trailofbits/skills-curated513—~945Automated safety check: NotesCC-BY-SA-4.02 mo ago
27

Measure JS SDK coverage or address measured coverage gaps. An agent skill from openai/openai-agents-js.

openai/openai-agents-js3.9k—~729Automated safety check: PassMITyesterday
28

Assemble a ChatGPT chat-reveal video ad from a thread + timeline JSON — one continuous Playwright recording of a ChatGPT mobile chat (user types with the iOS keyboard up → taps send → keyboard…

gooseworks-ai/goose-skills1.2k—~2.3kAutomated safety check: PassMITtoday
29

Deploys one assigned DynamoGraphDeployment and proves it with an OpenAI-compatible smoke test.

ai-dynamo/dynamo8.3k—~4.7kAutomated safety check: PassApache-2.0today
30

Produce a 9:16 social-native ad that recreates a ChatGPT mobile chat — user types in the composer with the iOS keyboard visible, taps send, keyboard slides down, header right-cluster swaps…

criptogus/agent-evolve-network288—~5kAutomated safety check: PassUnknown2 days ago
31

当用户要用本机已登录 Chrome 操作或验收网页(含测试自己的站、E2E、截图)、读取登录后台、抓取无 API 的数据、填表,或排查 OpenCLI doctor、会话、adapter 故障,或明确提到 opencli、浏览器自动化、标签页被抢时使用。公开文本先用 HTTP;话题调研交 agent-reach;SEO 与选词决策交 rankup;外链选点、提交与…

yan-labs/yan-skills213—~6.8kAutomated safety check: PassMITyesterday
32

Create a Mastra project using create-mastra and smoke test the studio in Chrome using Chrome MCP server

mastra-ai/mastra29k—~3.3kAutomated safety check: NotesUnknowntoday
33

Authenticate a browser against a local ChatJS app without OAuth.

FranciscoMoretti/chat-js1.2k—~141Automated safety check: PassApache-2.0yesterday
34

Visual quality assurance: analyze game screenshots for defects, compare against reference, check motion in frame sequences.

RandallLiuXin/GodotMaker550—~1.8kAutomated safety check: PassUnknown23 days ago
35

When you want to integrate an external tool, API, MCP server, or service into a project — the wizard walks you through auth, config, env vars, client wrapper code, example usage, and an optional…

coreyhaines31/makerskills851—~2.9kAutomated safety check: NotesMIT2 days ago
36

Creates a new GAIK software component as an installable Python package.

GAIK-project/gaik-toolkit100—~3kAutomated safety check: PassMIT2 days ago
37

Jev browser automation with indexed actions. An agent skill from openqa-cn/codexqa.

openqa-cn/codexqa152—~1.4kAutomated safety check: NotesMIT8 days ago
38

Connect external agents and MCP hosts (Claude, Claude Desktop, Claude Code, ChatGPT custom MCP apps, Codex, Cursor, Claude Cowork, VS Code GitHub Copilot, Goose, Postman, MCPJam) to an agent-native…

BuilderIO/agent-native7.1k—~7.2kAutomated safety check: NotesNo licenceyesterday
39

Assemble an Apple Notes list video ad from a note + end-card JSON — a frame-accurate fake iPhone screen recording of a short list being typed into Apple Notes (character by character, key pops…

gooseworks-ai/goose-skills1.2k—~1.6kAutomated safety check: PassMIT2 days ago
40

A skill your agent uses when debugging, comparing, or regression-testing LLM / agent calls — when the user wants to capture LLM traffic, see a conversation as a branchable DAG, fork an alternative…

ccplugins/awesome-claude-code-plugins970—~868Automated safety check: PassApache-2.01 mo ago
41
41.Openai Gh Fix CIOfficial

A skill your agent uses when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use gh to inspect checks and logs, summarize failure context, draft a fix plan, and…

trailofbits/skills-curated513—~955Automated safety check: NotesCC-BY-SA-4.02 mo ago