CodexBar Live QA
steipete/CodexBar
Runs live QA for the CodexBar app: provider usage matrix checks through its packaged CLI, config validation and menu checks, with 1Password-backed credentials handled safely.
Spins up an isolated Omnigent server, runner and mock model to prove a user-facing behavior or bug fix with recorded evidence instead of reasoning from code.
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install omnigent-ai/omnigent verify-omnigent --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/feature-map/skills/verify-omnigent .claude/skills/verify-omnigent && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .claude/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigentType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install omnigent-ai/omnigent verify-omnigent --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .agents/skills && cp -r skills-src/feature-map/skills/verify-omnigent .agents/skills/verify-omnigent && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .agents/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install omnigent-ai/omnigent verify-omnigent --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/feature-map/skills/verify-omnigent .cursor/skills/verify-omnigent && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .cursor/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/omnigent-ai/omnigent.git --path feature-map/skills/verify-omnigent--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install omnigent-ai/omnigent verify-omnigent --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/feature-map/skills/verify-omnigent .gemini/skills/verify-omnigent && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .gemini/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install omnigent-ai/omnigent verify-omnigentInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .github/skills && cp -r skills-src/feature-map/skills/verify-omnigent .github/skills/verify-omnigent && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .github/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add omnigent-ai/omnigent --skill verify-omnigent -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install omnigent-ai/omnigent verify-omnigent --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/omnigent-ai/omnigent.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/feature-map/skills/verify-omnigent .opencode/skills/verify-omnigent && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "verify-omnigent" agent skill from https://github.com/omnigent-ai/omnigent/tree/main/feature-map/skills/verify-omnigent into .opencode/skills/verify-omnigent/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "verify-omnigent", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
verify-omnigentSpins up an isolated Omnigent server, runner and mock model to prove a user-facing behavior or bug fix with recorded evidence instead of reasoning from code.
This skill has two parts: an isolated instance wrapping a repro-environment module that starts a server, runner and mock model server on private ports with their own config, data, Claude and Codex directories, never touching the real home install, a running host daemon, or another developer's server; and a feature map listing every user-facing feature's entry points, the tests that drive them, and known traps, where a fix only counts as verified once every listed entry point for its feature has proof.
Starting an instance installs dependencies, Chromium, and builds the web UI, then launches the instance and waits roughly ten seconds until the server, runner and mock model all answer; the instance writes its root path so later shell sessions can find it, and it stops itself automatically after a lease that defaults to 90 minutes. A doctor command checks readiness, supervisor health, and whether the runner and model server still answer, and should run before the first drive and after any failed one.
Readiness does not include checking that the browser actually launches, and native-terminal journeys need their own CLI and terminal prerequisites such as tmux beyond just a browser install; in CI, a repro-environment exec command reuses the environment the workflow already started instead of spinning up a second one.
2 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 2e1cd15. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/, which the agent can run.
Shell commands in SKILL.md call:
uvpnpmpythonFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use uv and pnpm, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Verify Omnigent End-to-End loads about 1.6k tokens when it runs. Until then it costs about 119 tokens; SKILL.md has 799 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from omnigent-ai/omnigent at commit 2e1cd15, republished under its Apache-2.0 licence (© omnigent-ai). 799 words, ~1,631 tokens.
.claude/skills/verify-omnigent/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.Use this skill to see a behavior happen in the real app and to prove a change, not to reason about it from code. It has two parts:
scripts/verify-env wraps
python -m dev.repro_env: a server, runner, and mock model server on private
ports, with their own config, data, Claude, and Codex directories. It never
touches ~/.omnigent, a running host daemon, or another developer server.Run all commands from the repository root. Put scripts/ on your path or call
feature-map/skills/verify-omnigent/scripts/verify-env directly.
Install dependencies, Chromium, and build the web UI for the checkout:
uv sync --frozen --extra all --group test
uv run --no-sync playwright install --with-deps chromium
pnpm install --frozen-lockfile --filter web && pnpm --filter web run buildStart an instance and load its paths:
verify-env start # waits until the runner is online, about 10 seconds
eval "$(verify-env paths)"start writes VERIFY_ROOT to .omnigent/verify/current, so later shells
find the same instance. The instance stops itself after its lease (default
90 minutes; --lease SECONDS, 60 to 21600).
Ready means the server answers, the runner reports online, and the mock model
server answers. start fails with the environment's error and log path
otherwise.
Ready does not check browser launch. Use the Chromium version installed by
this checkout's Playwright package; if the environment provides it through
PLAYWRIGHT_BROWSERS_PATH, keep that path available to the test process.
--ui-skip-build reuses the existing bundle, so rebuild after frontend edits.
Native-terminal journeys also need the relevant CLI and terminal prerequisites
(such as tmux); a browser installation alone does not provide them.
In CI, the repro workflow already runs this environment. Use
python -m dev.repro_env exec -- ... there instead of starting another.
Run verify-env doctor before the first drive, after any failed drive, and
whenever something looks off. It checks that the instance is ready, that its
supervisor is alive, that the runner and model server answer, and it warns when
the checkout has moved since launch. A warning about a moved checkout means the
instance runs old code: stop it and start a new one.
Open the matching file in feature map and list every entry point for the behavior in question.
For each entry point, run the named test through the instance, recording on:
verify-env run -- python -m pytest <tests/...py::test_name> \
--ui-skip-build --video=on --screenshot=on \
--output="$VERIFY_EVIDENCE/<feature>"Tests under tests/browser_ui/ need no instance:
uv run pytest <test> --browser-ui-skip-build --video=on --output=....
Tests the feature file marks "own environment" also run with plain
uv run pytest. Server/transport/component tests use plain pytest without
browser recording or build flags.
For an entry point with no test, hand-drive it with Playwright against
OMNIGENT_REPRO_SERVER_URL inside verify-env run -- python <script>, and
script model replies with the mock helpers named in the feature map README.
To reproduce a bug, run the same journey on the unfixed code first and keep that evidence; then run it again on the fix.
A test that fails before reaching the user action has a setup failure, not a
reproduction. Check its browser error and server/runner logs. When a test starts
its own server or runner from inside an agent, check whether it inherited the
parent's OMNIGENT_RUNNER_*, RUNNER_SERVER_URL, or
OMNIGENT_PROCESS_LOG_FILE. Use an isolated child environment for that test;
do not change the controlling agent's environment. Record setup failures,
skipped variants, and unavailable credentials separately from behavior results.
Everything under $VERIFY_EVIDENCE is the proof, one directory per feature.
Instance logs, the database, and the mock model's recorded requests stay in
$VERIFY_ENV. For each claim, record the checkout commit, installed harness versions,
feature file, entry point ID, command, and resulting artifact. Keep source
review, component tests, real process checks, browser drives, and live provider
checks distinct. A map contract pass verifies references and structure; it is
not evidence that the journeys ran. The proof standards are in the
feature map README. Mock runs prove
Omnigent's integration with Claude and Codex, not a live vendor model.
verify-env stop stops only this instance's supervisor, which stops its own
server, runner, and model processes and saves the model's request log. It never
deletes $VERIFY_ROOT, so the evidence survives; remove old roots under
.omnigent/verify/ or /tmp/verify-omnigent.* yourself once the evidence is
attached. Never kill Omnigent processes by name: a developer's own server and
host daemon may be running on the same machine.
scripts/verify-env is the only helper. Its subcommands are start [--lease SECONDS], doctor, run -- COMMAND..., paths, and stop; run it with no
arguments for usage.
On macOS the helper places the instance under /tmp/verify-omnigent.* when the
checkout path is long. macOS limits Unix socket paths to 104 bytes, and the
environment's socket relay only works around long paths on Linux.
When a reviewer says a change missed a surface, add that surface to the feature file in the same change. The checks and the weekly upkeep job are described in Keeping the map current.
© omnigent-ai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (scripts) in feature-map/skills/verify-omnigent of omnigent-ai/omnigent.
Open the folder on GitHubat commit 2e1cd15
Verify Omnigent End-to-End next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Verify Omnigent End-to-End this skillomnigent-ai/omnigent | 11k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | |
| CodexBar Live QAsteipete/CodexBar | 22k | — | ~1.2k | Automated safety check: Pass | MIT | |
| Senpi Agent QA Harnesscode-yeongyu/senpi | 472 | — | ~2.7k | Automated safety check: Notes | MIT | |
| tmux Real User TestingQwenLM/qwen-code | 28k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | |
| Glance TestDebugBase/glance | 156 | — | ~827 | Automated safety check: Pass | MIT | |
| Create a Verification Skillcursor/plugins | 10k | 8 repos | ~1.5k | Automated safety check: Pass | None |
steipete/CodexBar
Runs live QA for the CodexBar app: provider usage matrix checks through its packaged CLI, config validation and menu checks, with 1Password-backed credentials handled safely.
code-yeongyu/senpi
Checks changes to the senpi coding agent by driving the real CLI from source in an isolated sandbox, over RPC, terminal UI, mock model and CLI smoke channels.
QwenLM/qwen-code
Drives Qwen Code in a real tmux session the way a user would and saves a readable step-by-step transcript of each screen for maintainers to review.
DebugBase/glance
Run E2E browser tests on any web application using Glance MCP.
cursor/plugins
Generates a project-local skill that launches your app, exercises a feature the way a user would and captures evidence, for web, CLI, API or desktop projects.
twentyhq/twenty
Browser QA for a pull request against a running Twenty app: scenarios drawn from the diff, run in a real browser, checked in the database and logs, and closed with a verdict and report.
omnigent-ai/omnigent
Brings up the Omnigent server and Postgres as a Docker compose stack on any Docker host, and covers the Dockerfile's runtime and host build targets for extending it to a new platform.
omnigent-ai/omnigent
Scans Python agent code for framework imports and recommends the matching Omnigent executor type, or says when the framework is not natively supported yet.
omnigent-ai/omnigent
Runs the Omnigent load test with real hosts and multi-turn sessions against a mocked LLM, then explains the latency results from summary.md.
omnigent-ai/omnigent
Spins up a local Omnigent server and exercises the Antigravity (Gemini) SDK harness end to end: building agents, running real turns, smoke tests and bug-bashing.
omnigent-ai/omnigent
Gives patterns for generating a minimal, valid Omnigent agent directory: the config.yaml fields, the right executor type, and the files each agent needs.
omnigent-ai/omnigent
Spin up a live local Omnigent server and exercise the GitHub Copilot SDK harness end-to-end — build copilot agents, run real turns, smoke-test, and bug-bash.
Works with
Categories
Spins up an isolated Omnigent server, runner and mock model to prove a user-facing behavior or bug fix with recorded evidence instead of reasoning from code. This skill has two parts: an isolated instance wrapping a repro-environment module that starts a server, runner and mock model server on private ports with their own config, data, Claude and Codex directories, never touching the real home install, a running host daemon, or another developer's server; and a feature map listing every user-facing feature's entry points, the tests that drive them, and known traps, where a fix only counts as verified once every listed entry point for its feature has proof.
Verify Omnigent End-to-End fits situations like: reproducing a user-facing bug before attempting a fix; proving a fix works across every entry point for its feature; checking whether every native harness terminal journey is covered.
Run `npx skills add omnigent-ai/omnigent --skill verify-omnigent -a claude-code`. Or copy the skill folder (feature-map/skills/verify-omnigent in omnigent-ai/omnigent) into .claude/skills/verify-omnigent in your project. Claude Code loads it when a task matches its description.
Run `npx skills add omnigent-ai/omnigent --skill verify-omnigent -a codex`. Or copy the skill folder (feature-map/skills/verify-omnigent in omnigent-ai/omnigent) into .agents/skills/verify-omnigent in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add omnigent-ai/omnigent --skill verify-omnigent -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/verify-omnigent, .gemini/skills/verify-omnigent, .github/skills/verify-omnigent and .opencode/skills/verify-omnigent in your project.
Going by SKILL.md and its folder, Verify Omnigent End-to-End needs the command-line tools its instructions call (uv, pnpm and python). Our summary lists: uv; pnpm; Playwright with Chromium; tmux (for native-terminal journeys).
SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Verify Omnigent End-to-End is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.6k tokens (SKILL.md is roughly 6.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Verify Omnigent End-to-End: CodexBar Live QA (steipete/CodexBar, 22k stars), Senpi Agent QA Harness (code-yeongyu/senpi, 472 stars), tmux Real User Testing (QwenLM/qwen-code, 28k stars) and Glance Test (DebugBase/glance, 156 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
omnigent-ai (a GitHub organization) maintains it in omnigent-ai/omnigent, which has 10,691 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 9, 2026.
Source: omnigent-ai/omnigent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.