OmniRoute Chat CLI
diegosouzapw/OmniRoute
Sends chat completions, streams responses, and opens an interactive REPL against any OmniRoute-routed model provider.
Provides context about the CoStrict evals system structure in this monorepo.
$ npx skills add zgsm-ai/costrict --skill evals-context -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install zgsm-ai/costrict evals-context --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.roo/skills/evals-context .claude/skills/evals-context && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .claude/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-contextType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add zgsm-ai/costrict --skill evals-context -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install zgsm-ai/costrict evals-context --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.roo/skills/evals-context .agents/skills/evals-context && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .agents/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add zgsm-ai/costrict --skill evals-context -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install zgsm-ai/costrict evals-context --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.roo/skills/evals-context .cursor/skills/evals-context && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .cursor/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/zgsm-ai/costrict.git --path .roo/skills/evals-context--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add zgsm-ai/costrict --skill evals-context -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install zgsm-ai/costrict evals-context --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.roo/skills/evals-context .gemini/skills/evals-context && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .gemini/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install zgsm-ai/costrict evals-contextInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add zgsm-ai/costrict --skill evals-context -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .github/skills && cp -r skills-src/.roo/skills/evals-context .github/skills/evals-context && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .github/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add zgsm-ai/costrict --skill evals-context -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install zgsm-ai/costrict evals-context --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zgsm-ai/costrict.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.roo/skills/evals-context .opencode/skills/evals-context && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "evals-context" agent skill from https://github.com/zgsm-ai/costrict/tree/main/.roo/skills/evals-context into .opencode/skills/evals-context/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "evals-context", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
evals-contextProvides context about the CoStrict evals system structure in this monorepo.
Evals Context is an agent skill from zgsm-ai/costrict. Provides context about the CoStrict evals system structure in this monorepo. Use when tasks mention "evals", "evaluation", "eval runs", "eval exercises", or working with the evals infrastructure. Helps distinguish between the evals execution system (packages/evals, apps/web-evals) and the public website evals display page (apps/web-roo-code/src/app/evals).
Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in AI & LLM Engineering, covering LLM evaluation and Monorepo tooling. It works with Visual Studio Code, DeepSeek, Google Gemini and MiniMax. The repository describes itself as: Costrict - strict AI coder for enterprises, quality first, including AI Agent, AI CodeReview, AI Completion. The licence is Apache-2.0.
2 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit dd38f54. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pnpmnpxFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
github.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Evals Context loads about 1.9k tokens when it runs. Until then it costs about 93 tokens; SKILL.md has 380 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from zgsm-ai/costrict at commit dd38f54, republished under its Apache-2.0 licence (© zgsm-ai). 380 words, ~1,881 tokens.
.claude/skills/evals-context/SKILL.md (or your agent's skills folder).Use this skill when the task involves:
Do NOT use this skill when:
This monorepo has two distinct evals-related locations that can cause confusion:
| Component | Path | Purpose |
|---|---|---|
| Evals Execution System | packages/evals/ | Core eval infrastructure: CLI, DB schema, Docker configs |
| Evals Management UI | apps/web-evals/ | Next.js app for creating/monitoring eval runs (localhost:3446) |
| Website Evals Page | apps/web-roo-code/src/app/evals/ | Public roocode.com page displaying eval results |
| External Exercises Repo | Roo-Code-Evals | Actual coding exercises (NOT in this monorepo) |
packages/evals/ - Core Evals Packagepackages/evals/
├── ARCHITECTURE.md # Detailed architecture documentation
├── ADDING-EVALS.md # Guide for adding new exercises/languages
├── README.md # Setup and running instructions
├── docker-compose.yml # Container orchestration
├── Dockerfile.runner # Runner container definition
├── Dockerfile.web # Web app container
├── drizzle.config.ts # Database ORM config
├── src/
│ ├── index.ts # Package exports
│ ├── cli/ # CLI commands for running evals
│ │ ├── runEvals.ts # Orchestrates complete eval runs
│ │ ├── runTask.ts # Executes individual tasks in containers
│ │ ├── runUnitTest.ts # Validates task completion via tests
│ │ └── redis.ts # Redis pub/sub integration
│ ├── db/
│ │ ├── schema.ts # Database schema (runs, tasks)
│ │ ├── queries/ # Database query functions
│ │ └── migrations/ # SQL migrations
│ └── exercises/
│ └── index.ts # Exercise loading utilities
└── scripts/
└── setup.sh # Local macOS setup scriptapps/web-evals/ - Evals Management Web Appapps/web-evals/
├── src/
│ ├── app/
│ │ ├── page.tsx # Home page (runs list)
│ │ ├── runs/
│ │ │ ├── new/ # Create new eval run
│ │ │ └── [id]/ # View specific run status
│ │ └── api/runs/ # SSE streaming endpoint
│ ├── actions/ # Server actions
│ │ ├── runs.ts # Run CRUD operations
│ │ ├── tasks.ts # Task queries
│ │ ├── exercises.ts # Exercise listing
│ │ └── heartbeat.ts # Controller health checks
│ ├── hooks/ # React hooks (SSE, models, etc.)
│ └── lib/ # Utilities and schemasapps/web-roo-code/src/app/evals/ - Public Website Evals Pageapps/web-roo-code/src/app/evals/
├── page.tsx # Fetches and displays public eval results
├── evals.tsx # Main evals display component
├── plot.tsx # Visualization component
└── types.ts # EvalRun type (extends packages/evals types)This page displays eval results on the public roocode.com website. It imports types from @roo-code/evals but does NOT run evals.
The evals system is a distributed evaluation platform that runs AI coding tasks in isolated VS Code environments:
┌─────────────────────────────────────────────────────────────┐
│ Web App (apps/web-evals) ──────────────────────────────── │
│ │ │
│ ▼ │
│ PostgreSQL ◄────► Controller Container │
│ │ │ │
│ ▼ ▼ │
│ Redis ◄───► Runner Containers (1-25 parallel) │
└─────────────────────────────────────────────────────────────┘Key components:
packages/evals/ADDING-EVALS.md for structureEdit files in packages/evals/src/cli/:
runEvals.ts - Run orchestrationrunTask.ts - Task executionrunUnitTest.ts - Test validationEdit files in apps/web-evals/src/:
app/runs/new/new-run.tsx - New run formactions/runs.ts - Run server actionsEdit files in apps/web-roo-code/src/app/evals/:
packages/evals/src/db/schema.tscd packages/evals && pnpm drizzle-kit generatepnpm drizzle-kit migrate# From repo root
pnpm evals
# Opens web UI at http://localhost:3446Ports (defaults):
# packages/evals tests
cd packages/evals && npx vitest run
# apps/web-evals tests
cd apps/web-evals && npx vitest run@roo-code/evalsThe package exports are defined in packages/evals/src/index.ts:
getRuns, getTasks, getTaskMetrics, etc.Run, Task, TaskMetricsapps/web-evals and apps/web-roo-code© zgsm-ai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .roo/skills/evals-context of zgsm-ai/costrict.
Open the folder on GitHubat commit dd38f54
We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in zgsm-ai/costrict, which our catalogue first saw on October 7, 2026.
Evals Context next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Evals Context this skillzgsm-ai/costrict | 4.4k | 1 repos | ~1.9k | Automated safety check: Pass | Apache-2.0 | |
| OmniRoute Chat CLIdiegosouzapw/OmniRoute | 74k | — | ~345 | Automated safety check: Pass | MIT | |
| ModLens Image Vision Bridgeliustack/modlens | 4.2k | — | ~1.3k | Automated safety check: Notes | MIT | |
| Quality FlywheelGoogleCloudPlatform/vertex-ai-samples | 792 | — | ~2k | Automated safety check: Pass | Apache-2.0 | |
| Cross-Model Benchmarkgarrytan/gstack | 136k | — | ~4k | Automated safety check: Notes | MIT | |
| Frontierharness Evalfrontier-harness-eval/eval | 301 | — | ~8k | Automated safety check: Pass | None |
diegosouzapw/OmniRoute
Sends chat completions, streams responses, and opens an interactive REPL against any OmniRoute-routed model provider.
liustack/modlens
Gives text-only models sight by running the modlens CLI on an image path or URL and returning structured JSON evidence with transcribed text, layout and semantics.
GoogleCloudPlatform/vertex-ai-samples
Evaluate and improve GenAI models and agents using the Google GenAI Evaluation SDK.
garrytan/gstack
Sends one prompt to Claude, GPT through the Codex CLI and Gemini, then tabulates response time, token use and cost, with an optional judged quality score.
frontier-harness-eval/eval
Benchmark a third-party coding-agent harness against FrontierHarness Eval using Runta runtimes.
bitsky-tech/bridgic
LLM provider initialization for bridgic projects. An agent skill from bitsky-tech/bridgic.
zgsm-ai/costrict
Provides comprehensive guidelines for resolving merge conflicts intelligently using git history and commit context.
zgsm-ai/costrict
Provides comprehensive guidelines for translating and localizing CoStrict extension strings.
Categories
Provides context about the CoStrict evals system structure in this monorepo. Evals Context is an agent skill from zgsm-ai/costrict. Provides context about the CoStrict evals system structure in this monorepo.
Evals Context fits situations like: tasks mention evals; working with the evals infrastructure.
Run `npx skills add zgsm-ai/costrict --skill evals-context -a claude-code`. Or copy the skill folder (.roo/skills/evals-context in zgsm-ai/costrict) into .claude/skills/evals-context in your project. Claude Code loads it when a task matches its description.
Run `npx skills add zgsm-ai/costrict --skill evals-context -a codex`. Or copy the skill folder (.roo/skills/evals-context in zgsm-ai/costrict) into .agents/skills/evals-context in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add zgsm-ai/costrict --skill evals-context -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/evals-context, .gemini/skills/evals-context, .github/skills/evals-context and .opencode/skills/evals-context in your project.
Going by SKILL.md and its folder, Evals Context needs the command-line tools its instructions call (pnpm and npx). Our summary lists: Node.js; Docker.
SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Evals Context is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.9k tokens (SKILL.md is roughly 7.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Evals Context: OmniRoute Chat CLI (diegosouzapw/OmniRoute, 74k stars), ModLens Image Vision Bridge (liustack/modlens, 4.2k stars), Quality Flywheel (GoogleCloudPlatform/vertex-ai-samples, 792 stars) and Cross-Model Benchmark (garrytan/gstack, 136k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
zgsm-ai (a GitHub organization) maintains it in zgsm-ai/costrict, which has 4,448 GitHub stars. The repository holds 3 skills in this directory. The repository was last updated on September 30, 2026.
Source: zgsm-ai/costrict on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.