Agent skill

Cross Eval

by alirezarezvani in alirezarezvani/claude-skills

/cs:cross-eval <memo — Multi-model consensus on a board memo or strategy brief.

MITAuto-check passedDevelopment

Install Cross Eval

skills CLI
$ npx skills add alirezarezvani/claude-skills --skill cross-eval -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install alirezarezvani/claude-skills cross-eval --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/alirezarezvani/claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/c-level-agents/skills/cross-eval .claude/skills/cross-eval && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
cross-eval
GitHub stars
28k
Token cost
~1.1k tokens
SKILL.md length
348 words
Files
1
Skills in repo
342
Repo updated
First seen
Licence
MIT

At a glance

/cs:cross-eval <memo — Multi-model consensus on a board memo or strategy brief.

  • Works in 3 steps: Claude (primary, always available) — the… → Codex / OpenAI (if OPENAI_API_KEY or… → Gemini (if GEMINI_API_KEY or gemini CLI…
  • A high-stakes memo needs an independent sanity check before the boardroom — e.g
  • SKILL.md covers When to Run, Models Used (graceful…, Workflow and Output Format, plus 4 more sections
  • Needs OPENAI_API_KEY and GEMINI_API_KEY

What it does

Cross Eval is an agent skill from alirezarezvani/claude-skills. /cs:cross-eval <memo — Multi-model consensus on a board memo or strategy brief. Claude + Codex + Gemini cross-review with graceful degradation. Use when a high-stakes memo needs an independent sanity check before the boardroom — e.g. a bet-the-company pivot or fundraise terms.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Development, covering Error handling. It works with OpenAI. The repository describes itself as: 380 Claude Code skills & agent skills & plugins (30+ Agents, 70+ custom commands, 380+ skills, customizable references, scripts)for Claude Code, Codex, Gemini CLI, Cursor, and 8… The licence is MIT.

When your agent uses it

  • A high-stakes memo needs an independent sanity check before the boardroom — e.g
  • Tasks that involve Error handling

Example prompts

  • “/cross-eval”

Requirements

  • A credential in OPENAI_API_KEY
  • A credential in GEMINI_API_KEY

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Claude (primary, always available) — the boardroom's native voice
  2. Codex / OpenAI (if OPENAI_API_KEY or codex CLI available)
  3. Gemini (if GEMINI_API_KEY or gemini CLI available)

What it can do on your machine

Read from SKILL.md and the folder at commit 19392f7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENAI_API_KEY
    • GEMINI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Cross Eval loads about 1.1k tokens when it runs. Until then it costs about 72 tokens; SKILL.md has 348 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~72
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from alirezarezvani/claude-skills at commit 19392f7, republished under its MIT licence (© alirezarezvani). 348 words, ~1,076 tokens.

Download SKILL.mdSave it as .claude/skills/cross-eval/SKILL.md (or your agent's skills folder).
name
cross-eval
description
/cs:cross-eval <memo> — Multi-model consensus on a board memo or strategy brief. Claude + Codex + Gemini cross-review with graceful degradation. Use when a high-stakes memo needs an independent sanity check before the boardroom — e.g. a bet-the-company pivot or fundraise terms.

/cs:cross-eval — Multi-Model Consensus

Command: /cs:cross-eval <memo-or-brief>

Runs the same memo through multiple model providers and reconciles divergences. Use for high-stakes, irreversible decisions where single-model bias is too costly: M&A, major fundraises, layoffs, strategic pivots, regulatory commitments.

Adapted from gstack's /codex cross-review pattern, generalized to business memos instead of code PRs.

When to Run

  • Before signing a term sheet
  • Before announcing a layoff
  • Before committing to a regulated market
  • Before any decision where reversing costs > 6 months of company time
  • When the boardroom vote was split or had a CRITICAL dissent

Models Used (graceful degradation)

The command tries to invoke each available model in order:

  1. Claude (primary, always available) — the boardroom's native voice
  2. Codex / OpenAI (if OPENAI_API_KEY or codex CLI available)
  3. Gemini (if GEMINI_API_KEY or gemini CLI available)

If only Claude is available, the command runs Claude-only with adversarial mode — same model, different prompt seeds — and clearly labels the output as single-model.

Workflow

  1. Read the memo / brief
  2. Probe environment for available model CLIs / API keys
  3. For each available model:
    • Send the memo with this prompt prefix:

      "You are an independent C-suite reviewer. The following is a board memo from another company's boardroom. Identify the top 3 concerns, the top 3 supports, and your vote (APPROVE / REJECT / DEFER). Do not deferentially agree — assume the memo's reasoning is flawed until proven otherwise."

  4. Collect three independent reviews
  5. Reconcile: where do they agree? Where do they diverge?
  6. Surface the divergences as questions for the founder

Output Format

Saved to ~/.claude/cross-eval/YYYY-MM-DD-<slug>.md:

markdown
# Cross-Eval: <memo title>
**Date:** YYYY-MM-DD
**Memo reviewed:** <link>
**Models invoked:** Claude / Codex / Gemini (or noted fallbacks)

## Vote Tally
| Model | Vote | Confidence |
|---|---|---|
| Claude | APPROVE | High |
| Codex | DEFER | Med |
| Gemini | APPROVE | Low |

## Consensus Concerns (≥2 models flagged)
1. <concern> — flagged by Claude + Codex
2. <concern> — flagged by all 3

## Divergent Concerns (1 model flagged)
- <Codex only:> <concern> — worth a second look
- <Gemini only:> <concern> — likely noise, but check

## Consensus Supports (≥2 models endorsed)
1. <support>
2. <support>

## Recommendation
- 🟢 GO if 2+ models APPROVE and no CRITICAL concerns from any model
- 🟡 PAUSE if any model is DEFER or any concern is CRITICAL
- 🔴 STOP if 2+ models REJECT

## Open Questions for Founder
1. <question raised by divergence>
2. <question raised by divergence>

Why This Matters

Single-model recommendations have systematic biases. Claude trends helpful and may under-weight risk. Codex (OpenAI) trends more cautious on emerging-market and regulatory topics. Gemini trends more cautious on technical scale claims. Disagreement is signal, not noise.

This is the safety net before irreversibility — not a replacement for outside counsel or a real board.

Graceful Degradation

If only Claude is available:

markdown
**Models available:** Claude only
**Mode:** ADVERSARIAL — running 3 independent Claude passes with different system prompts:
  1. Standard reviewer
  2. Devil's advocate (must find 3 critical concerns)
  3. Steelman (must find 3 strongest reasons to approve)

This is weaker than true multi-model. Treat the result as suggestive, not conclusive.

Routing

  • /cs:decide — if consensus is GO
  • /cs:freeze — if consensus is PAUSE
  • /cs:boardroom (re-run) — if consensus is STOP

Version: 1.0.0

© alirezarezvani, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in c-level-agents/skills/cross-eval of alirezarezvani/claude-skills.

Open the folder on GitHubat commit 19392f7

Compare with similar skills

Cross Eval next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Cross Eval compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Cross Eval this skillalirezarezvani/claude-skills28k—~1.1kAutomated safety check: PassMIT
Error Handlingmajiayu000/litellm-rs116—~2kAutomated safety check: PassMIT
LLM Providercaliber-ai-org/ai-setup1.3k—~2.7kAutomated safety check: PassMIT
Agent Tool Builderomer-metin/skills-for-antigravity162—~705Automated safety check: PassApache-2.0
Get API Docs with chubandrewyng/context-hub14k2 repos~775Automated safety check: PassMIT
DeepChat Provider IntegrationThinkInAIXYZ/deepchat6.4k1 repos~1.2kAutomated safety check: PassApache-2.0

Similar skills

  • Error Handling

    majiayu000/litellm-rs

    LiteLLM-RS Error Handling Architecture. An agent skill from majiayu000/litellm-rs.

    116 GitHub stars~2k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • LLM Provider

    caliber-ai-org/ai-setup

    Adds a new LLM provider implementing LLMProvider interface with call() and stream() methods.

    1.3k GitHub stars~2.7k tokensUpdated 13 days ago
    Testing & QAAuto-check passed
  • Agent Tool Builder

    omer-metin/skills-for-antigravity

    Tools are how AI agents interact with the world. An agent skill from omer-metin/skills-for-antigravity.

    162 GitHub stars~705 tokensUpdated 8 mo ago
    AI & LLM EngineeringAuto-check passed
  • Get API Docs with chub

    andrewyng/context-hub

    Fetches current documentation for third-party APIs and SDKs with the chub CLI before the agent writes code against them, instead of relying on remembered API shapes.

    14k GitHub starsUsed in 2 repos~775 tokens
    DevelopmentAuto-check passed
  • DeepChat Provider Integration

    ThinkInAIXYZ/deepchat

    Guides adding an LLM provider to DeepChat through explicit source changes: collect the provider details, pick a transport path and add registry entries and tests.

    6.4k GitHub starsUsed in 1 repo~1.2k tokens
    DevelopmentAuto-check passed
  • Create Skill

    Hyk260/PureChat

    Create a new skill in the current repository. An agent skill from Hyk260/PureChat.

    546 GitHub starsUsed in 1 repo~823 tokens
    DevelopmentAuto-check passed

More from alirezarezvani/claude-skills

All 342 skills in this repo
  • Agile Product Owner

    alirezarezvani/claude-skills

    Writes INVEST-checked user stories with acceptance criteria, splits epics, plans sprints from velocity and ranks the backlog with a weighted score.

    28k GitHub starsUsed in 3 repos~3.2k tokens
    Auto-check passed
  • Product Strategist

    alirezarezvani/claude-skills

    OKR cascade toolkit for product leaders: generates aligned company-to-team OKRs from five strategy types and scores how well they line up.

    28k GitHub starsUsed in 2 repos~1.8k tokens
    Auto-check passed
  • App Store Optimization

    alirezarezvani/claude-skills

    App Store Optimization (ASO) toolkit for researching keywords, analyzing competitor rankings, generating metadata suggestions, and improving app visibility on Apple App Store and Google Play Store.

    28k GitHub starsUsed in 1 repo~4.2k tokens
    Auto-check passed
  • AWS Solution Architect

    alirezarezvani/claude-skills

    Design AWS architectures for startups using serverless patterns and IaC templates.

    28k GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • Campaign Analytics

    alirezarezvani/claude-skills

    Calculates attribution, funnel and ROI figures for marketing campaigns with three Python scripts that need only the standard library.

    28k GitHub starsUsed in 1 repo~2.1k tokens
    Auto-check passed
  • Code to PRD

    alirezarezvani/claude-skills

    Reverse-engineers a frontend, backend or fullstack codebase into a product requirements document with per-page docs, an enum dictionary and an API inventory.

    28k GitHub starsUsed in 1 repo~4.9k tokens
    Auto-check passed

Works with

Questions about Cross Eval

What does Cross Eval do?

/cs:cross-eval <memo — Multi-model consensus on a board memo or strategy brief. Cross Eval is an agent skill from alirezarezvani/claude-skills. /cs:cross-eval <memo — Multi-model consensus on a board memo or strategy brief.

When should I use Cross Eval?

Cross Eval fits situations like: A high-stakes memo needs an independent sanity check before the boardroom — e.g; tasks that involve Error handling.

How do I install Cross Eval in Claude Code?

Run `npx skills add alirezarezvani/claude-skills --skill cross-eval -a claude-code`. Or copy the skill folder (c-level-agents/skills/cross-eval in alirezarezvani/claude-skills) into .claude/skills/cross-eval in your project. Claude Code loads it when a task matches its description.

How do I install Cross Eval in Codex?

Run `npx skills add alirezarezvani/claude-skills --skill cross-eval -a codex`. Or copy the skill folder (c-level-agents/skills/cross-eval in alirezarezvani/claude-skills) into .agents/skills/cross-eval in your project. Codex loads it when a task matches its description.

Can I use Cross Eval in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add alirezarezvani/claude-skills --skill cross-eval -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/cross-eval, .gemini/skills/cross-eval, .github/skills/cross-eval and .opencode/skills/cross-eval in your project.

What does Cross Eval need to run?

Going by SKILL.md and its folder, Cross Eval needs credentials named OPENAI_API_KEY and GEMINI_API_KEY. Our summary lists: A credential in OPENAI_API_KEY; A credential in GEMINI_API_KEY.

Does Cross Eval access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Cross Eval safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Cross Eval use?

Cross Eval is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Cross Eval use?

About 1.1k tokens (SKILL.md is roughly 4.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Cross Eval?

Skills that share tags, products or a category with Cross Eval: Error Handling (majiayu000/litellm-rs, 116 stars), LLM Provider (caliber-ai-org/ai-setup, 1.3k stars), Agent Tool Builder (omer-metin/skills-for-antigravity, 162 stars) and Get API Docs with chub (andrewyng/context-hub, 14k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Cross Eval?

alirezarezvani (a GitHub user) maintains it in alirezarezvani/claude-skills, which has 27,788 GitHub stars. The repository holds 342 skills in this directory. The repository was last updated on August 30, 2026.

Source: alirezarezvani/claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.