Arize Evaluator
github/awesome-copilot
Handles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, running evaluations on spans or experiments, managing tasks, trigger-run operations, column mapping, and…
A skill your agent uses when evaluating team members — Core Values alignment and GWC (right people, right seats)
$ npx skills add bradfeld/ceos --skill ceos-people -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install bradfeld/ceos ceos-people --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/ceos-people .claude/skills/ceos-people && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .claude/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/bradfeld/ceos/tree/main/skills/ceos-peopleType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add bradfeld/ceos --skill ceos-people -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install bradfeld/ceos ceos-people --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/ceos-people .agents/skills/ceos-people && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .agents/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add bradfeld/ceos --skill ceos-people -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install bradfeld/ceos ceos-people --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/ceos-people .cursor/skills/ceos-people && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .cursor/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/bradfeld/ceos.git --path skills/ceos-people--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add bradfeld/ceos --skill ceos-people -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install bradfeld/ceos ceos-people --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/ceos-people .gemini/skills/ceos-people && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .gemini/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install bradfeld/ceos ceos-peopleInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add bradfeld/ceos --skill ceos-people -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/ceos-people .github/skills/ceos-people && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .github/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add bradfeld/ceos --skill ceos-people -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install bradfeld/ceos ceos-people --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/bradfeld/ceos.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/ceos-people .opencode/skills/ceos-people && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "ceos-people" agent skill from https://github.com/bradfeld/ceos/tree/main/skills/ceos-people into .opencode/skills/ceos-people/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "ceos-people", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
ceos-peopleA skill your agent uses when evaluating team members — Core Values alignment and GWC (right people, right seats)
Ceos People is an agent skill from bradfeld/ceos. Use when evaluating team members — Core Values alignment and GWC (right people, right seats)
Its SKILL.md is about 3.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
The repository describes itself as: EOS (Entrepreneurial Operating System) skills for Claude Code. Run your business on EOS with AI. The licence is MIT.
12 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit bdd5f01. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
gitFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Ceos People loads about 3.6k tokens when it runs. Until then it costs about 26 tokens; SKILL.md has 1,566 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from bradfeld/ceos at commit bdd5f01, republished under its MIT licence (© bradfeld). 1,566 words, ~3,644 tokens.
.claude/skills/ceos-people/SKILL.md (or your agent's skills folder).Evaluate whether team members are the right people (Core Values alignment) in the right seats (GWC: Get it, Want it, Capacity to do it). Manage people evaluations, run quarterly reviews, and flag below-the-bar situations for action.
Search upward from the current directory for the .ceos marker file. This file marks the root of the CEOS repository.
If .ceos is not found, stop and tell the user: "Not in a CEOS repository. Clone your CEOS repo and run setup.sh first."
Sync before use: Once you find the CEOS root, run git -C <ceos_root> pull --ff-only --quiet 2>/dev/null to get the latest data from teammates. If it fails (conflict or offline), continue silently with local data.
| File | Purpose |
|---|---|
data/people/ | Person evaluation files (one per person) |
data/people/alumni/ | Departed team members (historical reference) |
data/vision.md | Source of Core Values (read-only — use ceos-vto to modify) |
data/accountability.md | Source of seats and owners (reference for GWC) |
templates/people-analyzer.md | Template for new person evaluations |
Each person is a markdown file at data/people/firstname-lastname.md with YAML frontmatter:
name: "Brad Feld"
seat: "Visionary"
core_values:
# Each Core Value from vision.md, rated +, +/-, or -
status: right_person_right_seat # right_person_right_seat | below_bar | wrong_seat | evaluating
gwc:
get: true # true | false | null
want: true
capacity: true
last_evaluated: "2026-01-15"
created: "2026-01-02"
departed: falseFile naming: firstname-lastname.md — lowercase, hyphenated. Person-centric (survives role changes).
| Status | Meaning | When |
|---|---|---|
right_person_right_seat | Passes both Core Values and GWC | All Core Values are + or mostly +, all GWC = true |
below_bar | Fails Core Values OR GWC | Three strikes on values, or any GWC = false |
wrong_seat | Right person, wrong seat | Core Values pass but GWC fails for current seat |
evaluating | Not yet assessed | New hire (< 90 days) or incomplete evaluation |
| Rating | Meaning |
|---|---|
+ | Lives this value most of the time |
+/- | Sometimes demonstrates, sometimes doesn't |
- | Rarely or never demonstrates this value |
Three strikes rule: Three or more +/- or - ratings = "wrong person" (Core Values misalignment). This is a critical flag that requires action.
| Dimension | Question | Notes |
|---|---|---|
| Get it | Do they truly understand the role? | Intuitive grasp of the job, culture, systems |
| Want it | Do they genuinely want the work? | Not just title/pay — the actual daily work |
| Capacity | Can they do it? | Time, skill, knowledge, emotional capacity |
All three must be true for "right seat." Any single false = wrong seat.
Use when evaluating a specific person against Core Values and GWC.
Ask for the person's name. Check if data/people/firstname-lastname.md already exists.
templates/people-analyzer.mdRead Core Values from data/vision.md. For each Core Value, ask the user to rate: +, +/-, or -.
Display a rating table as you go:
Core Values Evaluation — [Person Name]
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
| Core Value | Rating | Notes |
|---------------|--------|-----------------|
| Integrity | + | Consistently |
| Innovation | +/- | Room to grow |
| Transparency | + | |
| Grit | - | Avoids hard work|After all Core Values are rated, count +/- and - ratings:
+/- or -): Flag immediately:⚠️ THREE STRIKES — Core Values misalignment detected.
[Person] has 3+ values rated +/- or -. This signals "wrong person."
Action required: coaching conversation, role change, or exit plan.+: "Strong Core Values alignment. Right person."Determine right_person: All or mostly + ratings = yes. Three strikes = no.
Read the person's current seat from data/accountability.md. If the person owns multiple seats, evaluate GWC for each seat separately.
For each seat, ask three binary questions:
Display the result:
GWC — [Person Name] as [Seat Name]
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Get it: ✓ Yes
Want it: ✓ Yes
Capacity: ✗ No — struggling with volume
Right seat? No (Capacity gap)Determine right_seat: All three must be true. Any false = wrong seat.
Based on the evaluation:
| Core Values | GWC | Suggested Status |
|---|---|---|
| Right person | Right seat | right_person_right_seat |
| Right person | Wrong seat | wrong_seat |
| Wrong person | Any | below_bar |
| Incomplete | Any | evaluating |
Present the suggestion: "Based on this evaluation, I'd suggest [status]. Do you agree, or would you set it differently?"
Always let the user confirm or override. Status is a leadership judgment call, not a formula.
Show the complete evaluation file before writing. Ask: "Save this evaluation?"
Update last_evaluated to today's date. Add a dated entry to the Evaluation History section.
If status is below_bar or wrong_seat, offer:
"[Person] is below the bar. Would you like to create an issue for a 30-day action plan? This will create a file in data/issues/open/ for IDS discussion."
If yes, use the issue template pattern from data/issues/open/ to create an issue with:
Use when reviewing the current state of all people evaluations.
Read all files from data/people/ (exclude alumni/ subdirectory). Parse the YAML frontmatter for each person.
If no files exist: "No people evaluations found. Run an Evaluate for your first team member."
People Analyzer — Team Overview
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
| Name | Seat | Status | Last Evaluated | Flag |
|---------------|-------------|-------------------------|----------------|---------|
| Brad Feld | Visionary | right_person_right_seat | 2026-01-15 | |
| Sarah Chen | Integrator | right_person_right_seat | 2026-01-15 | |
| Mike Torres | VP Sales | wrong_seat | 2025-12-01 | ⚠️ Seat |
| Alex Kim | VP Eng | below_bar | 2025-11-15 | 🔴 Bar |
| Jamie Lee | Marketing | evaluating | 2026-01-28 | 🆕 New |
The Bar: 3/5 (60%) at or above — Target: 80%+Flag the following:
🔴 — requires action plan⚠️ — person is right, seat is wrong (find a better fit)last_evaluated is > 120 days ago, flag: 📅 Overdue🆕 — new hire or incomplete evaluationCalculate: (right_person_right_seat count) / (total evaluated, excluding "evaluating").
If below 80%: "Below the 80% target. Consider bringing people discussions to the next L10."
Ask: "Want to drill into any person, or run a new Evaluate?"
Use for the formal quarterly review of all seats against the Accountability Chart.
Read data/accountability.md to get all seats and their current owners.
If the file doesn't exist or is empty: "No accountability chart found. Create one first with ceos-accountability or manually at data/accountability.md."
For each seat in the accountability chart:
Display the seat map:
Quarterly People Review
━━━━━━━━━━━━━━━━━━━━━━━
| Seat | Owner | Status | Action Needed? |
|-------------|---------------|-------------------------|-----------------|
| Visionary | Brad Feld | right_person_right_seat | No |
| Integrator | Sarah Chen | right_person_right_seat | No |
| VP Sales | Mike Torres | wrong_seat | Re-evaluate GWC |
| VP Eng | Alex Kim | below_bar | Action plan due |
| Marketing | (empty) | — | Hire needed |For each filled seat, ask: "Re-evaluate, update notes, or skip?"
Track progress:
Progress: 3/5 seats reviewed [████████░░░░] 60%For each empty seat, offer:
"The [Seat Name] seat is empty. Would you like to create an issue for hiring? This will go to data/issues/open/ for IDS discussion."
After reviewing all seats, display:
Quarterly People Review — Complete
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Seats filled: 4/5 (80%)
Right People, Right Seats: 2/4 (50%)
Below the bar: 1
Wrong seat: 1
Empty seats: 1
Evaluating: 0
Action items created: 2
Next quarterly review: [quarter-end date]If any seats are below bar or empty: "These should be discussed at the next L10. Bring them to the Issues List."
Evaluate: Show the complete evaluation file before writing. Display Core Values table and GWC results inline.
Review: Summary table with status flags, bar percentage. Offer drill-down.
Quarterly: Seat-by-seat walkthrough with progress tracker. End with quarterly summary.
+/- or -, always flag it prominently. Do not minimize.departed: true and move to data/people/alumni/.data/vision.md — never ask the user to list them (they're already defined in the V/TO).data/accountability.md rather than asking the user to recall it.ceos-vto, ceos-quarterly, and ceos-ids when relevant, but let the user decide when to switch workflows.ceos-people reads Core Values from data/vision.md for the Core Values evaluation. It does not write to the V/TO file.ceos-vto, existing people evaluations may need refreshing.ceos-people reads data/accountability.md for the person's seat(s) during GWC evaluation. It does not write to the accountability file.ceos-quarterly references People Analyzer evaluations from data/people/ during quarterly conversations. Core Values and GWC ratings serve as reference points, not re-evaluations.ceos-people.below_bar or wrong_seat, ceos-people offers to create an issue in data/issues/open/.ceos-ids for formal issue tracking of people-related action plans.ceos-annual references People Analyzer evaluations during the Organizational Checkup section (Section 4) of the annual planning session.Only ceos-people writes to data/people/. Other skills read person evaluations for reference. The quarterly conversation skill references evaluations but directs updates back to ceos-people.
© bradfeld, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/ceos-people of bradfeld/ceos.
Open the folder on GitHubat commit bdd5f01
Ceos People next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Ceos People this skillbradfeld/ceos | 156 | — | ~3.6k | Automated safety check: Pass | MIT | |
| Arize Evaluatorgithub/awesome-copilot | 40k | 1 repos | ~8.1k | Automated safety check: Notes | MIT | |
| LLM Evaluationdavila7/claude-code-templates | 33k | 12 repos | ~3.5k | Automated safety check: Pass | MIT | |
| Agent Evaluationsickn33/agentic-awesome-skills | 47k | 1 repos | ~2k | Automated safety check: Pass | MIT | |
| EvaluatorsArize-ai/phoenix | 12k | — | ~1.7k | Automated safety check: Pass | Custom licence | |
| Agent Evaluation Reportingsickn33/agentic-awesome-skills | 47k | 1 repos | ~2.1k | Automated safety check: Pass | MIT |
github/awesome-copilot
Handles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, running evaluations on spans or experiments, managing tasks, trigger-run operations, column mapping, and…
davila7/claude-code-templates
Master comprehensive evaluation strategies for LLM applications, from automated metrics to human evaluation and A/B testing.
sickn33/agentic-awesome-skills
Evaluate agent behavior with versioned cases and explicit verifiers.
Arize-ai/phoenix
Author or refine a Phoenix evaluator — code or LLM-as-a-judge — that scores a run's output.
sickn33/agentic-awesome-skills
A skill your agent uses when summarizing agent evaluations where autonomous, assisted, failed, timed-out, or invalid outcomes must remain distinct and comparable.
PostHog/posthog
Author continuously-running online evaluations in PostHog AI observability, grounded in real failure modes you've identified.
bradfeld/ceos
A skill your agent uses when viewing, updating, or auditing the Accountability Chart — seats, owners, and roles
bradfeld/ceos
A skill your agent uses when managing daily operational delegation through The Stack and structured leader-assistant standups
bradfeld/ceos
A skill your agent uses when managing the rolling market calendar - conferences, launches, fundraising milestones, partner events, and team constraints
bradfeld/ceos
A skill your agent uses when assessing and optimizing the 8 financial levers that drive cash flow and profitability
bradfeld/ceos
A skill your agent uses when taking a Clarity Break — stepping back from day-to-day work for strategic thinking time
bradfeld/ceos
A skill your agent uses when you want a quick snapshot of overall business health across all EOS components
A skill your agent uses when evaluating team members — Core Values alignment and GWC (right people, right seats). Ceos People is an agent skill from bradfeld/ceos.
Ceos People fits situations like: evaluating team members — Core Values alignment and GWC (right people.
Run `npx skills add bradfeld/ceos --skill ceos-people -a claude-code`. Or copy the skill folder (skills/ceos-people in bradfeld/ceos) into .claude/skills/ceos-people in your project. Claude Code loads it when a task matches its description.
Run `npx skills add bradfeld/ceos --skill ceos-people -a codex`. Or copy the skill folder (skills/ceos-people in bradfeld/ceos) into .agents/skills/ceos-people in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add bradfeld/ceos --skill ceos-people -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/ceos-people, .gemini/skills/ceos-people, .github/skills/ceos-people and .opencode/skills/ceos-people in your project.
Going by SKILL.md and its folder, Ceos People needs the command-line tools its instructions call (git).
SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Ceos People is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.6k tokens (SKILL.md is roughly 15k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Ceos People: Arize Evaluator (github/awesome-copilot, 40k stars), LLM Evaluation (davila7/claude-code-templates, 33k stars), Agent Evaluation (sickn33/agentic-awesome-skills, 47k stars) and Evaluators (Arize-ai/phoenix, 12k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
bradfeld (a GitHub user) maintains it in bradfeld/ceos, which has 156 GitHub stars. The repository holds 23 skills in this directory. The repository was last updated on September 13, 2026.
Source: bradfeld/ceos on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.