Acceptance Evidence for Deliveries
lobehub/lobehub
Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI.
End-to-end recipe for adding a new task under examples/ — the three pieces that have to line up (task.yaml, seed/, and grader/), what to put in each, the TaskGrader API surface, the coral validate →…
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Human-Agent-Society/CORAL coral-new-task --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/coral-new-task .claude/skills/coral-new-task && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .claude/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-taskType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Human-Agent-Society/CORAL coral-new-task --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.claude/skills/coral-new-task .agents/skills/coral-new-task && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .agents/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Human-Agent-Society/CORAL coral-new-task --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.claude/skills/coral-new-task .cursor/skills/coral-new-task && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .cursor/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Human-Agent-Society/CORAL.git --path .claude/skills/coral-new-task--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Human-Agent-Society/CORAL coral-new-task --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.claude/skills/coral-new-task .gemini/skills/coral-new-task && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .gemini/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Human-Agent-Society/CORAL coral-new-taskInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .github/skills && cp -r skills-src/.claude/skills/coral-new-task .github/skills/coral-new-task && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .github/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Human-Agent-Society/CORAL coral-new-task --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Human-Agent-Society/CORAL.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.claude/skills/coral-new-task .opencode/skills/coral-new-task && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "coral-new-task" agent skill from https://github.com/Human-Agent-Society/CORAL/tree/main/.claude/skills/coral-new-task into .opencode/skills/coral-new-task/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "coral-new-task", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
coral-new-taskEnd-to-end recipe for adding a new task under examples/ — the three pieces that have to line up (task.yaml, seed/, and grader/), what to put in each, the TaskGrader API surface, the coral validate →…
Coral New Task is an agent skill from Human-Agent-Society/CORAL. End-to-end recipe for adding a new task under examples/ — the three pieces that have to line up (task.yaml, seed/, and grader/), what to put in each, the TaskGrader API surface, the coral validate → smoke-test loop, and the common mistakes (repopath pointing at the wrong dir, score direction backwards, hidden answer keys leaking into seed/, grader writing to codebasepath which the daemon force-removes, private-vs-public confusion, missing run() signature). Use whenever the user wants to add a new CORAL task or…
Its SKILL.md is about 3.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Testing & QA, covering QA and bug reports. The repository describes itself as: Open-source autoresearch powered by autonomous coding agents. Run Claude Code, OpenCode, and Codex with grading, shared knowledge, and multi-agent evolution. Accepted at COLM 2026. The licence is Apache-2.0.
5 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 0123dfb. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
uvFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Coral New Task loads about 3.4k tokens when it runs. Until then it costs about 146 tokens; SKILL.md has 1,093 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Human-Agent-Society/CORAL at commit 0123dfb, republished under its Apache-2.0 licence (© Human-Agent-Society). 1,093 words, ~3,361 tokens.
.claude/skills/coral-new-task/SKILL.md (or your agent's skills folder).A CORAL task is three things that must line up:
examples/<task>/
├── task.yaml # config: name, description, grader entrypoint, agent count
├── seed/ # starter code agents see when they begin (the repo_path)
│ └── solution.py
└── grader/ # standalone Python package
├── pyproject.toml
└── src/<task>_grader/
├── __init__.py
└── grader.py # class Grader(TaskGrader): ...The packaged form is the only supported form. The package gives the grader its own venv and ships everything the eval needs — grader code, helper modules, and hidden data (see "Hidden data" below).
Look at these before writing anything new — copy the closest one and edit:
| Reference | When to copy it |
|---|---|
| examples/erdos/ | Minimal packaged grader, single grader file, numpy-only deps |
| examples/dna_design/ | Packaged grader with bundled data files (importlib.resources) and [ml] optional-deps for heavy libs |
| examples/swebench-verified/ | Tiered eval (different instance counts per tier), private answer keys, harbor integration |
| examples/circle_packing/ | Smallest packaged task end-to-end — single solution file, single grader file |
| examples/mnist/ | Packaged grader with a hidden answer key (note: secret data belongs under grader.private in a taskdata/ sibling of grader/, never inside the grader package) |
Whatever lives in seed/ is what the agent sees on first checkout — it's the working directory the grader will later score. The contract between seed/ and the grader is the program file: a Python file with a function the grader imports and calls.
The convention across examples is:
solution.py (or initial_program.py) defining a top-level run() function.program_file: "solution.py" via grader.args.run()'s signature is whatever the grader expects — usually () -> result or (input_path) -> result.Put a real, runnable baseline here. Agents should be able to coral eval immediately and get a non-zero score, so they have a starting point to improve. A no-op skeleton that crashes is not a good baseline.
If the task needs data files at runtime (training data, fixtures), put them under seed/data/ and reference them by relative path from solution.py. The grader will see them at <codebase_path>/data/....
grader/
├── pyproject.toml
└── src/<task>_grader/
├── __init__.py
└── grader.pypyproject.toml is a thin Hatchling package. Crib from examples/erdos/grader/pyproject.toml:
[project]
name = "<task>-grader"
version = "0.1.0"
description = "CORAL grader for the <task> task."
requires-python = ">=3.11"
dependencies = ["coral", "numpy"] # Whatever the grader actually imports.
[build-system]
requires = ["hatchling"]
build-backend = "hatchling.build"
[tool.hatch.build.targets.wheel]
packages = ["src/<task>_grader"]Subclass TaskGrader and implement evaluate():
# grader/src/<task>_grader/grader.py
from coral.grader import TaskGrader
from coral.types import ScoreBundle
class Grader(TaskGrader):
def evaluate(self) -> float | ScoreBundle:
program_file = self.args.get("program_file", "solution.py")
# self.codebase_path — the agent's commit checked out detached
# self.private_dir — .coral/private/ (your hidden answer keys live here)
# self.args — dict from task.yaml grader.args
# self.timeout — grader.timeout in seconds (or None)
# self.eval_logs_dir — write subprocess logs / artifacts the agent should see post-grade
try:
result = run_program_and_score(...)
except TimeoutError:
return self.fail(f"Evaluation timed out after {self.timeout}s")
except Exception as e:
return self.fail(f"Evaluation failed: {e}")
return self.score(result, explanation=f"score={result:.4f}")What you have available on self:
| Attribute / method | Use it for |
|---|---|
self.codebase_path | Path to the commit being graded (detached worktree). Read-only — anything written here is discarded after the eval. |
self.private_dir | .coral/private/. Your answer keys, hidden test data, anything from grader.private lives here. |
self.args | dict from task.yaml::grader.args. Use self.args.get("program_file", "solution.py") etc. |
self.timeout | Eval timeout in seconds (or None if grader.timeout: 0). |
self.eval_logs_dir | Per-attempt directory for logs/artifacts that should outlive the grader. Symlinked into each agent worktree as <shared_dir>/eval_logs/<hash>/. |
self.score(value, explanation=...) | Build a single-task ScoreBundle from a numeric score. |
self.fail(reason) | Return a fail ScoreBundle with reason as feedback. |
self.get_python_command() | List for the python binary inside the codebase's env (uses uv run if a pyproject.toml is present). Always use this instead of sys.executable so task-specific deps are visible. |
self.run_program(filename, *args) | Convenience: runs <codebase_path>/<filename> as a subprocess via get_python_command(). |
If the grader needs reference files (model weights, ground-truth answers, scoring fixtures), ship them inside the package and load via importlib.resources:
import importlib.resources
scorer_dir = str(importlib.resources.files("<task>_grader.scorers"))examples/dna_design/grader/src/dna_design_grader/grader.py is the canonical pattern — note the scorers/ subpackage. Add the directory to [tool.hatch.build.targets.wheel] if it has non-Python files.
If the grader wants torch, grelu, etc., put them in optional-dependencies and have the grader fall back gracefully when missing — see examples/dna_design/grader/pyproject.toml. Then grader.setup becomes ["uv pip install -e ./grader[ml]"] for the full version.
grader.privateAnswer keys, hidden test fixtures, and anything the agent must not see go under grader.private in task.yaml — CORAL copies those paths into .coral/private/<name> (which every agent runtime is denied read access to) and the grader reads them via self.private_dir:
grader:
private:
- "taskdata" # a sibling of grader/ → .coral/private/taskdata, read via Path(self.private_dir) / "taskdata"The single rule: everything inside the grader/ package is visible to agents, and secrets live in grader.private outside grader/. The whole grader source is surfaced read-only to agents at <shared_dir>/grader/ (a symlink to the real package) so they can read how they're scored — so anything under grader/ is readable, including a grader.private path that points inside it (it gets copied to .coral/private/ and leaked through the surfaced source). Keep taskdata/ as a sibling of grader/ (declare it taskdata, resolving to <task_dir>/taskdata); coral validate errors if a private path is inside the package.
Non-secret bundled data (lookup tables, reference configs, helper modules) may live inside grader/ and be read via Path(__file__).parent / ... — just remember it's visible, so never put a secret there. E.g. examples/ADRS/txn_scheduling/ puts helper modules on sys.path that way.
Fields that must be set; everything else has a sensible default.
task:
name: "My Task" # shows up in results/<slug>/
description: | # rendered into CORAL.md, agents read this
What the agent should do.
Reference the program file by name (e.g. solution.py and its run() signature).
tips: | # optional, also rendered into CORAL.md
- Eval timeout is N seconds.
- Constraints / scoring details / known baselines.
grader:
entrypoint: "<task>_grader.grader:Grader" # required
setup:
- "uv pip install -e ./grader" # runs once in .coral/private/grader_venv/
timeout: 600 # seconds; 0 disables, default 300
direction: maximize # or minimize — controls leaderboard ordering
args: # arbitrary dict, read inside grader as self.args
program_file: "solution.py"
private: [] # extra files copied into .coral/private/ (hidden from agents)
parallel:
max_workers: 1 # bump only when the grader is concurrency-safe
max_pending_per_agent: 1 # cap on in-flight submissions per agent
agents:
count: 1 # raise once the task is known to be stable
runtime: claude_code # claude_code | codex | cursor_agent | kiro | opencode
model: sonnet # default depends on runtime; see coral/agent/registry.py
workspace:
results_dir: "./results" # where each run lands
repo_path: "./examples/<task>/seed" # MUST point at the seed/ dir
setup: # runs in each agent worktree before agents start
- "uv pip install numpy" # task-runtime deps go here, NOT in grader/setup
run:
verbose: false
ui: false
session: tmux # local | tmux | dockerThe examples/README.md documents the full schema with every default. When in doubt, look there before adding fields.
coral validate examples/<task>This:
task.yaml and reports schema errors..coral/private/grader_venv/ and runs grader.setup.seed/ into a tempdir and runs the grader against it once.If coral validate succeeds, the grader can score the seed. That's the single most important checkpoint — most "agent stuck" issues trace back to a grader that crashes on the seed.
After validation, smoke-test with one agent:
coral start -c examples/<task>/task.yaml agents.count=1 run.session=local
# Wait for one eval, then:
coral stopAdd a one-line entry to the table in examples/README.md and a short ### <task> section under "Details" with the bullet points the others use (Agents / Timeout / Session). Skip this only for throwaway local tasks.
| Mistake | Symptom | Fix |
|---|---|---|
repo_path points at examples/<task>/ instead of examples/<task>/seed/ | Grader sees task.yaml and grader/ in codebase_path | Always point repo_path at the seed dir. |
direction: maximize for a loss / minimize for a benchmark ratio | Leaderboard ordered backwards | Score = "ratio against benchmark, >1 is better" → maximize. Score = "raw error" → minimize. |
Hidden answer key under seed/ or anywhere inside the grader/ package | Agents read it and game the score — seed/ is their repo, and the whole grader/ source is surfaced at <shared_dir>/grader/ | Put it under grader.private outside grader/ (e.g. a sibling taskdata/), read via self.private_dir. coral validate errors on a private path inside grader/. |
Grader writes results under self.codebase_path and reads them later | Files vanish — daemon force-removes the worktree after each eval | Write under self.eval_logs_dir. |
Grader uses sys.executable to run the agent's program | Misses task-specific deps installed via workspace.setup | Use self.get_python_command() (it switches to uv run --project when the codebase has a pyproject.toml). |
Heavy deps in main grader dependencies | coral validate is slow / fails on machines without the GPU stack | Move to optional-dependencies and fall back gracefully. See dna_design's [ml] extra. |
grader.setup tries to install task-runtime deps | The grader venv has them but the agent's worktree doesn't | Task-runtime deps go in workspace.setup. Grader-only deps go in grader.setup. |
parallel.max_workers > 1 with a non-concurrency-safe grader | Sporadic failures when two evals collide on Docker ports / GPU / scratch dirs | Leave at 1 unless the grader is provably safe. |
Forgetting coral validate | Agents start, fail every eval with the same error | Always validate first. |
© Human-Agent-Society, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .claude/skills/coral-new-task of Human-Agent-Society/CORAL.
Open the folder on GitHubat commit 0123dfb
Coral New Task next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Coral New Task this skillHuman-Agent-Society/CORAL | 1.1k | — | ~3.4k | Automated safety check: Pass | Apache-2.0 | |
| Acceptance Evidence for Deliverieslobehub/lobehub | 83k | — | ~9.7k | Automated safety check: Pass | Apache-2.0 | |
| Senpi Agent QA Harnesscode-yeongyu/senpi | 474 | — | ~2.7k | Automated safety check: Notes | MIT | |
| tmux Real User TestingQwenLM/qwen-code | 28k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | |
| StandardsItamarZand88/CLI-Anything-WEB | 231 | — | ~4.2k | Automated safety check: Pass | MIT | |
| Dev IssueFHIR/fhir-codegen | 155 | — | ~4k | Automated safety check: Pass | MIT |
lobehub/lobehub
Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI.
code-yeongyu/senpi
Checks changes to the senpi coding agent by driving the real CLI from source in an isolated sandbox, over RPC, terminal UI, mock model and CLI smoke channels.
QwenLM/qwen-code
Drives Qwen Code in a real tmux session the way a user would and saves a readable step-by-step transcript of each screen for maintainers to review.
ItamarZand88/CLI-Anything-WEB
Runs Phase 4 review/publish/verify for a cli-web- CLI: implementation review by 3 parallel agents, the tiered quality checklist (Tier 1 critical fail-fast, then comprehensive), pip install + smoke…
FHIR/fhir-codegen
Publishes a slot's feature request or bug report to GitHub as an issue, and keeps that issue in sync, in the role of a release-minded engineer.
daydreamlive/scope
Test Daydream Scope through its stdio MCP server, including disconnected startup, connecttoscope, direct Python stdio client fallback, and lightweight smoke tests.
Human-Agent-Society/CORAL
Author a new CORAL task — the three pieces that must line up (task.yaml, seed/, a packaged grader/), the coral init → coral validate → smoke-test loop, and how to pick a grader pattern (stdout…
Human-Agent-Society/CORAL
Run and manage CORAL experiments from the operator side — launch agents with coral start (dotlist overrides, model/count, tmux vs local), monitor with coral status / coral log / coral show / the web…
Human-Agent-Society/CORAL
Verify and debug changes to CORAL itself — smallest reproduce loop per area (grader / daemon / CLI / hooks / manager / workspace / hub / template / config / web), where to look when something breaks…
Human-Agent-Society/CORAL
A skill your agent uses when preparing, reviewing, resolving conflicts for, or merging a CORAL release pull request from the long-lived dev branch into main.
Human-Agent-Society/CORAL
One-time machine setup after installing the coral CLI — register local agent runtimes as named bindings with coral setup / coral setup agent, validate them with coral agents doctor (incl.
Human-Agent-Society/CORAL
Add a new component to the CORAL framework itself — a new agent runtime under coral/agent/builtin/ (claudecode/codex/cursoragent style), a new CLI command in coral/cli/, a new bundled skill or…
Categories
End-to-end recipe for adding a new task under examples/ — the three pieces that have to line up (task.yaml, seed/, and grader/), what to put in each, the TaskGrader API surface, the coral validate →…. Coral New Task is an agent skill from Human-Agent-Society/CORAL.yaml, seed/, and grader/), what to put in each, the TaskGrader API surface, the coral validate → smoke-test loop, and the common mistakes (repopath pointing at the wrong dir, score direction backwards, hidden answer keys leaking into seed/, grader writing to codebasepath which the daemon force-removes, private-vs-public confusion, missing run() signature).
Coral New Task fits situations like: the user wants to add a new CORAL task; port an existing benchmark into CORAL.
Run `npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a claude-code`. Or copy the skill folder (.claude/skills/coral-new-task in Human-Agent-Society/CORAL) into .claude/skills/coral-new-task in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a codex`. Or copy the skill folder (.claude/skills/coral-new-task in Human-Agent-Society/CORAL) into .agents/skills/coral-new-task in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Human-Agent-Society/CORAL --skill coral-new-task -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/coral-new-task, .gemini/skills/coral-new-task, .github/skills/coral-new-task and .opencode/skills/coral-new-task in your project.
Going by SKILL.md and its folder, Coral New Task needs the command-line tools its instructions call (uv). Our summary lists: Python 3; Docker.
SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Coral New Task is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.4k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Coral New Task: Acceptance Evidence for Deliveries (lobehub/lobehub, 83k stars), Senpi Agent QA Harness (code-yeongyu/senpi, 474 stars), tmux Real User Testing (QwenLM/qwen-code, 28k stars) and Standards (ItamarZand88/CLI-Anything-WEB, 231 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Human-Agent-Society (a GitHub organization) maintains it in Human-Agent-Society/CORAL, which has 1,060 GitHub stars. The repository holds 7 skills in this directory. The repository was last updated on September 8, 2026.
Source: Human-Agent-Society/CORAL on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.