Arize Evaluator
github/awesome-copilot
Handles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, running evaluations on spans or experiments, managing tasks, trigger-run operations, column mapping, and…
A skill your agent uses when a decision feels obvious or everyone agrees, when evaluating omeone's proposal or pitch, when in a team debate that's going nowhere, or when you suspect confirmation…
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidence --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .claude/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .claude/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidenceType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidence --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .agents/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .agents/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidence --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .cursor/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .cursor/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/apple-ouyang/book-to-skill.git --path skills/seeking-disconfirming-evidence--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidence --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .gemini/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .gemini/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidenceInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .github/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .github/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install apple-ouyang/book-to-skill seeking-disconfirming-evidence --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/apple-ouyang/book-to-skill.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/seeking-disconfirming-evidence .opencode/skills/seeking-disconfirming-evidence && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "seeking-disconfirming-evidence" agent skill from https://github.com/apple-ouyang/book-to-skill/tree/main/skills/seeking-disconfirming-evidence into .opencode/skills/seeking-disconfirming-evidence/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "seeking-disconfirming-evidence", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
seeking-disconfirming-evidenceA skill your agent uses when a decision feels obvious or everyone agrees, when evaluating omeone's proposal or pitch, when in a team debate that's going nowhere, or when you suspect confirmation…
Seeking Disconfirming Evidence is an agent skill from apple-ouyang/book-to-skill. Use when a decision feels obvious or everyone agrees, when evaluating omeone's proposal or pitch, when in a team debate that's going nowhere, or when you suspect confirmation bias in your own thinking.
Its SKILL.md is about 650 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/cases.md`).
The repository describes itself as: 把书拆成 AI Agent 可执行的 Skill,让书中的智慧变成你的决策副驾驶 | Turn books into executable AI Agent Skills. The licence is MIT.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit a24960a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Seeking Disconfirming Evidence loads about 654 tokens when it runs, and up to ~2.6k if it reads all its reference files. Until then it costs about 58 tokens; SKILL.md has 140 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from apple-ouyang/book-to-skill at commit a24960a, republished under its MIT licence (© apple-ouyang). 140 words, ~654 tokens.
.claude/skills/seeking-disconfirming-evidence/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.打破确认偏误:主动寻找能推翻自己结论的证据,用「条件法」把辩论变成协作探索。
研究发现,没有独立董事质疑的收购案,CEO 支付的溢价平均高出 41%——越是没人反对,越要警惕。
通用 CEO 斯隆的做法:「先生们,我们已经达成统一意见了?那把这个问题推迟到下次会议,让我们多些时间提出不同看法。」
指定一个人专门负责反对,而不是等待自然反对声音出现。
操作:在会议开始前,明确指定谁扮演反对角色,或者轮流担任。
来源:罗杰·马丁在因梅特矿业公司的实践。
高管想关闭铜矿,矿区经理想继续开采,双方争了几个小时毫无进展。马丁打断说:
「不要再争谁对谁错了。我们一次考虑一个选择,然后问:这个选择必须具备怎样的条件,才能成为正确的答案?」
结果:高管列出了「继续开矿」合理所需的生产目标;矿区经理认同了「如果铜价不反弹,关闭就是最优解」。会议结束时,5 个选项各自的成立条件都达成了共识。
操作模板:
对于选项 A,它要成为最佳选择,需要以下条件为真:
1. ___
2. ___
对于选项 B,它要成为最佳选择,需要以下条件为真:
1. ___
2. ___
现在我们来讨论:哪些条件更可能实现?NetApp 创始人戴夫·希茨的反直觉技巧:
「捍卫一个决定的最佳方法,是指出它的缺点。」
当有人反对你的 A 计划时,不要重复自己的论据,而是:
效果:对方会从防御状态转为倾听状态,因为他感受到了被理解。
对专家:刨根问底,问事实性细节
不要问「你有经验吗」,要问具体事实:
研究证明:问「它存在什么样的毛病?」比问「它没有任何毛病,对吗?」多获得 28% 的真实信息(89% vs 61%)。
对用户/小白:问开放式问题,不要引导
不要问「是不是这里痛?」,要问「你能描述一下是什么感觉吗?」
引导式问题会让对方顺着你的预设回答,你收集到的是你想听的,不是真相。
当你意识到自己有一个「理所当然」的假设,但从未验证过,可以故意违反它来测试。
操作:列出你「从未质疑过」的 2-3 个假设,选一个成本最低的,设计一个小实验来故意违反它。
场景:两派人各执一词,会议开了两小时没结论
操作:
场景:下属或合作方来推销一个方案,你有疑虑
操作:
场景:你已经倾向某个决定,想确认自己没有确认偏误
操作:
© apple-ouyang, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (references) in skills/seeking-disconfirming-evidence of apple-ouyang/book-to-skill.
Open the folder on GitHubat commit a24960a
Seeking Disconfirming Evidence next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Seeking Disconfirming Evidence this skillapple-ouyang/book-to-skill | 161 | — | ~654 | Automated safety check: Pass | MIT | |
| Arize Evaluatorgithub/awesome-copilot | 40k | 1 repos | ~8.1k | Automated safety check: Notes | MIT | |
| LLM Evaluationdavila7/claude-code-templates | 33k | 12 repos | ~3.5k | Automated safety check: Pass | MIT | |
| Agent Evaluationsickn33/agentic-awesome-skills | 47k | 1 repos | ~2k | Automated safety check: Pass | MIT | |
| EvaluatorsArize-ai/phoenix | 12k | — | ~1.7k | Automated safety check: Pass | Custom licence | |
| Agent Evaluation Reportingsickn33/agentic-awesome-skills | 47k | 1 repos | ~2.1k | Automated safety check: Pass | MIT |
github/awesome-copilot
Handles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, running evaluations on spans or experiments, managing tasks, trigger-run operations, column mapping, and…
davila7/claude-code-templates
Master comprehensive evaluation strategies for LLM applications, from automated metrics to human evaluation and A/B testing.
sickn33/agentic-awesome-skills
Evaluate agent behavior with versioned cases and explicit verifiers.
Arize-ai/phoenix
Author or refine a Phoenix evaluator — code or LLM-as-a-judge — that scores a run's output.
sickn33/agentic-awesome-skills
A skill your agent uses when summarizing agent evaluations where autonomous, assisted, failed, timed-out, or invalid outcomes must remain distinct and comparable.
PostHog/posthog
Author continuously-running online evaluations in PostHog AI observability, grounded in real failure modes you've identified.
apple-ouyang/book-to-skill
A skill your agent uses when facing a significant decision, evaluating options, or feeling stuck between choices.
apple-ouyang/book-to-skill
This skill should be used when the user asks to "create a team", "create agent team", "蜂群", "swarm", "团队协作", "parallel agents", "spawn teammates", "多 agent 协作", "组建团队", or executes "/swarm".
apple-ouyang/book-to-skill
A skill your agent uses when committing resources to a project, investment, or relationship.
apple-ouyang/book-to-skill
A skill your agent uses when tempted to predict, analyze, or debate whether something will work instead of testing it.
apple-ouyang/book-to-skill
把一本书拆解成 Claude Code 可执行的 Skill。当用户说"拆书"、"把这本书变成 Skill"、"book to skill"、"extract skill from book" 时触发。
A skill your agent uses when a decision feels obvious or everyone agrees, when evaluating omeone's proposal or pitch, when in a team debate that's going nowhere, or when you suspect confirmation…. Seeking Disconfirming Evidence is an agent skill from apple-ouyang/book-to-skill. Use when a decision feels obvious or everyone agrees, when evaluating omeone's proposal or pitch, when in a team debate that's going nowhere, or when you suspect confirmation bias in your own thinking.
Seeking Disconfirming Evidence fits situations like: A decision feels obvious; everyone agrees; evaluating omeones proposal; in a team debate thats going nowhere.
Run `npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a claude-code`. Or copy the skill folder (skills/seeking-disconfirming-evidence in apple-ouyang/book-to-skill) into .claude/skills/seeking-disconfirming-evidence in your project. Claude Code loads it when a task matches its description.
Run `npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a codex`. Or copy the skill folder (skills/seeking-disconfirming-evidence in apple-ouyang/book-to-skill) into .agents/skills/seeking-disconfirming-evidence in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add apple-ouyang/book-to-skill --skill seeking-disconfirming-evidence -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/seeking-disconfirming-evidence, .gemini/skills/seeking-disconfirming-evidence, .github/skills/seeking-disconfirming-evidence and .opencode/skills/seeking-disconfirming-evidence in your project.
SKILL.md names no scripts, command-line tools or credentials: Seeking Disconfirming Evidence is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Seeking Disconfirming Evidence is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 654 tokens (SKILL.md is roughly 2.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Seeking Disconfirming Evidence: Arize Evaluator (github/awesome-copilot, 40k stars), LLM Evaluation (davila7/claude-code-templates, 33k stars), Agent Evaluation (sickn33/agentic-awesome-skills, 47k stars) and Evaluators (Arize-ai/phoenix, 12k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
apple-ouyang (a GitHub user) maintains it in apple-ouyang/book-to-skill, which has 161 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on February 25, 2026.
Source: apple-ouyang/book-to-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.