Agent skill

Sycophancy Challenger

by mohitagw15856 in mohitagw15856/pm-claude-skills

Flip Claude’s default from validation to adversarial critique.

MITAuto-check passed

Install Sycophancy Challenger

skills CLI
$ npx skills add mohitagw15856/pm-claude-skills --skill sycophancy-challenger -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install mohitagw15856/pm-claude-skills sycophancy-challenger --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/mohitagw15856/pm-claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/sycophancy-challenger .claude/skills/sycophancy-challenger && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
sycophancy-challenger
GitHub stars
1.4k
Token cost
~2.3k tokens
SKILL.md length
1,159 words
Files
1
Skills in repo
1,348
Repo updated
First seen
Licence
MIT

At a glance

Flip Claude’s default from validation to adversarial critique.

  • Works in 5 steps: Assume the idea hasn't been stress-tested → Find the strongest case against it → Identify the weakest element → …
  • You are about to make a high-stakes decision
  • SKILL.md covers Required Inputs, Output Structure, Instructions for Claude and Quality Checks, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Sycophancy Challenger is an agent skill from mohitagw15856/pm-claude-skills. Flip Claude’s default from validation to adversarial critique. Use when you are about to make a high-stakes decision, commit to a plan, or pitch something you have not stress-tested. Produces structured challenges, steelmanned counter-arguments, and the strongest case against your position — a genuine thinking partner, not a mirror.

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: 1255 professional Agent Skills for Claude, ChatGPT, Gemini, Cursor & Codex — PRDs, postmortems, leases, medical bills, layoffs, go-bags, new countries. Plain markdown, MIT, in… The licence is MIT.

When your agent uses it

  • You are about to make a high-stakes decision
  • Commit to a plan
  • Pitch something you have not stress-tested

Example prompts

  • “/sycophancy-challenger”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Assume the idea hasn't been stress-tested
  2. Find the strongest case against it
  3. Identify the weakest element
  4. Surface the required assumptions
  5. Report what holds up (only if true)

What it can do on your machine

Read from SKILL.md and the folder at commit 1cbf1f0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Sycophancy Challenger loads about 2.3k tokens when it runs. Until then it costs about 89 tokens; SKILL.md has 1,159 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~89
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from mohitagw15856/pm-claude-skills at commit 1cbf1f0, republished under its MIT licence (© mohitagw15856). 1,159 words, ~2,260 tokens.

Download SKILL.mdSave it as .claude/skills/sycophancy-challenger/SKILL.md (or your agent's skills folder).
name
sycophancy-challenger
description
Flip Claude’s default from validation to adversarial critique. Use when you are about to make a high-stakes decision, commit to a plan, or pitch something you have not stress-tested. Produces structured challenges, steelmanned counter-arguments, and the strongest case against your position — a genuine thinking partner, not a mirror.

Sycophancy Challenger

Claude defaults to validating. You bring a decision, it finds three reasons your instinct is solid, and you leave more confident but not more right. That's actively dangerous when the stakes are high — a hiring call, a pricing change, a strategy pivot, a public commitment. This skill flips the default: Claude argues against your idea first, holds its position under pushback, and only concedes when you give it new evidence. Not when you express displeasure.

Credit: Originally created by Joel Salinas (Leadership in Change) — adapted and extended for this library.


Required Inputs

InputFormatNotes
Your idea, decision, plan, or assumptionDescribe it in plain languageMore context = sharper challenge. Include reasoning if you have it.

No other setup required. Activating the skill is enough — describe your idea and Claude will challenge it immediately.


Output Structure

Every response in this mode follows this exact format:

## Strongest Case AGAINST This

[The single most damaging criticism of the idea. Not a list of concerns — the
one argument that, if true, would kill this. Stated directly, without softening.]


## The Weakest Element

[The specific part of the idea most likely to fail, be wrong, or break under
real-world conditions. Named precisely. Not "execution risk" — the actual thing.]


## What You'd Need to Prove to Make This Work

[The assumptions that must be true for this idea to succeed. Written as testable
claims, not as encouragement. If an assumption can't be tested, that's noted.]


## What I Can't Find Fault With

[Only appears when a genuine search finds nothing damaging. States clearly what
holds up and why — doesn't invent weak praise to fill the section. If everything
is actually fine, says so plainly and explains why the challenge came up short.]

No additional sections. No summary. No "overall, this is a solid idea." The format ends when the four sections are complete.


Instructions for Claude

On activation

Do not open with agreement, validation, or any form of "I see where you're coming from." Begin the challenge immediately. The first word of your response should advance the criticism, not soften the user's expectations.

Step 1: Assume the idea hasn't been stress-tested

Treat the idea as if the user believes in it strongly and has not actively looked for reasons it fails. Your job is to be the adversary they didn't have in the room.

Step 2: Find the strongest case against it

Not a balanced view. Not pros and cons. The strongest case against. Ask:

  • What's the most likely way this fails?
  • What's the assumption that, if wrong, makes everything else irrelevant?
  • Who would argue against this, and what's the best version of their argument?
  • What does this idea get wrong about how people, markets, or systems actually behave?

State the strongest case directly. Do not list multiple criticisms in this section — lead with the one that does the most damage.

Step 3: Identify the weakest element

This is different from the strongest case against. The weakest element is the most fragile specific component — the thing most likely to crack under execution, scrutiny, or changed conditions. Name it precisely. Examples of insufficient answers:

  • "The timeline might be tight" → insufficient
  • "The assumption that customers will pay $99/month before experiencing the product is the element most likely to break this, because you have no evidence of willingness-to-pay at that price point" → correct level of specificity
Step 4: Surface the required assumptions

List what must be true for this to work. Write each assumption as a testable claim:

For this to work, the following must be true:
1. [Assumption stated as a claim that can be verified or falsified]
2. [Assumption stated as a claim]
3. [Assumption stated as a claim]

If an assumption cannot be tested — it's based on hope, belief, or unprovable prediction — flag it explicitly: "This assumption cannot currently be tested. That's a risk."

Step 5: Report what holds up (only if true)

Search genuinely for what the idea gets right or where the challenge fails. If you find it, state it clearly. If you can't find a real flaw, say exactly that: "I've looked for the failure points and I can't find them. Here's what actually holds up: [specific things]." Do not invent praise. Do not invent flaws either.

Handling pushback

If the user pushes back:

  • New evidence or new information: update your position based on the evidence. State what changed and why.
  • Emotional pushback, repetition, or displeasure: do not move. Restate the criticism calmly. Example: "I understand you feel strongly about this — I'm not backing off the point about X because that hasn't changed. If there's something I'm missing, tell me what it is."
  • A clarification that changes the picture: acknowledge the clarification, adjust if warranted, and explain exactly what the clarification changed.

Do not soften a position because the user seems upset. Do not move back to validation mode mid-conversation.

When the skill ends

The session is complete when the user has either:

  1. Strengthened their idea by addressing the core criticism with real evidence or a genuine plan adjustment, or
  2. Identified a real flaw they're going to fix.

Not when they've expressed satisfaction. Not when a certain number of exchanges have happened. The measure is whether something actually changed or was genuinely defended.

Show full SKILL.md (459 more words)Show less
Prohibitions

These prohibitions do more work than the rules above. Follow them absolutely:

  • Never open with agreement or validation. Not "That's an interesting approach," not "I can see why you'd think that." Start with the challenge.
  • Never say "great question," "great point," or "I see where you're coming from" as a lead. These are validation openers, not neutral transitions.
  • Never soften a criticism with "however, there are also positives." If the positives are real, they go in the "What I Can't Find Fault With" section, not as a counterweight to every criticism.
  • Never back down because the user expressed displeasure. Only move if given new evidence.
  • Never invent a flaw that isn't real. If the idea is actually solid, say so. Inventing fake criticisms is as useless as fake validation.
  • Never use the word "valid" to describe the user's perspective mid-challenge. It's a validation signal disguised as a neutral word.

Quality Checks

  • Response opened with the challenge — not with a softening phrase or acknowledgment
  • "Strongest Case Against" section contains one argument, not a list
  • "Weakest Element" is specific — names the actual component, not a category of risk
  • "What You'd Need to Prove" lists testable assumptions, not encouragement
  • Untestable assumptions are explicitly flagged as risks
  • "What I Can't Find Fault With" only appears if the search was genuine and something held up
  • No invented flaws — every criticism connects to something real in what the user described
  • Pushback was met with a position restatement, not a retreat (unless new evidence was provided)
  • The session ended because something changed or was genuinely defended — not because the user seemed satisfied
  • None of the prohibited phrases or patterns appear anywhere in the response

Anti-Patterns

  • Do not open with a softening phrase or acknowledgment before the challenge — the first sentence must be the critique
  • Do not retreat from a position when the user pushes back without providing new evidence — update only when genuinely persuaded
  • Do not invent flaws — every criticism must connect to something real in what the user described
  • Do not provide a list of weak objections — identify the single strongest case against the idea
  • Do not end the session because the user seems satisfied — end only when something genuinely changed or was defended

Example Trigger Phrases

  • "Use the sycophancy-challenger skill — here's my plan: [describe it]"
  • "Challenge this idea before I commit to it: [describe it]"
  • "I've already decided to do X — tell me why I'm wrong"
  • "Be the devil's advocate on this hire: [describe the candidate and the role]"
  • "I'm about to pitch this to investors — tear it apart first: [describe it]"
  • "Don't validate this, challenge it: [idea or assumption]"
  • "Stress-test this strategy: [describe it]"
  • "What's the strongest argument against doing this: [decision]"
  • "I think I'm right about X — what am I missing?"

© mohitagw15856, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/sycophancy-challenger of mohitagw15856/pm-claude-skills.

Open the folder on GitHubat commit 1cbf1f0

Compare with similar skills

Sycophancy Challenger next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Sycophancy Challenger compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Sycophancy Challenger this skillmohitagw15856/pm-claude-skills1.4k—~2.3kAutomated safety check: PassMIT
Design Critiquepaperclipai/paperclip100k—~1.2kAutomated safety check: PassMIT
Agent Challengesruvnet/ruflo74k2 repos~995Automated safety check: PassMIT
Critique Theaternexu-io/open-design100k—~640Automated safety check: PassApache-2.0
Critiqueplugin87/ux-ui-agent-skills1.6k—~1.3kAutomated safety check: PassMIT
Challengealirezarezvani/claude-skills28k1 repos~1.7kAutomated safety check: PassMIT

Similar skills

  • Design Critique

    paperclipai/paperclip

    Give a structured product design critique — user job clarity, hierarchy, affordance, error states, accessibility, and consistency — focused on what to change, in what order, and why.

    100k GitHub stars~1.2k tokensUpdated today
    Media & CreativeAuto-check passed
  • Agent Challenges

    ruvnet/ruflo

    Agent skill for challenges - invoke with $agent-challenges. An agent skill from ruvnet/ruflo.

    74k GitHub starsUsed in 2 repos~995 tokens
    Auto-check passed
  • Critique Theater

    nexu-io/open-design

    Five-dimension design quality review — score the artifact against craft, brand, accessibility, and copy, then fix what falls short before handing it over.

    100k GitHub stars~640 tokensUpdated yesterday
    Frontend & DesignAuto-check passed
  • Critique

    plugin87/ux-ui-agent-skills

    Adversarial design critique of the current work — render it, look at it, and argue for rejection.

    1.6k GitHub stars~1.3k tokensUpdated 3 days ago
    Frontend & DesignAuto-check passed
  • Challenge

    alirezarezvani/claude-skills

    Pre-mortem plan analysis. An agent skill from alirezarezvani/claude-skills.

    28k GitHub starsUsed in 1 repo~1.7k tokens
    Business, Finance & HRAuto-check passed
  • Design Critique

    Owl-Listener/designer-skills

    Facilitate a structured team critique — framing, feedback rules, and actionable outcomes.

    2.9k GitHub starsUsed in 1 repo~513 tokens
    Media & CreativeAuto-check passed

More from mohitagw15856/pm-claude-skills

All 1,348 skills in this repo
  • Car Tco

    mohitagw15856/pm-claude-skills

    Compare the total cost of car ownership across buy-new, buy-used, lease, and keep-your-current-car — depreciation, insurance, maintenance ramp, and fuel over a real horizon, not just the monthly…

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Cs Health Scorecard

    mohitagw15856/pm-claude-skills

    Build a customer health scorecard for a specific account. An agent skill from mohitagw15856/pm-claude-skills.

    1.4k GitHub stars~2.4k tokensUpdated yesterday
    Auto-check passed
  • Exit Waterfall

    mohitagw15856/pm-claude-skills

    Compute who gets what at each exit price from a cap table — liquidation preferences, conversion points, and where the founders' share collapses.

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Feature Prioritisation

    mohitagw15856/pm-claude-skills

    Apply prioritisation frameworks (RICE, MoSCoW, Kano, ICE, Opportunity Scoring) to rank features and backlog items.

    1.4k GitHub stars~2k tokensUpdated yesterday
    Auto-check passed
  • Fire Number

    mohitagw15856/pm-claude-skills

    Compute a financial-independence (FIRE) target and years-to-reach with every assumption labeled as an assumption — plus a sensitivity table instead of a single false-precision answer.

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Freelance Rate

    mohitagw15856/pm-claude-skills

    Derive a freelance day/hourly rate backwards from target income, honest billable utilization, overhead, and the self-employment tax premium — the arithmetic that proves a rate is not salary÷2000.

    1.4k GitHub stars~1.2k tokensUpdated yesterday
    Auto-check passed

Questions about Sycophancy Challenger

What does Sycophancy Challenger do?

Flip Claude’s default from validation to adversarial critique. Sycophancy Challenger is an agent skill from mohitagw15856/pm-claude-skills. Flip Claude’s default from validation to adversarial critique.

When should I use Sycophancy Challenger?

Sycophancy Challenger fits situations like: you are about to make a high-stakes decision; commit to a plan; pitch something you have not stress-tested.

How do I install Sycophancy Challenger in Claude Code?

Run `npx skills add mohitagw15856/pm-claude-skills --skill sycophancy-challenger -a claude-code`. Or copy the skill folder (skills/sycophancy-challenger in mohitagw15856/pm-claude-skills) into .claude/skills/sycophancy-challenger in your project. Claude Code loads it when a task matches its description.

How do I install Sycophancy Challenger in Codex?

Run `npx skills add mohitagw15856/pm-claude-skills --skill sycophancy-challenger -a codex`. Or copy the skill folder (skills/sycophancy-challenger in mohitagw15856/pm-claude-skills) into .agents/skills/sycophancy-challenger in your project. Codex loads it when a task matches its description.

Can I use Sycophancy Challenger in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mohitagw15856/pm-claude-skills --skill sycophancy-challenger -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/sycophancy-challenger, .gemini/skills/sycophancy-challenger, .github/skills/sycophancy-challenger and .opencode/skills/sycophancy-challenger in your project.

What does Sycophancy Challenger need to run?

SKILL.md names no scripts, command-line tools or credentials: Sycophancy Challenger is instructions for the agent only.

Does Sycophancy Challenger access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Sycophancy Challenger safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Sycophancy Challenger use?

Sycophancy Challenger is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Sycophancy Challenger use?

About 2.3k tokens (SKILL.md is roughly 9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Sycophancy Challenger?

Skills that share tags, products or a category with Sycophancy Challenger: Design Critique (paperclipai/paperclip, 100k stars), Agent Challenges (ruvnet/ruflo, 74k stars), Critique Theater (nexu-io/open-design, 100k stars) and Critique (plugin87/ux-ui-agent-skills, 1.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Sycophancy Challenger?

mohitagw15856 (a GitHub user) maintains it in mohitagw15856/pm-claude-skills, which has 1,434 GitHub stars. The repository holds 1,348 skills in this directory. The repository was last updated on October 9, 2026.

Source: mohitagw15856/pm-claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.