Agent skill

Arecon Evidence Standards

by brycewang-stanford in brycewang-stanford/Awesome-Journal-Skills

A skill your agent uses when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair…

MITAuto-check passedResearch & Science

Install Arecon Evidence Standards

skills CLI
$ npx skills add brycewang-stanford/Awesome-Journal-Skills --skill arecon-evidence-standards -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install brycewang-stanford/Awesome-Journal-Skills arecon-evidence-standards --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/brycewang-stanford/Awesome-Journal-Skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/Annual-Review-of-Economics-Skills/skills/arecon-evidence-standards .claude/skills/arecon-evidence-standards && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
arecon-evidence-standards
GitHub stars
1.2k
Token cost
~1.7k tokens
SKILL.md length
789 words
Files
1
Skills in repo
2,387
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair…

  • Appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes
  • SKILL.md covers When to trigger, Appraising evidence you did…, Execution bridge — when rating… and Weighing, not vote-counting, plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • When calibrating comprehensiveness

What it does

Arecon Evidence Standards is an agent skill from brycewang-stanford/Awesome-Journal-Skills. Use when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair treatment of competing views. Weighs evidence and audits even-handedness; it does not design the spine (arecon-organizing-framework) or run your own identification.

Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Research & Science. The repository describes itself as: Journal-specific Claude Code/Codex skill packs covering mainstream journals — AER, QJE, Nature, Cell, 管理世界, 经济研究 & 200+ more — your fast track to getting published. | 覆盖主流期刊的… The licence is MIT.

When your agent uses it

  • Appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes
  • When calibrating comprehensiveness
  • Fair treatment of competing views

Example prompts

  • “/arecon-evidence-standards”

What it can do on your machine

Read from SKILL.md and the folder at commit 932eb23. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Arecon Evidence Standards loads about 1.7k tokens when it runs. Until then it costs about 91 tokens; SKILL.md has 789 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~91
When it runs · the whole SKILL.md, loaded when a task matches
~1.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from brycewang-stanford/Awesome-Journal-Skills at commit 932eb23, republished under its MIT licence (© brycewang-stanford). 789 words, ~1,733 tokens.

Download SKILL.mdSave it as .claude/skills/arecon-evidence-standards/SKILL.md (or your agent's skills folder).
name
arecon-evidence-standards
description
Use when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair treatment of competing views. Weighs evidence and audits even-handedness; it does not design the spine (arecon-organizing-framework) or run your own identification.

Evidence Appraisal & Balance (arecon-evidence-standards)

When to trigger

  • The framework is set and you are filling cells with conflicting empirical results
  • You must decide whether "the literature finds X" is actually supported, or only loosely
  • The field has rival schools, a live controversy, or estimates that disagree
  • You are a contributor to this literature and worry the review tilts toward your own work

Appraising evidence you did not produce

An ARE review runs no identification of its own — there is no design to defend, no replication package of your data. Instead you act as the field's referee-of-record for adjacent readers: you judge how much weight each primary study can bear so the review weighs the evidence correctly. Make the appraisal explicit, not implicit:

  • DID / event study: does the cited paper use TWFE on staggered timing without addressing heterogeneity bias? A study that predates Callaway–Sant'Anna / Sun–Abraham corrections may not support the magnitude you attribute to it — flag it.
  • IV: is the first stage strong and the exclusion restriction defended, or is it a convenience instrument? Weak-IV results carry less weight in your synthesis.
  • RDD: density/manipulation checks, bandwidth robustness — a fragile RDD is a fragile data point.
  • Structural / calibration: are the parameters identified and the counterfactual policy-invariant, or is it calibration presented as estimation?
  • Experiments: pre-registration, balance, attrition, multiple-hypothesis adjustment, external validity.

You are not re-running these — you are rating their credibility in the evidence matrix so the review's conclusions track the best evidence, not the loudest paper.

Execution bridge — when rating is not enough

A few moments in an ARE review justify running something instead of only rating it: a survey table pools estimates and wants a formal meta-analytic average or publication-bias check; a pivotal magnitude predates the staggered-DiD corrections and its replication data are public, so you can report what callaway_santanna or bacon_decomposition actually does to it rather than speculate; or a weak-instrument worry could be settled by an effective_f_test on the archived first stage. For these, hand off to execution-with-mcp, which maps each design and reviewer objection to the callable StatsPAI / Stata MCP chain (detect_design → fit with as_handle=true → audit_result). Label any such figure as your re-analysis, and never report a number you did not compute.

Weighing, not vote-counting

Conflicting results are reconciled by credibility and by what each study estimates, never by tallying "7 studies positive, 4 negative." Two estimates that disagree often measure different objects (different populations, estimands, time horizons); say so, and let the framework's cells carry the distinction. A pooled "consensus" across non-comparable designs manufactures false agreement that ARE's methodologically literate readers will catch.

Comprehensiveness vs. selectivity: the ARE contract

A review must be comprehensive in coverage yet selective in emphasis — and stay accessible in ~25–40 pages. Tier the corpus:

TierTreatment
Foundational / field-definingdiscussed in text, with what they established and their limits
Important contributionsgrouped and weighed within framework cells; cited with their finding
Confirmatory / incrementalcited in clusters ("see also …") to show coverage without bloating prose
Tangentialcited only where they bear on a specific claim

Comprehensiveness is proven by the citation set + saturation log (arecon-literature-synthesis); selectivity is exercised in the prose.

Show full SKILL.md (271 more words)Show less

Fairness and the self-citation trap

ARE referees are frequently the surveyed authors themselves, so balance is strategic as well as ethical:

  • Steelman every camp. State each school's strongest case in terms its proponents would accept before noting weaknesses.
  • Attribute ideas to originators, not popularizers (a recurring referee complaint).
  • Handle live controversies without resolving by fiat. Lay out the disagreement, what evidence would settle it, and where your own read sits — labelled as your read, not as consensus.
  • Audit self-citation. Your own work appears at the tier its importance to the field warrants — no more; rivals get their strongest statement; a reader who does not know the author cannot tell from the emphasis.

Checklist

  • Each pivotal primary study carries a credibility appraisal (design, identification, robustness)
  • Outdated or fragile designs flagged where the review leans on their magnitudes
  • Conflicting findings reconciled by credibility + estimand, not vote-counting
  • Corpus tiered; prose emphasis matches tier; coverage provable from the saturation log
  • Every rival school stated at its strongest before critique (steelman)
  • Idea attribution traces to originators
  • Live controversies presented with what evidence would settle them; author's read labelled
  • Self-citation audited: own work at warranted tier; emphasis is identity-blind

Anti-patterns

  • Citing a study's headline number without noting its identification is now known to be biased
  • Vote-counting conflicting results instead of weighing credibility and estimand
  • Pooling non-comparable estimates into one "the literature shows…" magnitude
  • Comprehensiveness theatre: equal-length summaries of every paper (no editorial judgment)
  • Strawmanning the camp the author disagrees with
  • A review that doubles as the author's CV (the most-punished ARE balance failure)
  • Declaring a live controversy "resolved" by assertion rather than the evidentiary state

Output format

text
【Credibility appraisal】pivotal studies rated (design/identification/robustness)? Y/N
【Conflict handling】reconciled by credibility + estimand (not vote-count)? Y/N
【Tiering】corpus split foundational/important/confirmatory/tangential? Y/N
【Comprehensiveness】saturation log supports "nothing important missing"? Y/N
【Steelman】each rival school stated at its strongest? Y/N
【Controversy】evidence-to-settle stated; author's read labelled? Y/N
【Self-citation audit】own work at warranted tier; emphasis identity-blind? Y/N
【Next step】→ arecon-tables-figures (who-found-what tables) → arecon-writing-style

© brycewang-stanford, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in Annual-Review-of-Economics-Skills/skills/arecon-evidence-standards of brycewang-stanford/Awesome-Journal-Skills.

Open the folder on GitHubat commit 932eb23

Compare with similar skills

Arecon Evidence Standards next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Arecon Evidence Standards compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Arecon Evidence Standards this skillbrycewang-stanford/Awesome-Journal-Skills1.2k—~1.7kAutomated safety check: PassMIT
Hypothesis Generationspacering-net/codeg3.9k14 repos~3.6kAutomated safety check: NotesMIT
GitHub Deep Researchbytedance/deer-flow84k4 repos~1.3kAutomated safety check: PassMIT
Nature Paper CardYuan1z0825/nature-skills47k2 repos~2.1kAutomated safety check: PassApache-2.0
Content Research Writerweapp-tailwindcss/weapp-tailwindcss1.9k25 repos~3.5kAutomated safety check: PassMIT
Last30daysmvanhorn/last30days-skill64k—~7.9kAutomated safety check: NotesMIT

Similar skills

  • Hypothesis Generation

    spacering-net/codeg

    Structured hypothesis formulation from observations. An agent skill from spacering-net/codeg.

    3.9k GitHub starsUsed in 14 repos~3.6k tokens
    Research & ScienceAuto-check: notes
  • GitHub Deep Research

    bytedance/deer-flow

    Researches a GitHub repository over four rounds using the GitHub API and web search, then writes a structured markdown report with timeline, metrics and Mermaid diagrams.

    84k GitHub starsUsed in 4 repos~1.3k tokens
    Research & ScienceAuto-check passed
  • Nature Paper Card

    Yuan1z0825/nature-skills

    Builds a structured deep-reading card for one scientific paper, covering methods, how experiments support claims, limitations and research ideas, with a script to prepare the source.

    47k GitHub starsUsed in 2 repos~2.1k tokens
    Research & ScienceAuto-check passed
  • Content Research Writer

    weapp-tailwindcss/weapp-tailwindcss

    Assists in writing high-quality content by conducting research, adding citations, improving hooks, iterating on outlines, and providing real-time feedback on each section.

    1.9k GitHub starsUsed in 25 repos~3.5k tokens
    Research & ScienceAuto-check passed
  • Last30days

    mvanhorn/last30days-skill

    Research what people actually say about any topic in the last 30 days.

    64k GitHub stars~7.9k tokensUpdated yesterday
    Research & ScienceAuto-check: notes
  • Peer Review

    spacering-net/codeg

    Structured manuscript/grant review with checklist-based evaluation.

    3.9k GitHub starsUsed in 17 repos~5.9k tokens
    Research & ScienceAuto-check: notes

More from brycewang-stanford/Awesome-Journal-Skills

All 2,387 skills in this repo
  • Aaag Data Analysis

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when running and reporting the analysis for an Annals of the American Association of Geographers manuscript — spatial statistics and modeling, remote-sensing accuracy, or…

    1.2k GitHub stars~1.3k tokensUpdated 14 days ago
    Auto-check passed
  • Aaag Literature Positioning

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when positioning an Annals of the American Association of Geographers manuscript in the literature — engaging geographic scholarship across the relevant area and the…

    1.2k GitHub stars~1.3k tokensUpdated 14 days ago
    Auto-check passed
  • Aaag Rebuttal

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when responding to an Annals of the American Association of Geographers decision letter (major/minor revision) — building a point-by-point response to the subject editor and…

    1.2k GitHub stars~1.4k tokensUpdated 14 days ago
    Auto-check passed
  • Aaag Research Design

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when defending the research design of an Annals of the American Association of Geographers manuscript — spatial/quantitative analysis and GIScience, remote-sensing and…

    1.2k GitHub stars~1.4k tokensUpdated 14 days ago
    Auto-check passed
  • Aaag Review Process

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when you need to understand how the Annals of the American Association of Geographers evaluates a manuscript — double-anonymous review routed through a subject editor by…

    1.2k GitHub stars~1.3k tokensUpdated 14 days ago
    Auto-check passed
  • Aaag Submission

    brycewang-stanford/Awesome-Journal-Skills

    A skill your agent uses when running the final pre-submission preflight for the Annals of the American Association of Geographers via ScholarOne Manuscripts — area/article-type selection…

    1.2k GitHub stars~1.6k tokensUpdated 14 days ago
    Auto-check passed

Questions about Arecon Evidence Standards

What does Arecon Evidence Standards do?

A skill your agent uses when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair…. Arecon Evidence Standards is an agent skill from brycewang-stanford/Awesome-Journal-Skills. Use when appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes, and when calibrating comprehensiveness, balance, and fair treatment of competing views.

When should I use Arecon Evidence Standards?

Arecon Evidence Standards fits situations like: appraising the credibility of the primary studies an Annual Review of Economics (ARE) review synthesizes; when calibrating comprehensiveness; fair treatment of competing views.

How do I install Arecon Evidence Standards in Claude Code?

Run `npx skills add brycewang-stanford/Awesome-Journal-Skills --skill arecon-evidence-standards -a claude-code`. Or copy the skill folder (Annual-Review-of-Economics-Skills/skills/arecon-evidence-standards in brycewang-stanford/Awesome-Journal-Skills) into .claude/skills/arecon-evidence-standards in your project. Claude Code loads it when a task matches its description.

How do I install Arecon Evidence Standards in Codex?

Run `npx skills add brycewang-stanford/Awesome-Journal-Skills --skill arecon-evidence-standards -a codex`. Or copy the skill folder (Annual-Review-of-Economics-Skills/skills/arecon-evidence-standards in brycewang-stanford/Awesome-Journal-Skills) into .agents/skills/arecon-evidence-standards in your project. Codex loads it when a task matches its description.

Can I use Arecon Evidence Standards in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add brycewang-stanford/Awesome-Journal-Skills --skill arecon-evidence-standards -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arecon-evidence-standards, .gemini/skills/arecon-evidence-standards, .github/skills/arecon-evidence-standards and .opencode/skills/arecon-evidence-standards in your project.

What does Arecon Evidence Standards need to run?

SKILL.md names no scripts, command-line tools or credentials: Arecon Evidence Standards is instructions for the agent only.

Does Arecon Evidence Standards access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Arecon Evidence Standards safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Arecon Evidence Standards use?

Arecon Evidence Standards is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Arecon Evidence Standards use?

About 1.7k tokens (SKILL.md is roughly 6.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Arecon Evidence Standards?

Skills that share tags, products or a category with Arecon Evidence Standards: Hypothesis Generation (spacering-net/codeg, 3.9k stars), GitHub Deep Research (bytedance/deer-flow, 84k stars), Nature Paper Card (Yuan1z0825/nature-skills, 47k stars) and Content Research Writer (weapp-tailwindcss/weapp-tailwindcss, 1.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Arecon Evidence Standards?

brycewang-stanford (a GitHub user) maintains it in brycewang-stanford/Awesome-Journal-Skills, which has 1,231 GitHub stars. The repository holds 2,387 skills in this directory. The repository was last updated on September 27, 2026.

Source: brycewang-stanford/Awesome-Journal-Skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.