Agent skill

Paper Autoraters

by Ar9av in Ar9av/PaperOrchestra

Run the four paper-quality autoraters from PaperOrchestra (arXiv:2604.05018, App.

Custom licenceAuto-check passedResearch & Science

Install Paper Autoraters

skills CLI
$ npx skills add Ar9av/PaperOrchestra --skill paper-autoraters -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Ar9av/PaperOrchestra paper-autoraters --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Ar9av/PaperOrchestra.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/paper-autoraters .claude/skills/paper-autoraters && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
paper-autoraters
GitHub stars
679
Used in
1 other repo
Token cost
~1.6k tokens
SKILL.md length
670 words
Files
6 (incl. scripts, references)
Skills in repo
9
Repo updated
First seen
Licence
Custom licence

At a glance

Run the four paper-quality autoraters from PaperOrchestra (arXiv:2604.05018, App.

  • Works in 2 steps: Partition the reference lists into P0 / P1 → Resolve references to entity IDs and…
  • The user asks to score this paper draft
  • SKILL.md covers The four autoraters, Workflow and Resources
  • Runs Python scripts from its folder; calls python

What it does

Paper Autoraters is an agent skill from Ar9av/PaperOrchestra. Run the four paper-quality autoraters from PaperOrchestra (arXiv:2604.05018, App. F.3) — Citation F1 (P0/P1 partition + Precision/Recall/F1), Literature Review Quality (6-axis 0-100 with anti-inflation rules), SxS Overall Paper Quality (side-by-side), and SxS Literature Review Quality (side-by-side). TRIGGER when the user asks to "score this paper draft", "evaluate against the benchmark", "compare two papers", or "run the autoraters".

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts and reference files (for example `references/citation-f1-prompt.md`, `references/litreview-quality-prompt.md` and `references/sxs-litreview-prompt.md`).

It sits in Research & Science, covering Literature review, Academic paper search and Citation management. It works with arXiv and Semantic Scholar. The repository describes itself as: An automated AI research-paper writer based off Google's PaperOrchestra paper's implementation through a skills - benchmark + autoraters using any coding agent (Claude Code…

When your agent uses it

  • The user asks to score this paper draft
  • Evaluate against the benchmark
  • Compare two papers
  • Run the autoraters

Example prompts

  • “score this paper draft”
  • “evaluate against the benchmark”
  • “compare two papers”
  • “/paper-autoraters”

Requirements

  • Python 3

Workflow steps

2 steps, taken from the step headings in SKILL.md.

  1. Partition the reference lists into P0 / P1
  2. Resolve references to entity IDs and compute F1

What it can do on your machine

Read from SKILL.md and the folder at commit 36c3cc4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Paper Autoraters loads about 1.6k tokens when it runs, and up to ~5.8k if it reads all its reference files. Until then it costs about 114 tokens; SKILL.md has 670 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~114
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~5.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 670 words (~1,637 tokens).

“Faithful implementation of the four LLM-as-judge autoraters used in PaperOrchestra (Song et al., 2026, arXiv:2604.05018, §5 and App. F.3).”

— opening of SKILL.md by Ar9av, Custom licence
name
paper-autoraters

Read the full SKILL.md on GitHub

Files

SKILL.md and 5 other files (scripts, references) in skills/paper-autoraters of Ar9av/PaperOrchestra.

  • SKILL.md
  • references/citation-f1-prompt.md
  • references/litreview-quality-prompt.md
  • references/sxs-litreview-prompt.md
  • references/sxs-paper-quality-prompt.md
  • scripts/compute_f1.py

Open the folder on GitHubat commit 36c3cc4

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Ar9av/PaperOrchestra, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Paper Autoraters next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Paper Autoraters compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Paper Autoraters this skillAr9av/PaperOrchestra6791 repos~1.6kAutomated safety check: PassCustom licence
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Paper Research on arXivXiaomiMiMo/MiMo-Code14k—~1.5kAutomated safety check: PassMIT
Deep Research Literature SurveyHKUSTDial/Supervisor-Skills8.8k—~2.4kAutomated safety check: PassCC-BY-NC-SA-4.0
Literature ReviewK-Dense-AI/scientific-agent-skills48k1 repos~3.2kAutomated safety check: NotesMIT
Literature ReviewNorman-bury/research-writing-skill3.4k—~2.2kAutomated safety check: NotesMIT

Similar skills

  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Paper Research on arXiv

    XiaomiMiMo/MiMo-Code

    Searches arXiv, fetches metadata, generates BibTeX, downloads PDFs and finds citations and related papers using a bundled Python script.

    14k GitHub stars~1.5k tokensUpdated yesterday
    Research & ScienceAuto-check passed
  • Deep Research Literature Survey

    HKUSTDial/Supervisor-Skills

    Runs a survey-grade literature investigation: fixes the research questions, searches from adversarial angles, verifies citations and writes an evidence-first report.

    8.8k GitHub stars~2.4k tokensUpdated 1 mo ago
    Research & ScienceAuto-check passed
  • Literature Review

    K-Dense-AI/scientific-agent-skills

    Runs systematic, scoping or narrative literature reviews across PubMed, arXiv, bioRxiv and Semantic Scholar, with citation checks and Markdown or PDF output.

    48k GitHub starsUsed in 1 repo~3.2k tokens
    Research & ScienceAuto-check: notes
  • Literature Review

    Norman-bury/research-writing-skill

    A skill your agent uses when writing literature review sections - guides searching, organizing, and synthesizing academic sources

    3.4k GitHub stars~2.2k tokensUpdated 4 mo ago
    Research & ScienceAuto-check: notes
  • Paper Navigator

    AI4Scientist/nano-scientist

    Find and read academic papers: disambiguate queries, discover papers (search, citation traversal, recommendations, arXiv monitoring, trending, GitHub search), evaluate (TLDR, citations, code, SOTA)…

    128 GitHub stars~7.7k tokensUpdated 4 mo ago
    Research & ScienceAuto-check: notes

More from Ar9av/PaperOrchestra

All 9 skills in this repo
  • Agent Research Aggregator

    Ar9av/PaperOrchestra

    Pre-pipeline aggregator that scans AI agent cache directories (.claude, .cursor, .antigravity, .openclaw) or any user-specified directory for experimentation logs, extracts insights and numeric…

    679 GitHub starsUsed in 1 repo~3.5k tokens
    Auto-check passed
  • Literature Review Agent

    Ar9av/PaperOrchestra

    Step 3 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~5.2k tokens
    Auto-check passed
  • Outline Agent

    Ar9av/PaperOrchestra

    Step 1 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~1.6k tokens
    Auto-check passed
  • Paper Orchestra

    Ar9av/PaperOrchestra

    Orchestrate the full PaperOrchestra (Song et al., 2026, arXiv:2604.05018) five-agent pipeline to turn unstructured research materials (idea, experimental log, LaTeX template, conference guidelines…

    679 GitHub starsUsed in 1 repo~3.5k tokens
    Auto-check passed
  • Plotting Agent

    Ar9av/PaperOrchestra

    Step 2 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~2.3k tokens
    Auto-check passed
  • Section Writing Agent

    Ar9av/PaperOrchestra

    Step 4 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.

    679 GitHub starsUsed in 1 repo~3.3k tokens
    Auto-check passed

Questions about Paper Autoraters

What does Paper Autoraters do?

Run the four paper-quality autoraters from PaperOrchestra (arXiv:2604.05018, App. Paper Autoraters is an agent skill from Ar9av/PaperOrchestra.05018, App.

When should I use Paper Autoraters?

Paper Autoraters fits situations like: the user asks to score this paper draft; evaluate against the benchmark; compare two papers; run the autoraters.

How do I install Paper Autoraters in Claude Code?

Run `npx skills add Ar9av/PaperOrchestra --skill paper-autoraters -a claude-code`. Or copy the skill folder (skills/paper-autoraters in Ar9av/PaperOrchestra) into .claude/skills/paper-autoraters in your project. Claude Code loads it when a task matches its description.

How do I install Paper Autoraters in Codex?

Run `npx skills add Ar9av/PaperOrchestra --skill paper-autoraters -a codex`. Or copy the skill folder (skills/paper-autoraters in Ar9av/PaperOrchestra) into .agents/skills/paper-autoraters in your project. Codex loads it when a task matches its description.

Can I use Paper Autoraters in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Ar9av/PaperOrchestra --skill paper-autoraters -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/paper-autoraters, .gemini/skills/paper-autoraters, .github/skills/paper-autoraters and .opencode/skills/paper-autoraters in your project.

What does Paper Autoraters need to run?

Going by SKILL.md and its folder, Paper Autoraters needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Paper Autoraters access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Paper Autoraters safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Paper Autoraters use?

Paper Autoraters has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Paper Autoraters use?

About 1.6k tokens (SKILL.md is roughly 6.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4.1k tokens, read only when the agent opens those files.

What are the alternatives to Paper Autoraters?

Skills that share tags, products or a category with Paper Autoraters: Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Paper Research on arXiv (XiaomiMiMo/MiMo-Code, 14k stars), Deep Research Literature Survey (HKUSTDial/Supervisor-Skills, 8.8k stars) and Literature Review (K-Dense-AI/scientific-agent-skills, 48k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Paper Autoraters?

Ar9av (a GitHub user) maintains it in Ar9av/PaperOrchestra, which has 679 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on September 21, 2026.

Source: Ar9av/PaperOrchestra on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.