Agent skill

Arxiv Paper Reader

by huangruiteng in huangruiteng/CS-Notes

Fetch and extract full text from arxiv paper HTML pages. An agent skill from huangruiteng/CS-Notes.

MITAuto-check passedResearch & Science

Install Arxiv Paper Reader

skills CLI
$ npx skills add huangruiteng/CS-Notes --skill arxiv-paper-reader -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install huangruiteng/CS-Notes arxiv-paper-reader --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/huangruiteng/CS-Notes.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.trae/openclaw-skills/arxiv-paper-reader .claude/skills/arxiv-paper-reader && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
arxiv-paper-reader
GitHub stars
4k
Token cost
~393 tokens
SKILL.md length
82 words
Files
2 (incl. scripts)
Skills in repo
39
Repo updated
First seen
Licence
MIT

At a glance

Fetch and extract full text from arxiv paper HTML pages. An agent skill from huangruiteng/CS-Notes.

  • Works in 2 steps: abs 页面:https://arxiv.org/abs/ — 只有摘要 → HTML 页面:https://arxiv.org/html/v1 —…
  • Asks to read an arxiv paper
  • SKILL.md covers 适用场景, 核心原理, 使用步骤 and 示例, plus 2 more sections
  • Runs Python scripts from its folder; calls python; reaches arxiv.org

What it does

Arxiv Paper Reader is an agent skill from huangruiteng/CS-Notes. Fetch and extract full text from arxiv paper HTML pages. Invoke when user asks to read an arxiv paper, analyze a paper's content, or when given an arxiv ID/URL.

Its SKILL.md is about 390 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/arxiv_reader.py`).

It sits in Research & Science, covering Academic paper search. It works with arXiv. The licence is MIT.

When your agent uses it

  • Asks to read an arxiv paper
  • Analyze a papers content
  • Given an arxiv ID/URL

Example prompts

  • “/arxiv-paper-reader”

Requirements

  • Python 3

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. abs 页面:https://arxiv.org/abs/ — 只有摘要
  2. HTML 页面:https://arxiv.org/html/v1 — 完整论文文本(LaTeXML 渲染)

What it can do on your machine

Read from SKILL.md and the folder at commit f7b4e92. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • arxiv.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Arxiv Paper Reader loads about 393 tokens when it runs. Until then it costs about 45 tokens; SKILL.md has 82 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~45
When it runs · the whole SKILL.md, loaded when a task matches
~393

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from huangruiteng/CS-Notes at commit f7b4e92, republished under its MIT licence (© huangruiteng). 82 words, ~393 tokens.

Download SKILL.mdSave it as .claude/skills/arxiv-paper-reader/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
arxiv-paper-reader
description
Fetch and extract full text from arxiv paper HTML pages. Invoke when user asks to read an arxiv paper, analyze a paper's content, or when given an arxiv ID/URL.

Arxiv Paper Reader

从 arxiv 论文的 HTML 页面抓取并提取全文纯文本。arxiv 的 HTML 版本比 PDF 更容易程序化提取,且包含完整内容。

适用场景

  • 用户给出 arxiv 论文 ID 或 URL,要求阅读/分析论文内容
  • 需要获取论文的完整文本用于笔记整理、技术分析
  • WebFetch 工具无法直接访问 arxiv 页面时的替代方案

核心原理

arxiv 论文有两种可读格式:

  1. abs 页面:https://arxiv.org/abs/<ID> — 只有摘要
  2. HTML 页面:https://arxiv.org/html/<ID>v1 — 完整论文文本(LaTeXML 渲染)

HTML 页面是完整论文的网页版,可通过 curl 获取后提取纯文本。这比 PDF 解析更可靠。

使用步骤

  1. 从用户输入中提取 arxiv ID(如 2507.02259)
  2. 运行脚本:
bash
python .trae/skills/arxiv-paper-reader/scripts/arxiv_reader.py <arxiv_id_or_url>
  1. 可选参数:
    • --output FILE:保存到文件(默认输出到 stdout)
    • --sections:添加章节分隔符
    • --raw:跳过数学符号清理

示例

bash
# 基本用法
python .trae/skills/arxiv-paper-reader/scripts/arxiv_reader.py 2507.02259

# 从 URL 提取
python .trae/skills/arxiv-paper-reader/scripts/arxiv_reader.py https://arxiv.org/abs/2507.02259

# 保存到文件并添加章节分隔
python .trae/skills/arxiv-paper-reader/scripts/arxiv_reader.py 2507.02259 --output /tmp/paper.txt --sections

# 读取后用 Read 工具分段查看
python .trae/skills/arxiv-paper-reader/scripts/arxiv_reader.py 2507.02259 --output /tmp/paper.txt
# 然后 Read /tmp/paper.txt

输出说明

  • 纯文本格式,保留论文的章节结构
  • 数学公式会被简化清理(LaTeXML 渲染的 LaTeX 标记会被部分还原)
  • 包含标题、摘要、正文、参考文献等完整内容

注意事项

  • 依赖 curl 命令(macOS/Linux 默认可用)
  • 部分论文可能没有 HTML 版本(较老的论文),此时脚本会报错
  • 版本号默认使用 v1,如需其他版本可手动构造 URL
  • 对于超长论文,建议 --output 保存到文件后用 Read 工具分段查看

© huangruiteng, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in .trae/openclaw-skills/arxiv-paper-reader of huangruiteng/CS-Notes.

  • SKILL.md
  • scripts/arxiv_reader.py

Open the folder on GitHubat commit f7b4e92

Compare with similar skills

Arxiv Paper Reader next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Arxiv Paper Reader compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Arxiv Paper Reader this skillhuangruiteng/CS-Notes4k—~393Automated safety check: PassMIT
Read arXiv Paperkarpathy/nanochat59k1 repos~494Automated safety check: PassMIT
Literature Reviewneflibata-feng/MyArxiv-Agent12620 repos~5.9kAutomated safety check: NotesMIT
Openalex Databaseneflibata-feng/MyArxiv-Agent12612 repos~3kAutomated safety check: PassCustom licence
Citation ManagementK-Dense-AI/claude-scientific-writer2.4k2 repos~3.9kAutomated safety check: NotesMIT
Citation Managementneflibata-feng/MyArxiv-Agent12619 repos~8.1kAutomated safety check: NotesMIT

Similar skills

  • Read arXiv Paper

    karpathy/nanochat

    Fetches the TeX source of an arXiv paper from its URL, reads it and writes a markdown summary tied to the nanochat project.

    59k GitHub starsUsed in 1 repo~494 tokens
    Research & ScienceAuto-check passed
  • Literature Review

    neflibata-feng/MyArxiv-Agent

    Conduct comprehensive, systematic literature reviews using multiple academic databases (PubMed, arXiv, bioRxiv, Semantic Scholar, etc.).

    126 GitHub starsUsed in 20 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Openalex Database

    neflibata-feng/MyArxiv-Agent

    Query and analyze scholarly literature using the OpenAlex database.

    126 GitHub starsUsed in 12 repos~3k tokens
    Research & ScienceAuto-check passed
  • Citation Management

    K-Dense-AI/claude-scientific-writer

    Finds papers in OpenAlex, PubMed and Google Scholar, turns DOIs, PMIDs and arXiv IDs into clean BibTeX, and validates citations for a manuscript or thesis.

    2.4k GitHub starsUsed in 2 repos~3.9k tokens
    Research & ScienceAuto-check: notes
  • Citation Management

    neflibata-feng/MyArxiv-Agent

    Comprehensive citation management for academic research. An agent skill from neflibata-feng/MyArxiv-Agent.

    126 GitHub starsUsed in 19 repos~8.1k tokens
    Research & ScienceAuto-check: notes
  • Searches arXiv across many papers on one topic, extracts each paper's methodology and findings in parallel, and synthesizes a cited literature review.

    84k GitHub starsUsed in 2 repos~4.3k tokens
    Research & ScienceAuto-check passed

More from huangruiteng/CS-Notes

All 39 skills in this repo
  • CLI Creator

    huangruiteng/CS-Notes

    Build a composable CLI for Codex from API docs, an OpenAPI spec, existing curl examples, an SDK, a web app, an admin tool, or a local script.

    4k GitHub starsUsed in 2 repos~2.7k tokens
    Auto-check passed
  • Codex Thread Heartbeat

    huangruiteng/CS-Notes

    Inspect and manage guarded Codex App-native or launchd heartbeats for Codex main control threads.

    4k GitHub stars~1.5k tokensUpdated 2 days ago
    Auto-check passed
  • Slack

    huangruiteng/CS-Notes

    A skill your agent uses when you need to control Slack from Clawdbot via the slack tool, including reacting to messages or pinning/unpinning items in Slack channels or DMs.

    4k GitHub starsUsed in 10 repos~578 tokens
    Auto-check passed
  • Codex Thread Reader

    huangruiteng/CS-Notes

    Locate and read a Codex thread by a codex thread link, thread id, or rollout path across all local CODEXHOME directories (~/.codex, ~/.codex-gpt, ...).

    4k GitHub stars~1.4k tokensUpdated 2 days ago
    Auto-check passed
  • Research Material Scout

    huangruiteng/CS-Notes

    A skill your agent uses when the user asks Codex to research, find learning materials, process "素材:" links, "请你读" / "精读" a material, build a material radar, or use SenSight-like broad information…

    4k GitHub stars~8.3k tokensUpdated 2 days ago
    Auto-check passed
  • GitHub

    huangruiteng/CS-Notes

    Interact with GitHub using the gh CLI. An agent skill from huangruiteng/CS-Notes.

    4k GitHub starsUsed in 27 repos~279 tokens
    Auto-check passed

Works with

Questions about Arxiv Paper Reader

What does Arxiv Paper Reader do?

Fetch and extract full text from arxiv paper HTML pages. An agent skill from huangruiteng/CS-Notes. Arxiv Paper Reader is an agent skill from huangruiteng/CS-Notes. Fetch and extract full text from arxiv paper HTML pages.

When should I use Arxiv Paper Reader?

Arxiv Paper Reader fits situations like: asks to read an arxiv paper; analyze a papers content; given an arxiv ID/URL.

How do I install Arxiv Paper Reader in Claude Code?

Run `npx skills add huangruiteng/CS-Notes --skill arxiv-paper-reader -a claude-code`. Or copy the skill folder (.trae/openclaw-skills/arxiv-paper-reader in huangruiteng/CS-Notes) into .claude/skills/arxiv-paper-reader in your project. Claude Code loads it when a task matches its description.

How do I install Arxiv Paper Reader in Codex?

Run `npx skills add huangruiteng/CS-Notes --skill arxiv-paper-reader -a codex`. Or copy the skill folder (.trae/openclaw-skills/arxiv-paper-reader in huangruiteng/CS-Notes) into .agents/skills/arxiv-paper-reader in your project. Codex loads it when a task matches its description.

Can I use Arxiv Paper Reader in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add huangruiteng/CS-Notes --skill arxiv-paper-reader -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arxiv-paper-reader, .gemini/skills/arxiv-paper-reader, .github/skills/arxiv-paper-reader and .opencode/skills/arxiv-paper-reader in your project.

What does Arxiv Paper Reader need to run?

Going by SKILL.md and its folder, Arxiv Paper Reader needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python 3.

Does Arxiv Paper Reader access the network?

SKILL.md names 1 domain. In commands or code: arxiv.org; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Arxiv Paper Reader safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Arxiv Paper Reader use?

Arxiv Paper Reader is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Arxiv Paper Reader use?

About 393 tokens (SKILL.md is roughly 1.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Arxiv Paper Reader?

Skills that share tags, products or a category with Arxiv Paper Reader: Read arXiv Paper (karpathy/nanochat, 59k stars), Literature Review (neflibata-feng/MyArxiv-Agent, 126 stars), Openalex Database (neflibata-feng/MyArxiv-Agent, 126 stars) and Citation Management (K-Dense-AI/claude-scientific-writer, 2.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Arxiv Paper Reader?

huangruiteng (a GitHub user) maintains it in huangruiteng/CS-Notes, which has 4,001 GitHub stars. The repository holds 39 skills in this directory. The repository was last updated on October 8, 2026.

Source: huangruiteng/CS-Notes on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.