Paper Orchestra
Ar9av/PaperOrchestra
Orchestrate the full PaperOrchestra (Song et al., 2026, arXiv:2604.05018) five-agent pipeline to turn unstructured research materials (idea, experimental log, LaTeX template, conference guidelines…
Download and parse LaTeX source files from arXiv preprints. An agent skill from wentorai/research-plugins.
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install wentorai/research-plugins arxiv-latex-source --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .claude/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .claude/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-sourceType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install wentorai/research-plugins arxiv-latex-source --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .agents/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .agents/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install wentorai/research-plugins arxiv-latex-source --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .cursor/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .cursor/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/wentorai/research-plugins.git --path skills/literature/fulltext/arxiv-latex-source--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install wentorai/research-plugins arxiv-latex-source --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .gemini/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .gemini/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install wentorai/research-plugins arxiv-latex-sourceInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .github/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .github/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add wentorai/research-plugins --skill arxiv-latex-source -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install wentorai/research-plugins arxiv-latex-source --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/wentorai/research-plugins.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/literature/fulltext/arxiv-latex-source .opencode/skills/arxiv-latex-source && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "arxiv-latex-source" agent skill from https://github.com/wentorai/research-plugins/tree/main/skills/literature/fulltext/arxiv-latex-source into .opencode/skills/arxiv-latex-source/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "arxiv-latex-source", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
arxiv-latex-sourceDownload and parse LaTeX source files from arXiv preprints. An agent skill from wentorai/research-plugins.
Arxiv Latex Source is an agent skill from wentorai/research-plugins. Download and parse LaTeX source files from arXiv preprints
Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Research & Science, covering Academic paper search and LaTeX. It works with LaTeX and arXiv. The repository describes itself as: 350+ academic research skills, MCP configs, and plugins for Research-Claw and AI agents. The licence is MIT.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit bf44b3c. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
curlFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
arxiv.orgexport.arxiv.orgAlso links to:
info.arxiv.orgFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Arxiv Latex Source loads about 2k tokens when it runs. Until then it costs about 19 tokens; SKILL.md has 487 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from wentorai/research-plugins at commit bf44b3c, republished under its MIT licence (© wentorai). 487 words, ~1,995 tokens.
.claude/skills/arxiv-latex-source/SKILL.md (or your agent's skills folder).arXiv stores the original LaTeX source files for the vast majority of its 2.4 million+ preprints. Accessing LaTeX source provides major advantages over PDF parsing: exact mathematical notation as written by the author, structured sections and labels, machine-readable bibliography entries, and intact figure captions, table data, and cross-references.
For formula extraction, citation graph construction, section-level text analysis, or training data curation for scientific language models, LaTeX source is the gold standard. PDF parsing introduces OCR errors in equations, loses structural hierarchy, and mangles complex tables.
The e-print endpoint serves source bundles as gzip-compressed tarballs (.tar.gz) containing .tex files, figures, .bib/.bbl bibliography files, style files, and supplementary materials. No authentication is required.
No authentication or API key is required. The e-print endpoint is publicly accessible. However, arXiv asks that automated tools set a descriptive User-Agent header and comply with rate limits.
URL: GET https://arxiv.org/e-print/{arxiv_id}
Response: application/gzip — a .tar.gz archive containing the source files
Parameters:
| Param | Type | Required | Description |
|---|---|---|---|
| arxiv_id | string | Yes | arXiv identifier, e.g. 2301.00001 or 2301.00001v2 for a specific version |
Example:
# Download source archive (response: 200, application/gzip, ~1.3 MB)
curl -sL -o source.tar.gz "https://arxiv.org/e-print/2301.00001"
# List archive contents
tar tz -f source.tar.gz | head -10
# ACM-Reference-Format.bbx
# ACM-Reference-Format.bst
# Image_1.jpg
# README.txt
# acmart.clsContent-Disposition header: attachment; filename="arXiv-2301.00001v1.tar.gz"
ETag: SHA-256 hash provided for caching: sha256:f1ffe8ec...
The endpoint almost always returns a gzip-compressed tar archive. Rare cases (very old or single-file submissions) may return a single gzip-compressed .tex file without tar wrapper. Always verify format before extracting:
curl -sL "https://arxiv.org/e-print/{arxiv_id}" -o source.gz
file source.gz # "gzip compressed data, was 'XXXX.tar', ..."Pair source downloads with the arXiv Atom API for structured metadata:
GET https://export.arxiv.org/api/query?id_list={arxiv_id}<title>, <author>, <summary>, <category>, <published>curl -s "https://export.arxiv.org/api/query?id_list=2301.00001"A source archive typically contains multiple files. To find the main document:
\documentclass in .tex files — this marks the root documentREADME.txt that may specify the main file.tex files contain \documentclass, prefer the one with \begin{document}import tarfile, re
def find_main_tex(tar_path):
with tarfile.open(tar_path, 'r:gz') as tar:
tex_files = [m for m in tar.getmembers() if m.name.endswith('.tex')]
for member in tex_files:
content = tar.extractfile(member).read().decode('utf-8', errors='ignore')
if r'\documentclass' in content and r'\begin{document}' in content:
return member.name, content
return None, NoneLaTeX sections follow a predictable hierarchy:
import re
def extract_sections(tex_content):
pattern = r'\\(section|subsection|subsubsection)\{([^}]+)\}'
sections = re.findall(pattern, tex_content)
return [(level, title) for level, title in sections]
# [('section', 'Introduction'), ('section', 'Related Work'), ...]def extract_equations(tex_content):
patterns = [
r'\\\[(.+?)\\\]',
r'\\begin\{equation\}(.+?)\\end\{equation\}',
r'\\begin\{align\*?\}(.+?)\\end\{align\*?\}',
]
equations = []
for pat in patterns:
equations.extend(re.findall(pat, tex_content, re.DOTALL))
return equationsParse .bib files (BibTeX entries) or .bbl files (compiled \bibitem commands):
def extract_bibliography(tar_path):
refs = []
with tarfile.open(tar_path, 'r:gz') as tar:
for member in tar.getmembers():
if member.name.endswith('.bib'):
content = tar.extractfile(member).read().decode('utf-8', errors='ignore')
refs.extend(re.findall(r'@\w+\{([^,]+),(.+?)\n\}', content, re.DOTALL))
elif member.name.endswith('.bbl'):
content = tar.extractfile(member).read().decode('utf-8', errors='ignore')
refs.extend(re.findall(r'\\bibitem.*?\{(.+?)\}', content))
return refsMyTool/1.0 (mailto:user@university.edu).bib/.bbl files for exact reference keys to construct citation graphsimport requests, tarfile, io, re, time, gzip
def download_arxiv_source(arxiv_id, delay=1.0):
"""Download and extract all .tex files from an arXiv paper's source."""
url = f"https://arxiv.org/e-print/{arxiv_id}"
headers = {"User-Agent": "ResearchTool/1.0 (mailto:user@example.com)"}
resp = requests.get(url, headers=headers)
resp.raise_for_status()
time.sleep(delay)
buf = io.BytesIO(resp.content)
try:
with tarfile.open(fileobj=buf, mode='r:gz') as tar:
return {m.name: tar.extractfile(m).read().decode('utf-8', errors='ignore')
for m in tar.getmembers() if m.name.endswith('.tex') and m.isfile()}
except tarfile.ReadError:
buf.seek(0)
return {"main.tex": gzip.decompress(buf.read()).decode('utf-8', errors='ignore')}
# Usage
sources = download_arxiv_source("2301.00001")
for fname, content in sources.items():
if r'\documentclass' in content:
sections = re.findall(r'\\section\{([^}]+)\}', content)
equations = re.findall(r'\\begin\{equation\}(.+?)\\end\{equation\}', content, re.DOTALL)
print(f"{fname}: {len(sections)} sections, {len(equations)} equations")© wentorai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/literature/fulltext/arxiv-latex-source of wentorai/research-plugins.
Open the folder on GitHubat commit bf44b3c
We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in wentorai/research-plugins, which our catalogue first saw on October 7, 2026.
Arxiv Latex Source next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Arxiv Latex Source this skillwentorai/research-plugins | 298 | 1 repos | ~2k | Automated safety check: Pass | MIT | |
| Paper OrchestraAr9av/PaperOrchestra | 679 | 1 repos | ~3.5k | Automated safety check: Pass | Custom licence | |
| Section Writing AgentAr9av/PaperOrchestra | 679 | 1 repos | ~3.3k | Automated safety check: Pass | Custom licence | |
| SearchMuuuun/luxas | 1.2k | — | ~2.3k | Automated safety check: Warn | MIT | |
| Oleafly Pre SubmissionOleafly/Oleafly | 212 | — | ~2.9k | Automated safety check: Pass | MIT | |
| Anmath Submissionfranklee16/academic-research-skills | 223 | 1 repos | ~1.1k | Automated safety check: Pass | None |
Ar9av/PaperOrchestra
Orchestrate the full PaperOrchestra (Song et al., 2026, arXiv:2604.05018) five-agent pipeline to turn unstructured research materials (idea, experimental log, LaTeX template, conference guidelines…
Ar9av/PaperOrchestra
Step 4 of the PaperOrchestra pipeline (arXiv:2604.05018). An agent skill from Ar9av/PaperOrchestra.
Muuuun/luxas
Unified academic paper search, citation chains, paper download (arXiv LaTeX/PDF, Sci-Hub), figure extraction from papers, LaTeX source reading, BibTeX fetching, web search, and browser automation…
Oleafly/Oleafly
Run a pre-flight pass over the project before uploading to arXiv or a venue, and write a pass or fail checklist.
franklee16/academic-research-skills
A skill your agent uses when running the final pre-submission preflight for an Annals of Mathematics manuscript — AMS-LaTeX compile, theorem environments, abstract and MSC, references, arXiv…
Mathews-Tom/armory
Package a TeX/LaTeX project into a clean tarball or zip for arXiv upload: file selection, build-artifact exclusion, 00README.XXX generation, ancillary file organization, archive validation.
wentorai/research-plugins
Craft structured research abstracts that maximize clarity and journal acceptance
wentorai/research-plugins
Manage academic citations across BibTeX, APA, MLA, and Chicago formats
wentorai/research-plugins
Summarize academic papers with structured extraction of key elements
wentorai/research-plugins
Evidence-based study techniques for academic learning and retention
wentorai/research-plugins
Adjust writing tone and register for academic audiences and venues
wentorai/research-plugins
Academic translation, post-editing, and Chinglish correction guide
Categories
Download and parse LaTeX source files from arXiv preprints. An agent skill from wentorai/research-plugins. Arxiv Latex Source is an agent skill from wentorai/research-plugins.
Arxiv Latex Source fits situations like: tasks that involve Academic paper search; tasks that involve LaTeX.
Run `npx skills add wentorai/research-plugins --skill arxiv-latex-source -a claude-code`. Or copy the skill folder (skills/literature/fulltext/arxiv-latex-source in wentorai/research-plugins) into .claude/skills/arxiv-latex-source in your project. Claude Code loads it when a task matches its description.
Run `npx skills add wentorai/research-plugins --skill arxiv-latex-source -a codex`. Or copy the skill folder (skills/literature/fulltext/arxiv-latex-source in wentorai/research-plugins) into .agents/skills/arxiv-latex-source in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add wentorai/research-plugins --skill arxiv-latex-source -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/arxiv-latex-source, .gemini/skills/arxiv-latex-source, .github/skills/arxiv-latex-source and .opencode/skills/arxiv-latex-source in your project.
Going by SKILL.md and its folder, Arxiv Latex Source needs the command-line tools its instructions call (curl). Our summary lists: Python 3.
SKILL.md names 3 domains. In commands or code: arxiv.org and export.arxiv.org; the agent is likely to contact these when it follows the instructions. As links in the text: info.arxiv.org. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Arxiv Latex Source is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2k tokens (SKILL.md is roughly 8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Arxiv Latex Source: Paper Orchestra (Ar9av/PaperOrchestra, 679 stars), Section Writing Agent (Ar9av/PaperOrchestra, 679 stars), Search (Muuuun/luxas, 1.2k stars) and Oleafly Pre Submission (Oleafly/Oleafly, 212 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
wentorai (a GitHub user) maintains it in wentorai/research-plugins, which has 298 GitHub stars. The repository holds 405 skills in this directory. The repository was last updated on June 19, 2026.
Source: wentorai/research-plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.