Agent skill

Import

by ZimoLiao in ZimoLiao/scholaraio

A skill your agent uses when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content.

MITAuto-check passedDocuments & Office

Install Import

skills CLI
$ npx skills add ZimoLiao/scholaraio --skill import -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ZimoLiao/scholaraio import --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ZimoLiao/scholaraio.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/import .claude/skills/import && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
import
GitHub stars
577
Token cost
~790 tokens
SKILL.md length
151 words
Files
1
Skills in repo
43
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content.

  • Works in 4 steps: 批量 PDF→MD:云端模式使用… → Abstract 补全:从 markdown 中提取缺失的摘要 → TOC + L3 提取:LLM 提取目录结构和结论段 → …
  • The user wants to import papers from Endnote XML/RIS
  • SKILL.md covers Endnote 导入, Zotero 导入, 补充 PDF(单篇) and 批量补转 PDF(已入库论文)
  • Needs COLLECTION_KEY

What it does

Import is an agent skill from ZimoLiao/scholaraio. Use when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content.

Its SKILL.md is about 790 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Documents & Office, covering PDF and Citation management. It works with Zotero and SQLite. The repository describes itself as: Scholar All-In-One: A research infrastructure for AI agents. The licence is MIT.

When your agent uses it

  • The user wants to import papers from Endnote XML/RIS
  • Match PDFs to records
  • Supplement records with PDF content

Example prompts

  • “/import”

Requirements

  • Python 3
  • A credential in COLLECTION_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. 批量 PDF→MD:云端模式使用 convert_pdfs_cloud_batch() 批量转换(批次大小由 config.yaml ingest.mineru_batch_size 控制,默认 20)
  2. Abstract 补全:从 markdown 中提取缺失的摘要
  3. TOC + L3 提取:LLM 提取目录结构和结论段
  4. Embed + Index:更新语义向量和全文索引

What it can do on your machine

Read from SKILL.md and the folder at commit 777628b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash, yaml and python).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • COLLECTION_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Import loads about 790 tokens when it runs. Until then it costs about 44 tokens; SKILL.md has 151 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~44
When it runs · the whole SKILL.md, loaded when a task matches
~790

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ZimoLiao/scholaraio at commit 777628b, republished under its MIT licence (© ZimoLiao). 151 words, ~790 tokens.

Download SKILL.mdSave it as .claude/skills/import/SKILL.md (or your agent's skills folder).
name
import
description
Use when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content.

导入外部文献管理工具数据 / 补充 PDF

支持从 Endnote / Zotero 批量导入,或为已入库论文补充/重拉 PDF(attach-pdf、fetch-pdf)。

Endnote 导入

支持 Endnote 导出的 XML 和 RIS 格式文件。

bash
# 完整导入:元数据 + PDF 匹配 + MinerU 批量转换 + enrich (toc/l3/abstract) + embed + index
scholaraio import-endnote <file.xml>

# 多文件导入
scholaraio import-endnote file1.xml file2.ris

# 仅导入元数据和 PDF,跳过 MinerU 转换和 enrich
scholaraio import-endnote <file.xml> --no-convert

# 预览模式
scholaraio import-endnote <file.xml> --dry-run

# 离线模式
scholaraio import-endnote <file.xml> --no-api
PDF 自动匹配

对 Endnote XML 文件,自动解析 internal-pdf:// 链接,从 <library>.Data/PDF/ 目录匹配 PDF:

  • 多个 PDF 时自动排除 SI/补充材料
  • 默认通过 MinerU 批量转换为 paper.md
导入后自动处理

默认行为(不带 --no-convert)下,导入完成后自动执行完整 pipeline:

  1. 批量 PDF→MD:云端模式使用 convert_pdfs_cloud_batch() 批量转换(批次大小由 config.yaml ingest.mineru_batch_size 控制,默认 20)
  2. Abstract 补全:从 markdown 中提取缺失的摘要
  3. TOC + L3 提取:LLM 提取目录结构和结论段
  4. Embed + Index:更新语义向量和全文索引

使用 --no-convert 跳过以上全部后处理(仅导入元数据 + PDF 复制 + embed + index)。

Zotero 导入

支持 Web API 和本地 SQLite 两种模式。

Web API 模式
bash
# 列出 collections
scholaraio import-zotero --api-key KEY --library-id ID --list-collections

# 完整导入
scholaraio import-zotero --api-key KEY --library-id ID

# 仅导入指定 collection
scholaraio import-zotero --api-key KEY --library-id ID --collection COLLECTION_KEY

# 导入后将 collections 创建为工作区
scholaraio import-zotero --api-key KEY --library-id ID --import-collections
本地 SQLite 模式
bash
scholaraio import-zotero --local /path/to/zotero.sqlite
配置文件(可选)

在 config.local.yaml 中配置 Zotero 凭据:

yaml
zotero:
  api_key: "your-zotero-api-key"
  library_id: "your-library-id"

补充 PDF(单篇)

已有本地 PDF 文件时:

bash
scholaraio attach-pdf <paper-id> <path/to/paper.pdf> [--force]

自动把原始 PDF 保存到论文目录中(与 paper.md 同级,使用论文目录同名 stem),调用 MinerU 转换 PDF → markdown,补全缺失的 abstract,增量更新 embed + index。若目标目录已有 canonical PDF,默认拒绝覆盖;确认要替换时使用 --force。

当前网络环境有正版访问权限、希望从出版社页面或 DOI 拉取 PDF 时:

bash
# 拉取新论文 PDF 到 configured inbox;校园网直连时加 --direct 可绕过代理环境变量
scholaraio fetch-pdf 10.xxxx/example --direct

# 拉取后立刻走普通论文入库流程
scholaraio fetch-pdf 10.xxxx/example --direct --ingest

# 为已入库单篇论文重新拉取 canonical PDF;已有 PDF 时需要 --force
scholaraio fetch-pdf --paper <paper-id> --direct --force

# 为指定多篇论文批量重新拉取 PDF
scholaraio fetch-pdf --paper <paper-id-1> <paper-id-2> --direct --force

# 批量为全库论文重新拉取 PDF
scholaraio fetch-pdf --all --direct --force

fetch-pdf 只做 PDF 获取,不做访问绕过;能否下载取决于用户当前网络、机构权限和 publisher 返回的 PDF 链接。它会优先使用 source_url,否则使用 DOI。重拉已有论文不会自动重新转换 paper.md。

批量补转 PDF(已入库论文)

对已入库但缺少 paper.md 的论文(如首次导入时用了 --no-convert),可通过 Python 调用批量转换:

python
from scholaraio.core.config import load_config
from scholaraio.services.ingest.pipeline import batch_convert_pdfs

cfg = load_config()
stats = batch_convert_pdfs(cfg, enrich=True)

自动扫描 configured papers library 中有 PDF 无 paper.md 的论文,云端模式使用批量 API 转换,完成后运行 abstract backfill + toc + l3 + embed + index。

© ZimoLiao, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .claude/skills/import of ZimoLiao/scholaraio.

Open the folder on GitHubat commit 777628b

Compare with similar skills

Import next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Import compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Import this skillZimoLiao/scholaraio577—~790Automated safety check: PassMIT
Docsagentdocsagent/docsagent625—~834Automated safety check: PassNone
Ref Downloaderltczding-gif/ref-downloader139—~5.9kAutomated safety check: PassMIT
Obsidian Paper VaultAperivue/medsci-skills333—~1.6kAutomated safety check: PassMIT
Research Vault Literature Retrievalcheneternity/Research-Vault-Literature-Retrieval292—~2.3kAutomated safety check: PassNone
Experiment Detail Comparatoraipoch/medical-research-skills1.9k—~3.3kAutomated safety check: NotesMIT

Similar skills

  • Docsagent

    docsagent/docsagent

    Search and manage private, local document collections (PDF, PPTX, DOCX) offline.

    625 GitHub stars~834 tokensUpdated 12 days ago
    Documents & OfficeAuto-check passed
  • Ref Downloader

    ltczding-gif/ref-downloader

    A skill your agent uses when the user asks to batch-download academic PDFs with ref-downloader — either ALL references of one paper (Mode A: DOI or PDF input), OR a custom batch of papers (Mode B…

    139 GitHub stars~5.9k tokensUpdated 4 mo ago
    Research & ScienceAuto-check passed
  • Obsidian Paper Vault

    Aperivue/medsci-skills

    A skill your agent uses when turning a folder of research PDFs into Obsidian notes, even if Obsidian is not named.

    333 GitHub stars~1.6k tokensUpdated 5 days ago
    Research & ScienceAuto-check passed
  • Research Vault Literature Retrieval

    cheneternity/Research-Vault-Literature-Retrieval

    ResearchVault 文献知识问题的默认检索技能。先从当前 Vault 的 Analytical Notes 定位相关论文,再按需要定向进入对应 Fulltext,必要时回到 Zotero PDF 验证。若用户明确要求基于当前项目文件检索,优先执行严格的项目文件检索后再回答。纯 Skill、代码、Git、文件整理和转换工具调试任务不自动触发文献检索。

    292 GitHub stars~2.3k tokensUpdated 10 days ago
    Research & ScienceAuto-check passed
  • Experiment Detail Comparator

    aipoch/medical-research-skills

    Compare experimental method details between two Zotero PDF papers, identify protocol differences (ratios, dosages, timing, conditions), search supporting literature to explain why they differ, and…

    1.9k GitHub stars~3.3k tokensUpdated 23 days ago
    Research & ScienceAuto-check: notes
  • Annotate Paper

    54yyyu/zotero-mcp

    Read the open paper and write study annotations into its PDF with zotero-cli - a context box on the title, a four-part summary on the abstract, role-coded abstract highlights, one box per figure…

    5.3k GitHub stars~1.5k tokensUpdated 2 days ago
    Research & ScienceAuto-check passed

More from ZimoLiao/scholaraio

All 43 skills in this repo
  • Document

    ZimoLiao/scholaraio

    A skill your agent uses when the user wants to create or inspect DOCX, PPTX, or XLSX files, generate a downloadable Office deliverable, or verify its structure and layout warnings with scholaraio…

    577 GitHub stars~674 tokensUpdated 15 days ago
    Auto-check passed
  • Academic Writing

    ZimoLiao/scholaraio

    A skill your agent uses when the user needs help choosing or organizing an academic-writing workflow by deliverable, stage, or format, including review articles, guided reading, paper sections, PPT…

    577 GitHub stars~827 tokensUpdated 15 days ago
    Auto-check passed
  • Arxiv

    ZimoLiao/scholaraio

    A skill your agent uses when the user wants to browse arXiv preprints, search arXiv directly, fetch a PDF by arXiv ID or URL, or send a preprint into the ScholarAIO ingest pipeline.

    577 GitHub stars~799 tokensUpdated 15 days ago
    Auto-check passed
  • Bioinformatics

    ZimoLiao/scholaraio

    A skill your agent uses when working on bioinformatics workflows such as alignment, variant calling, phylogenetics, or protein-structure analysis, especially across BLAST, minimap2, samtools…

    577 GitHub stars~1.4k tokensUpdated 15 days ago
    Auto-check passed
  • Citation Check

    ZimoLiao/scholaraio

    A skill your agent uses when the user wants to verify citations in AI-generated or human-written text against the local knowledge base and catch hallucinated, wrong, or missing references.

    577 GitHub stars~454 tokensUpdated 15 days ago
    Auto-check passed
  • Draw

    ZimoLiao/scholaraio

    A skill your agent uses when the user wants diagrams, flowcharts, architecture visuals, data relationships, timelines, concept maps, Mermaid, Graphviz, drawio, or polished paper figures generated…

    577 GitHub stars~1.3k tokensUpdated 15 days ago
    Auto-check: notes

Works with

Questions about Import

What does Import do?

A skill your agent uses when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content. Import is an agent skill from ZimoLiao/scholaraio. Use when the user wants to import papers from Endnote XML/RIS, Zotero Web API or local SQLite, attach PDFs, match PDFs to records, or supplement records with PDF content.

When should I use Import?

Import fits situations like: the user wants to import papers from Endnote XML/RIS; match PDFs to records; supplement records with PDF content.

How do I install Import in Claude Code?

Run `npx skills add ZimoLiao/scholaraio --skill import -a claude-code`. Or copy the skill folder (.claude/skills/import in ZimoLiao/scholaraio) into .claude/skills/import in your project. Claude Code loads it when a task matches its description.

How do I install Import in Codex?

Run `npx skills add ZimoLiao/scholaraio --skill import -a codex`. Or copy the skill folder (.claude/skills/import in ZimoLiao/scholaraio) into .agents/skills/import in your project. Codex loads it when a task matches its description.

Can I use Import in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ZimoLiao/scholaraio --skill import -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/import, .gemini/skills/import, .github/skills/import and .opencode/skills/import in your project.

What does Import need to run?

Going by SKILL.md and its folder, Import needs credentials named COLLECTION_KEY. Our summary lists: Python 3; A credential in COLLECTION_KEY.

Does Import access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Import safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Import use?

Import is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Import use?

About 790 tokens (SKILL.md is roughly 3.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Import?

Skills that share tags, products or a category with Import: Docsagent (docsagent/docsagent, 625 stars), Ref Downloader (ltczding-gif/ref-downloader, 139 stars), Obsidian Paper Vault (Aperivue/medsci-skills, 333 stars) and Research Vault Literature Retrieval (cheneternity/Research-Vault-Literature-Retrieval, 292 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Import?

ZimoLiao (a GitHub user) maintains it in ZimoLiao/scholaraio, which has 577 GitHub stars. The repository holds 43 skills in this directory. The repository was last updated on September 25, 2026.

Source: ZimoLiao/scholaraio on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.