Agent skill

Research Topic Extractor

by huangwb8 in huangwb8/ChineseResearchLaTeX

当用户明确要求"从文件/图片/网页/描述中提取综述主题"、"生成主题+关键词+核心问题结构化输出",或要求使用旧名 get-review-theme skill 时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 research-literature-review…

MITAuto-check passedDocuments & Office

Install Research Topic Extractor

skills CLI
$ npx skills add huangwb8/ChineseResearchLaTeX --skill research-topic-extractor -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install huangwb8/ChineseResearchLaTeX research-topic-extractor --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/huangwb8/ChineseResearchLaTeX.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/research-topic-extractor .claude/skills/research-topic-extractor && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
research-topic-extractor
GitHub stars
2.9k
Token cost
~666 tokens
SKILL.md length
179 words
Files
5 (incl. references)
Skills in repo
27
Repo updated
First seen
Licence
MIT

At a glance

当用户明确要求"从文件/图片/网页/描述中提取综述主题"、"生成主题+关键词+核心问题结构化输出",或要求使用旧名 get-review-theme skill 时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 research-literature-review…

  • Tasks that involve Literature review
  • SKILL.md covers 流程 and 约束
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Tasks that involve LaTeX

What it does

Research Topic Extractor is an agent skill from huangwb8/ChineseResearchLaTeX. 当用户明确要求"从文件/图片/网页/描述中提取综述主题"、"生成主题+关键词+核心问题结构化输出",或要求使用旧名 get-review-theme skill 时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 research-literature-review 及其他文献综述技能的结构化输出。

Its SKILL.md is about 670 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files (for example `CHANGELOG.md`, `README.md` and `config.yaml`).

It sits in Documents & Office, covering Literature review, LaTeX and PDF. The licence is MIT.

When your agent uses it

  • Tasks that involve Literature review
  • Tasks that involve LaTeX
  • Tasks that involve PDF

Example prompts

  • “从文件/图片/网页/描述中提取综述主题”
  • “生成主题+关键词+核心问题结构化输出”
  • “/research-topic-extractor”

What it can do on your machine

Read from SKILL.md and the folder at commit b8b4142. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Research Topic Extractor loads about 666 tokens when it runs, and up to ~1.1k if it reads all its reference files. Until then it costs about 58 tokens; SKILL.md has 179 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~58
When it runs · the whole SKILL.md, loaded when a task matches
~666
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from huangwb8/ChineseResearchLaTeX at commit b8b4142, republished under its MIT licence (© huangwb8). 179 words, ~666 tokens.

Download SKILL.mdSave it as .claude/skills/research-topic-extractor/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
research-topic-extractor
description
当用户明确要求"从文件/图片/网页/描述中提取综述主题"、"生成主题+关键词+核心问题结构化输出",或要求使用旧名 get-review-theme skill 时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 research-literature-review 及其他文献综述技能的结构化输出。
metadata.author
Bensz Conan

Research Topic Extractor

  • 从文件、图片、网页、文件夹或自然语言描述中提取结构化综述主题。
  • 输出直接服务 research-literature-review 或其他文献综述工作流。
  • 兼容旧名 get-review-theme 的 prompt 触发;系统级旧目录由安装器清理。
  • 最高原则:主题要可操作、关键词要能检索、核心问题要具体。
输入

必需:

  • {输入源}:文件路径、URL、文件夹路径、图片路径,或直接文本描述

可选:

  • {输出格式}:text / yaml / json,默认 text

流程

输入

按用户请求和配置文件提供必要输入;缺失信息应明确列出并停止依赖该输入的步骤。

执行步骤
  • 当用户环境中出现因本 skill 设计缺陷导致的 bug 时,优先使用 bensz-collect-bugs 按规范记录到 ~/.bensz-skills/bugs/,严禁直接修改用户本地 Claude Code / Codex 中已安装的 skill 源码。
  • 若 AI 仍可通过 workaround 继续完成用户任务,应先记录 bug,再继续完成当前任务。
  • 当用户明确要求“report bensz skills bugs”等公开上报动作时,调用本地 gh 与 bensz-collect-bugs,仅上传新增 bug 到 huangwb8/bensz-bugs;不要 pull / clone 整个 bug 仓库。
识别输入类型
  • 自然语言描述
  • 图片
  • URL
  • 文本文件
  • PDF
  • Word
  • 文件夹
提取内容
  • 自然语言:直接使用
  • 图片:依赖 LLM 原生视觉能力
  • URL:优先网页读取工具,失败则请用户提供正文
  • 文本 / PDF / Word:直接读取
  • 文件夹:递归扫描并合并 .md/.txt/.pdf 等核心材料

原则:

  • 优先用宿主原生能力和现有标准工具
  • 工具不可用时优雅降级,不额外引入脚本依赖
语义提取

围绕以下任务输出:

  • 用一句话概括主题
  • 提取 5-10 个英文标准术语
  • 提取 2-5 个具体研究问题或挑战
格式化
  • text:适合直接复制给下游 skill

  • yaml / json:适合结构化衔接

  • topic 可直接喂给 research-literature-review

  • keywords 可补充检索策略

  • core_questions 可作为综述边界和纳排参考

输出

始终包含三项:

  • 主题
  • 关键词
  • 核心问题

格式由用户选择:

  • text
  • yaml
  • json
输出管理

本 Skill 的新任务中间文件统一写入 ./.bensz-api/task-{yyyymmdd-hhmm}-{简短描述}/{skill名}/input|output|log/。同一任务复用一个任务根目录;多 Skill 协作才创建 shared/。正式交付物不写入该目录,历史隐藏目录只允许显式兼容读取、迁移或清理。

校验
  • 主题要包含研究对象与核心问题或方法
  • 关键词优先用标准检索术语
  • 核心问题必须具体,避免“意义重大/挑战很多”这种空话
失败与恢复
  • 文件不存在:提示用户改路径或直接粘贴内容
  • 格式不支持:提示转换
  • 内容提取失败:让用户手动提供文本
  • URL 解析失败:让用户复制网页正文或提供 PDF
  • 图片语义不清:请用户补一句描述

约束

遵守以下公共约束,并执行本 Skill 的专属边界。

公共硬约束
  • 任务需要落盘时,使用唯一的 ./.bensz-api/task-{yyyymmdd-hhmm}-{简短描述}/ 根目录;共享材料放入 shared/,Skill 专属材料放入该 Skill 的 input/、output/、log/。
  • 正式交付物、源代码和正式计划按项目约定保存,不写入任务工作区;未经授权不覆盖、删除、迁移或远程写入。
  • 项目维护变更检查 BAC 可用性并记录需求、AI 产出、工具结果、文件改动和验证摘要;BAC 只做过程审计,不替代署名、责任或合规判断。
  • 不记录 API Key、访问令牌、密码、Cookie、环境/凭据文件、私有 Prompt、身份信息、本地用户名、主机名或不必要的大体积原始数据。
  • 文件路径必须规范化并限制在授权项目范围内;外部 URL、子进程和网络访问遵循最小权限,防止路径遍历、SSRF 和命令注入。
  • Skill 版本唯一记录在自身 config.yaml:skill_info.version;公开 API、协议、目录或配置变更同步文档与 CHANGELOG.md。
  • 仅将 Skill 或 Bensz 基础设施本身的设计缺陷交给 bensz-collect-bugs;先脱敏写入 ~/.bensz-skills/bugs/,当前任务不中断,只有用户明确要求才公开上报,禁止直接修改用户已安装的 Skill 源码。
<!-- End of canonical common constraints. -->
Skill 专属约束

不得超出本 Skill description 和上方流程所声明的范围;不将未验证的信息伪装成确定结论。

© huangwb8, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files (references) in skills/research-topic-extractor of huangwb8/ChineseResearchLaTeX.

  • SKILL.md
  • CHANGELOG.md
  • README.md
  • config.yaml
  • references/prompt_templates.md

Open the folder on GitHubat commit b8b4142

Compare with similar skills

Research Topic Extractor next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Research Topic Extractor compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Research Topic Extractor this skillhuangwb8/ChineseResearchLaTeX2.9k—~666Automated safety check: PassMIT
Get Review ThemeInternScience/DrClaw172—~1.3kAutomated safety check: PassNone
Check Review AlignmentInternScience/DrClaw172—~1.5kAutomated safety check: PassNone
Typst Paperbahayonghang/academic-writing-skills498—~3.6kAutomated safety check: PassNone
Paper Interpreterchujianyun/skills740—~810Automated safety check: PassCustom licence
Evidence Ledgerwanshuiyin/Anti-Autoresearch161—~11kAutomated safety check: NotesMIT

Similar skills

  • Get Review Theme

    InternScience/DrClaw

    当用户明确要求"从文件/图片/网页/描述中提取综述主题"或"生成主题+关键词+核心问题结构化输出"时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 systematic-literature-review 及其他文献综述技能的结构化输出。

    172 GitHub stars~1.3k tokensUpdated 6 mo ago
    Documents & OfficeAuto-check passed
  • Check Review Alignment

    InternScience/DrClaw

    当用户明确要求"核查/优化综述 {主题}review.tex 的正文引用"或"运行 check-review-alignment"时使用。通过宿主 AI 的语义理解逐条核查引用是否与文献内容吻合,只在发现致命性引用错误时对"包含引用的句子"做最小化改写,并复用 systematic-literature-review 的渲染脚本输出…

    172 GitHub stars~1.5k tokensUpdated 6 mo ago
    Research & ScienceAuto-check passed
  • Typst Paper

    bahayonghang/academic-writing-skills

    Typst paper assistant for existing .typ manuscripts in English or Chinese.

    498 GitHub stars~3.6k tokensUpdated 11 days ago
    Documents & OfficeAuto-check passed
  • Paper Interpreter

    chujianyun/skills

    论文解读助手。适用于用户发送 arXiv 论文链接,并希望下载论文、解读论文、生成读书笔记、做论文拆解或输出详细报告时使用。会在工作目录创建论文文件夹、下载 PDF 与 TeX Source(如有)、生成中文 Markdown 报告。默认先交付初稿,不自动复查;如果用户明确同意,再安排后续复查。不适用于只要简短推荐语的情况。

    740 GitHub stars~810 tokensUpdated today
    Documents & OfficeAuto-check passed
  • Evidence Ledger

    wanshuiyin/Anti-Autoresearch

    Build the deterministic evidence ledger (artifactmanifest.json + claims.json) that every other Anti-Autoresearch auditor reads.

    161 GitHub stars~11k tokensUpdated 2 days ago
    Documents & OfficeAuto-check: notes
  • Paper Audit

    brycewang-stanford/Auto-Empirical-Research-Skills

    Deep-review-first audit for Chinese and English academic papers across LaTeX, Typst, and PDF formats.

    4.5k GitHub stars~3.6k tokensUpdated 4 days ago
    Documents & OfficeAuto-check passed

More from huangwb8/ChineseResearchLaTeX

All 27 skills in this repo
  • NSFC Grant Rationale Writer

    huangwb8/ChineseResearchLaTeX

    Writes, restructures, reviews and polishes the rationale section of NSFC research grant applications in LaTeX, with backups and a diff before every write.

    2.9k GitHub stars~945 tokensUpdated 5 days ago
    Auto-check passed
  • LaTeX Example Content Generator

    huangwb8/ChineseResearchLaTeX

    Fills an existing LaTeX project with sample sections, tables and figure narratives, protecting the template structure and previewing before writing.

    2.9k GitHub starsUsed in 1 repo~744 tokens
    Auto-check passed
  • Journal Selector for Manuscripts

    huangwb8/ChineseResearchLaTeX

    Recommends journals for a manuscript by filtering a bundled impact-factor catalog, verifying scope and quality online, and writing a ranked Markdown report.

    2.9k GitHub stars~1.7k tokensUpdated 5 days ago
    Auto-check passed
  • NSFC Abstract Writer

    huangwb8/ChineseResearchLaTeX

    Writes Chinese and English abstracts for NSFC grant applications, with a recommended title and five alternatives, within set character limits.

    2.9k GitHub starsUsed in 1 repo~1.3k tokens
    Auto-check passed
  • NSFC Budget Justification Writer

    huangwb8/ChineseResearchLaTeX

    Writes a submission-ready NSFC budget justification as a LaTeX project and renders budget.pdf from your grant proposal text and supporting materials.

    2.9k GitHub starsUsed in 1 repo~1.4k tokens
    Auto-check passed
  • NSFC Application Code Recommender

    huangwb8/ChineseResearchLaTeX

    Recommends five pairs of NSFC application codes, primary and secondary, from a grant proposal's text and writes the reasons to a Markdown report.

    2.9k GitHub starsUsed in 1 repo~1.1k tokens
    Auto-check passed

Questions about Research Topic Extractor

What does Research Topic Extractor do?

当用户明确要求"从文件/图片/网页/描述中提取综述主题"、"生成主题+关键词+核心问题结构化输出",或要求使用旧名 get-review-theme skill 时使用。支持文件(PDF/Word/Markdown/Tex)、文件夹、图片、自然语言描述、网页 URL 等多种输入源,自动识别输入类型并提取内容,生成可直接用于 research-literature-review…. Research Topic Extractor is an agent skill from huangwb8/ChineseResearchLaTeX.

When should I use Research Topic Extractor?

Research Topic Extractor fits situations like: tasks that involve Literature review; tasks that involve LaTeX; tasks that involve PDF.

How do I install Research Topic Extractor in Claude Code?

Run `npx skills add huangwb8/ChineseResearchLaTeX --skill research-topic-extractor -a claude-code`. Or copy the skill folder (skills/research-topic-extractor in huangwb8/ChineseResearchLaTeX) into .claude/skills/research-topic-extractor in your project. Claude Code loads it when a task matches its description.

How do I install Research Topic Extractor in Codex?

Run `npx skills add huangwb8/ChineseResearchLaTeX --skill research-topic-extractor -a codex`. Or copy the skill folder (skills/research-topic-extractor in huangwb8/ChineseResearchLaTeX) into .agents/skills/research-topic-extractor in your project. Codex loads it when a task matches its description.

Can I use Research Topic Extractor in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add huangwb8/ChineseResearchLaTeX --skill research-topic-extractor -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/research-topic-extractor, .gemini/skills/research-topic-extractor, .github/skills/research-topic-extractor and .opencode/skills/research-topic-extractor in your project.

What does Research Topic Extractor need to run?

SKILL.md names no scripts, command-line tools or credentials: Research Topic Extractor is instructions for the agent only.

Does Research Topic Extractor access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Research Topic Extractor safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Research Topic Extractor use?

Research Topic Extractor is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Research Topic Extractor use?

About 666 tokens (SKILL.md is roughly 2.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 451 tokens, read only when the agent opens those files.

What are the alternatives to Research Topic Extractor?

Skills that share tags, products or a category with Research Topic Extractor: Get Review Theme (InternScience/DrClaw, 172 stars), Check Review Alignment (InternScience/DrClaw, 172 stars), Typst Paper (bahayonghang/academic-writing-skills, 498 stars) and Paper Interpreter (chujianyun/skills, 740 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Research Topic Extractor?

huangwb8 (a GitHub user) maintains it in huangwb8/ChineseResearchLaTeX, which has 2,870 GitHub stars. The repository holds 27 skills in this directory. The repository was last updated on October 4, 2026.

Source: huangwb8/ChineseResearchLaTeX on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.