Topic · AI & LLM Engineering

Best natural language processing skills for Claude Code, Codex and other agents.

Skills that process text with classical and neural NLP techniques.
skills
141
official
3

Natural language processing skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Natural language processing skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Helps write and debug JavaScript or TypeScript that uses the compromise English NLP library for matching, entity extraction, tagging and sentence transforms.

spencermountain/compromise12k—~2kAutomated safety check: PassMITyesterday
2

Google API integration for blog performance: PageSpeed Insights, CrUX Core Web Vitals with 25-week history, Search Console performance, URL Inspection, Indexing API, GA4 organic traffic, NLP entity…

AgriciDaniel/claude-blog2.3k1 repo~3.3kAutomated safety check: NotesMIT5 days ago
3

Guide for writing correct code with the compromise rule-based NLP library: tagging, match syntax, in-place transforms and common tasks like tense changes and redaction.

spencermountain/compromise12k—~1.8kAutomated safety check: PassMITyesterday
4

Three modes for CS-conference papers (CVPR/ICCV/ECCV vision, ACL/EMNLP/NAACL NLP, ICLR/NeurIPS/ICML/AAAI ML).

Spark-To-Paper-Skills/paperjury1.2k—~5.3kAutomated safety check: PassMIT1 mo ago
5

Shows how to load, train and use fast Hugging Face tokenizers, with BPE, WordPiece and Unigram models, padding, truncation and alignment tracking.

Orchestra-Research/AI-Research-SKILLs13k7 repos~3.4kAutomated safety check: PassMIT3 mo ago
6

Fills in a model card for an OpenMed clinical NER or de-identification model from its evaluation reports: intended use, metrics, subgroups and limitations.

maziyarpanahi/openmed5.5k—~1.8kAutomated safety check: PassApache-2.0yesterday
7

Tokenize, tag, and analyze natural language text using Apple's NaturalLanguage framework and translate between languages with the Translation framework.

dpearson2699/swift-ios-skills1.2k1 repo~3.5kAutomated safety check: PassUnknown2 mo ago
8

Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems.

ModelCloud/GPTQModel1.3k—~1.1kAutomated safety check: PassUnknowntoday
9

Suggests candidate ICD-10-CM diagnosis and ICD-10-PCS procedure codes for clinical text extracted by OpenMed, with rationale for a certified coder to review.

maziyarpanahi/openmed5.5k—~2kAutomated safety check: PassApache-2.0yesterday
10

Global-installable, project-level academic rebuttal strategy skill for AI/ML/CV/NLP/Robotics papers.

xiongqi123123/awesome-rebuttal305—~3.3kAutomated safety check: PassMIT3 mo ago
11

Drive native macOS apps via interceptor macos : AX trees, background click/type/keys/drag/scroll, occluded or minimized window capture, browser chrome, URL bars, OS dialogs, Apple Events, trusted OS…

Hacker-Valley-Media/Interceptor514—~2kAutomated safety check: PassUnknown5 days ago
12

Search 2500+ curated ChatGPT and LLM open-source repositories.

taishi-i/awesome-ChatGPT-repositories3.3k—~3.8kAutomated safety check: PassCC0-1.03 days ago
13

Maps OpenMed-extracted, terminology-coded conditions, drugs and measurements into OMOP CDM v5.4 tables for OHDSI and ATLAS analytics.

maziyarpanahi/openmed5.5k—~1.9kAutomated safety check: PassApache-2.0yesterday
14

Detect crisis signals in user content using NLP, mental health sentiment analysis, and safe intervention protocols.

curiositech/some_claude_skills2433 repos~3.8kAutomated safety check: PassMIT1 mo ago
15

End-to-end Stellar development playbook. An agent skill from VelaPayments/vela-payments.

VelaPayments/vela-payments131—~1.8kAutomated safety check: PassMITyesterday
16

Direct access to Google's own SEO data via Search Console (Search Analytics, URL Inspection, Sitemaps), PageSpeed Insights v5, CrUX field data with 25-week history, Indexing API v3, GA4 organic…

seranking/seo-skills160—~4.8kAutomated safety check: PassMIT3 mo ago
17

Finds social risks such as housing instability or food insecurity in clinical notes and proposes matching ICD-10-CM Z-codes for a coder to confirm.

maziyarpanahi/openmed5.5k—~1.9kAutomated safety check: PassApache-2.0yesterday
18

Applies the mental models and frameworks of Andrej Karpathy (deep learning, former Director of AI at Tesla, founding member of OpenAI, Eureka Labs).

K-Dense-AI/mimeo282—~1.9kAutomated safety check: PassMIT1 mo ago
19

Analyze text content using both traditional NLP and LLM-enhanced methods.

liangdabiao/claude-data-analysis-ultra-main2901 repo~1.7kAutomated safety check: NotesNo licence5 mo ago
20

Analyze the token length of an OT-Agent conversation-format (ShareGPT-style) dataset — the per-trace distribution (median/p90/max) and/or counts under a token threshold + a metadata predicate (e.g.

open-thoughts/OpenThoughts-Agent301—~1.5kAutomated safety check: PassApache-2.09 days ago
21

This skill should be used when the user asks to "learn from Kaggle", "study Kaggle solutions", "analyze Kaggle competitions", or mentions Kaggle competition URLs.

Galaxy-Dawn/claude-scholar5.7k2 repos~940Automated safety check: PassMIT14 days ago
22

Compare several Japanese NLP libraries, models, or datasets for a keyword (a specific tool name, or a function/task like '形態素解析') across a handful of criteria chosen for that comparison, rendered as…

taishi-i/awesome-japanese-nlp-resources1k—~4.1kAutomated safety check: NotesCC0-1.0yesterday
23

Generates text embeddings locally with the sentence-transformers library for RAG, semantic search, clustering and similarity, with model picks for general, multilingual and legal text.

Orchestra-Research/AI-Research-SKILLs13k3 repos~1.6kAutomated safety check: PassMIT3 mo ago
24

Trains and uses SentencePiece tokenizers on raw text, with BPE or Unigram models, for multilingual and CJK projects that need a reproducible vocabulary.

Orchestra-Research/AI-Research-SKILLs13k3 repos~1.4kAutomated safety check: NotesMIT3 mo ago
25

Analyze current trends and challenges in Japanese NLP for a topic.

taishi-i/awesome-japanese-nlp-resources1k—~3.5kAutomated safety check: NotesCC0-1.0yesterday
26

學術研究實驗設計技能——從研究假設到可重現實驗計畫的完整流程。當使用者需要規劃實驗、設計 ablation study、選擇 baseline、確定評估指標,或問「我應該跑哪些實驗」時,一定要使用此技能。觸發詞包括:實驗設計、experiment design、ablation、baseline、跑什麼實驗、evaluation metric、如何驗證方法。適用於機器學習、NLP、CV…

voidful/academic-skills132—~1.2kAutomated safety check: PassMIT6 mo ago
27

Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets.

davila7/claude-code-templates32k12 repos~1.2kAutomated safety check: PassMITtoday
28

Uses the Deepgram Python SDK's Read API to analyze text for sentiment, summaries, topics and intents with client.read.v1.text.analyze, from raw text or a hosted URL.

deepgram/deepgram-python-sdk469—~1.4kAutomated safety check: PassMITyesterday
29

Consolidates BERTopic, LDA or NMF topic output into a theory-driven classification framework and writes the final labels back to an Excel file.

TyrealQ/q-skills108—~1kAutomated safety check: PassMIT14 days ago
30

Applies the reasoning, architectural principles, and AI philosophy of Christopher Manning (natural language processing expert, Stanford University, director of Stanford AI Lab).

K-Dense-AI/mimeo282—~1.8kAutomated safety check: PassMIT1 mo ago
31

Expert developer for Calcpad.Highlighter - tokenization, linting, content resolution, and language tooling.

imartincei/CalcpadCE109—~1kAutomated safety check: NotesMITyesterday
32

Search all Japanese NLP resources (libraries, models, datasets, tutorials, dictionaries, Hugging Face).

taishi-i/awesome-japanese-nlp-resources1k—~4.3kAutomated safety check: NotesCC0-1.0yesterday
33

Add or modify a Lizard language reader. An agent skill from terryyin/lizard.

terryyin/lizard2.5k—~1.1kAutomated safety check: PassUnknowntoday
34

Given a Japanese NLP GitHub repo/model/dataset (URL / owner/repo / tool name) OR a topic, find what's already in awesome-japanese-nlp-resources and discover related resources NOT yet listed…

taishi-i/awesome-japanese-nlp-resources1k—~6.5kAutomated safety check: NotesCC0-1.0yesterday
35
35.Transformers.jsOfficial

Runs pre-trained Hugging Face models in JavaScript or TypeScript with Transformers.js, in browsers or Node.js, Bun and Deno, for text, vision, audio and multimodal tasks.

huggingface/skills11k1 repo~6.2kAutomated safety check: PassApache-2.06 days ago
36

나노바나나 프롬프트의 정확성을 한글 번역본과 비교하여 검증하고, WebSearch로 팩트체크한 후 이슈별로 사용자 확인을 거쳐 수정합니다.

team-attention/stanford-cs146s-kr294—~2.2kAutomated safety check: PassNo licence7 mo ago
37

Azure AI Text Analytics SDK for sentiment analysis, entity recognition, key phrases, language detection, PII, and healthcare NLP.

microsoft/skills3.1k6 repos~2.4kAutomated safety check: PassMITyesterday
38

Scores news, announcements and macro events with the LLM, stores them in an event CSV and blends the decaying event signal with technical signals in signal_engine.py.

HKUDS/Vibe-Trading35k—~2.1kAutomated safety check: PassMITyesterday
39

Build, inspect, prepare, generate and export Overmind datasets in Data Workshop.

overmind-core/overmind544—~875Automated safety check: PassAGPL-3.0yesterday
40

Supports Gtars for local genomic interval models and set algebra, overlaps and counts, consensus and coverage, tokenization, fragment processing, and refget/BEDbase planning across Python, Rust, and…

K-Dense-AI/scientific-agent-skills48k1 repo~3.8kAutomated safety check: NotesMIT2 days ago
41

Query the PHEE pharmacovigilance event extraction dataset. An agent skill from QSong-github/DrugClaw.

QSong-github/DrugClaw1161 repo~694Automated safety check: PassNo licence1 mo ago
42

Score an OpenMed clinical or biomedical NER model against a user-supplied gold corpus with entity-level precision, recall, and F1, then break errors down per label.

maziyarpanahi/openmed5.5k—~1.7kAutomated safety check: PassApache-2.0yesterday
43

Combine OpenMed clinical NLP with Microsoft Presidio, spaCy, or LangChain through OpenMed's built-in interop adapter registry (openmed.interop).

maziyarpanahi/openmed5.5k—~2.2kAutomated safety check: PassApache-2.0yesterday
44

Orient and bootstrap any project that uses OpenMed, the on-device clinical and biomedical NLP library, for named-entity recognition, PHI de-identification, FHIR export, and evaluation.

maziyarpanahi/openmed5.5k—~1.4kAutomated safety check: PassApache-2.0yesterday
45

Authors computable phenotype and cohort definitions in the OHDSI ATLAS / CIRCE style over the OMOP CDM, combining standard concept sets with NLP-derived features that OpenMed extracts.

maziyarpanahi/openmed5.5k—~1.9kAutomated safety check: PassApache-2.0yesterday
46

Run clinical and biomedical named-entity recognition on medical text with OpenMed's analyzetext.

maziyarpanahi/openmed5.5k—~1.9kAutomated safety check: PassApache-2.0yesterday
47

Replace detected PHI with realistic, type-matched fake values in OpenMed so clinical notes stay readable and parseable instead of full of [REDACTED] markers.

maziyarpanahi/openmed5.5k—~1.8kAutomated safety check: PassApache-2.0yesterday
48

Extract arbitrary, custom entity types from clinical or biomedical text with no fine-tuning using OpenMed's GLiNER / GLiNER2 zero-shot support.

maziyarpanahi/openmed5.5k—~1.7kAutomated safety check: PassApache-2.0yesterday

Questions, answered from the data.

What is the best natural language processing skill?

Compromise NLP Library from spencermountain/compromise ranks first of the 141 natural language processing skills listed here, with the highest score: its repository has 12k GitHub stars, its SKILL.md loads about 2k tokens and it passes the automated safety check with no findings. Next come Blog Google and Compromise NLP for JavaScript.

Which natural language processing skills are official?

3 of the 141 natural language processing skills are official, published by the vendor's own GitHub organization: Transformers.js, Azure AI Textanalytics Py and Azure AI Language Conversations Py.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.