Agent skill

Speech Pathology AI

by curiositech in curiositech/some_claude_skills

Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication…

MITAuto-check passedAI & LLM Engineering

Install Speech Pathology AI

skills CLI
$ npx skills add curiositech/some_claude_skills --skill speech-pathology-ai -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install curiositech/some_claude_skills speech-pathology-ai --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/curiositech/some_claude_skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/speech-pathology-ai .claude/skills/speech-pathology-ai && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
speech-pathology-ai
GitHub stars
244
Used in
2 other repos
Token cost
~1.9k tokens
SKILL.md length
744 words
Files
7 (incl. references)
Skills in repo
95
Repo updated
First seen
Licence
MIT

At a glance

Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication…

  • Tasks that involve Speech recognition and synthesis
  • SKILL.md covers Python Dependencies, When to Use This Skill, Core Competencies and Anti-Patterns, plus 2 more sections
  • Calls pip
  • Tasks that involve Language learning

What it does

Speech Pathology AI is an agent skill from curiositech/some_claude_skills. Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication technology. Activate on 'speech therapy', 'articulation', 'phoneme analysis', 'voice disorder', 'fluency', 'stuttering', 'AAC', 'pronunciation', 'speech recognition', 'mellifluo.us'. NOT for general audio processing, music production, or voice acting coaching without clinical context.

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 8 other files, including reference files (for example `.claude-plugin/plugin.json`, `CHANGELOG.md` and `references/acoustic-analysis.md`).

It sits in AI & LLM Engineering, covering Speech recognition and synthesis and Language learning. The repository describes itself as: Claude skills that make my life easier. The licence is MIT.

When your agent uses it

  • Tasks that involve Speech recognition and synthesis
  • Tasks that involve Language learning

Example prompts

  • “speech therapy”
  • “articulation”
  • “phoneme analysis”
  • “/speech-pathology-ai”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Read, Write, Edit, Bash(python:*,pip:*), mcp__firecrawl__firecrawl_search, WebFetch, mcp__ElevenLabs__text_to_speech, mcp__ElevenLabs__speech_to_text

What it can do on your machine

Read from SKILL.md and the folder at commit 6713fc7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Write
    • Edit
    • Bash(python:*
    • pip:*)
    • mcp__firecrawl__firecrawl_search
    • WebFetch
    • mcp__ElevenLabs__text_to_speech
    • mcp__ElevenLabs__speech_to_text

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Speech Pathology AI loads about 1.9k tokens when it runs, and up to ~12k if it reads all its reference files. Until then it costs about 124 tokens; SKILL.md has 744 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~124
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~12k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from curiositech/some_claude_skills at commit 6713fc7, republished under its MIT licence (© curiositech). 744 words, ~1,942 tokens.

Download SKILL.mdSave it as .claude/skills/speech-pathology-ai/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
speech-pathology-ai
description
Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication technology. Activate on 'speech therapy', 'articulation', 'phoneme analysis', 'voice disorder', 'fluency', 'stuttering', 'AAC', 'pronunciation', 'speech recognition', 'mellifluo.us'. NOT for general audio processing, music production, or voice acting coaching without clinical context.
allowed-tools
Read, Write, Edit, Bash(python:*,pip:*), mcp__firecrawl__firecrawl_search, WebFetch, mcp__ElevenLabs__text_to_speech, mcp__ElevenLabs__speech_to_text
metadata.category
AI & Machine Learning
metadata.tags
speech-therapy, phonemes, articulation, voice, aac

Speech-Language Pathology AI Expert

You are an expert speech-language pathologist (SLP) with deep knowledge of phonetics, articulation disorders, voice therapy, fluency disorders, and AI-powered speech analysis. You specialize in building technology-assisted interventions, real-time feedback systems, and accessible communication tools.

Python Dependencies

bash
pip install praat-parselmouth librosa torch transformers numpy scipy

When to Use This Skill

Use for:

  • Phoneme-level accuracy scoring and feedback
  • Articulation disorder assessment tools
  • AI-powered speech therapy platforms
  • Real-time pronunciation feedback systems
  • Fluency (stuttering/cluttering) intervention tools
  • AAC (Augmentative and Alternative Communication) systems
  • Child speech recognition and analysis
  • mellifluo.us platform development

NOT for:

  • General audio/music production (use sound-engineer)
  • Voice acting or performance coaching
  • Accent modification without clinical indication
  • Diagnosing speech disorders (only licensed SLPs diagnose)

Core Competencies

Phonetics & Phonology
Consonant Classification by Place of Articulation
  • Bilabial: /p/, /b/, /m/ (both lips)
  • Labiodental: /f/, /v/ (lip + teeth)
  • Dental: /θ/, /ð/ (tongue + teeth) [think, this]
  • Alveolar: /t/, /d/, /n/, /s/, /z/, /l/, /r/ (tongue + alveolar ridge)
  • Postalveolar: /ʃ/, /ʒ/, /tʃ/, /dʒ/ [sh, zh, ch, j]
  • Palatal: /j/ [yes]
  • Velar: /k/, /g/, /ŋ/ [king, go, sing]
  • Glottal: /h/
Manner of Articulation
  • Stops: /p/, /b/, /t/, /d/, /k/, /g/ (complete blockage)
  • Fricatives: /f/, /v/, /θ/, /ð/, /s/, /z/, /ʃ/, /ʒ/, /h/ (turbulent air)
  • Affricates: /tʃ/, /dʒ/ (stop + fricative)
  • Nasals: /m/, /n/, /ŋ/ (air through nose)
  • Liquids: /l/, /r/ (partial obstruction)
  • Glides: /w/, /j/ (vowel-like)
Vowel Space (F1/F2 Formants)
         Front    Central    Back
High     /i/      /ɪ/        /u/    [ee, ih, oo]
                  /ə/               [schwa - unstressed]
Mid      /e/                 /o/    [ay, oh]
         /ɛ/      /ʌ/        /ɔ/    [eh, uh, aw]
Low      /æ/                 /ɑ/    [a, ah]

Diphthongs: /aɪ/, /aʊ/, /ɔɪ/ [eye, ow, oy]
State-of-the-Art AI Models (2024-2025)
PERCEPT-R Classifier (ASHA 2024)
  • Performance: 94.2% agreement with human SLP ratings
  • Architecture: GRU + wav2vec 2.0 with multi-head attention
  • Use case: Phoneme-level accuracy scoring in real-time
wav2vec 2.0 XLS-R for Children's Speech
  • Cross-lingual model fine-tuned for pediatric populations
  • Research shows 45% faster mastery with AI-guided practice
  • Fine-tuned on MyST (My Speech Technology) dataset

For detailed implementations, see /references/ai-models.md

Speech Analysis & Recognition

Acoustic Analysis Capabilities:

  • Formant extraction using Linear Predictive Coding (LPC)
  • MFCC (Mel-Frequency Cepstral Coefficients) for speech recognition
  • Voice Onset Time (VOT) detection for stop consonant analysis
  • Articulation precision measurement via formant space distance

For signal processing implementations, see /references/acoustic-analysis.md

Therapy Intervention Strategies

Evidence-Based Techniques:

  • Minimal Pair Contrast Therapy: Word pairs differing by single phoneme
  • Easy Onset: Gentle voice initiation for fluency
  • Prolonged Speech: Slow, stretched speech pattern for stuttering
  • AAC Integration: Symbol boards, word prediction, voice synthesis

For therapy implementations, see /references/therapy-interventions.md

mellifluo.us Platform Integration

Platform Architecture:

  • Real-time phoneme analysis with < 200ms latency
  • Adaptive practice engine with spaced repetition
  • Progress tracking and clinical dashboards
  • Gamification for engagement

Performance Benchmarks:

  • Latency: < 200ms end-to-end (audio → feedback)
  • Accuracy: 94.2% agreement with human SLP (PERCEPT-R)
  • Learning Gains: 45% faster mastery vs traditional therapy

For platform details, see /references/mellifluo-platform.md

Anti-Patterns

"One-Size-Fits-All" Therapy

What it looks like: Using the same exercises for all clients regardless of specific needs. Why it's wrong: Speech disorders are highly individual; what works for /r/ may not work for /s/. Instead: Individualize based on phoneme-specific challenges and baseline assessment.

Show full SKILL.md (284 more words)Show less
Technology Replacing Clinical Judgment

What it looks like: Relying solely on AI scores without SLP interpretation. Why it's wrong: AI is a tool, not a replacement for clinical expertise. Instead: Use AI for augmentation; trained SLPs interpret results and make treatment decisions.

Ignoring Generalization

What it looks like: Mastering sounds in isolation but never progressing to real conversation. Why it's wrong: The goal is functional communication, not perfect production in drills. Instead: Systematically progress: isolation → syllables → words → sentences → conversation.

Cultural Insensitivity

What it looks like: Treating bilingual speech patterns as disorders. Why it's wrong: Bilingualism is not a disorder; dialectal variations are normal. Instead: Distinguish between difference (normal variation) and disorder (clinical concern).

Best Practices

✅ DO:
  • Use evidence-based practices (cite SLP research)
  • Provide immediate feedback (visual + auditory)
  • Make therapy fun and engaging (gamification)
  • Track progress systematically (data-driven decisions)
  • Personalize to individual needs (adaptive difficulty)
  • Respect client autonomy (client chooses activities)
  • Ensure accessibility (multiple input methods)
  • Collaborate with families/caregivers (home practice)
❌ DON'T:
  • Diagnose without proper credentials (only licensed SLPs diagnose)
  • Provide one-size-fits-all therapy (individualize!)
  • Overwhelm with too many targets (focus on 1-2 sounds)
  • Ignore cultural/linguistic diversity (bilingualism is not a disorder)
  • Rely solely on drills (functional communication matters)
  • Forget to celebrate progress (even small wins)
  • Neglect carryover to real life (generalization is the goal)
  • Assume technology replaces human SLPs (it's a tool, not a replacement)

Integration with Other Skills

  • hrv-alexithymia-expert: Emotional awareness training for speech anxiety
  • sound-engineer: Audio processing and quality optimization

Remember: The goal of speech therapy is functional communication in real-life contexts. Technology should empower, engage, and accelerate progress—but the therapeutic relationship, clinical expertise, and individualized care remain irreplaceable. Make tools that SLPs love to use and clients are excited to practice with.

© curiositech, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (references) in .claude/skills/speech-pathology-ai of curiositech/some_claude_skills.

  • SKILL.md
  • .claude-plugin/plugin.json
  • CHANGELOG.md
  • references/acoustic-analysis.md
  • references/ai-models.md
  • references/mellifluo-platform.md
  • references/therapy-interventions.md

Open the folder on GitHubat commit 6713fc7

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in curiositech/some_claude_skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Speech Pathology AI next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Speech Pathology AI compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Speech Pathology AI this skillcuriositech/some_claude_skills2442 repos~1.9kAutomated safety check: PassMIT
TriageTalAter/annyang6.8k—~810Automated safety check: NotesMIT
Yichen Asrmcncarl/yichen-skills4.4k—~780Automated safety check: PassCustom licence
Dingtalk MinutesDingTalk-Real-AI/dingtalk-workspace-cli3.2k—~2.3kAutomated safety check: PassApache-2.0
Youtube FetcherJimmySadek/youtube-fetcher-to-markdown485—~3.1kAutomated safety check: PassMIT
Yichen Web Researchmcncarl/yichen-skills4.4k—~1.9kAutomated safety check: PassCustom licence

Similar skills

  • Triage

    TalAter/annyang

    Triage and close GitHub issues on TalAter/annyang. An agent skill from TalAter/annyang.

    6.8k GitHub stars~810 tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check: notes
  • Yichen Asr

    mcncarl/yichen-skills

    逸尘自用的统一音视频转写入口,在 StepFun Step ASR 与火山引擎豆包 ASR 之间按输出需求、安全边界和可用状态路由。用于本地音频或视频的纯文本转写、时间戳、SRT 字幕、口播粗剪,以及转写前体检;用户明确指定服务商时不得静默切换。Use when a local audio or video file needs transcription and the correct…

    4.4k GitHub stars~780 tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Dingtalk Minutes

    DingTalk-Real-AI/dingtalk-workspace-cli

    钉钉 AI 听记。Use when 查询或修改听记摘要、完整逐字稿、关键词、标签、行动项、录音、上传、思维导图、发言人洞察、ASR 热词/识别词配置或分享权限。写文档走 dingtalk-doc;建待办走 dingtalk-todo;日程走 dingtalk-calendar。命令前缀:dws minutes。

    3.2k GitHub stars~2.3k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Youtube Fetcher

    JimmySadek/youtube-fetcher-to-markdown

    Retrieve transcripts from YouTube, Instagram, TikTok, X, Vimeo and other video sites, summarize or analyze what was said (and shown on screen), or save an Obsidian-ready Markdown knowledge-base note…

    485 GitHub stars~3.1k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Yichen Web Research

    mcncarl/yichen-skills

    逸尘自用的互联网研究总入口。用于跨平台且跨阶段、用户尚未确定工具,或明确要求对公司、产品、人物、技术、行业和领域做横纵分析、发展史加现状对比或有来源约束的系统深度研究;先生成有截止日期和证据闸门的计划,再把搜索发现、候选核验、有限归档、按需转写和证据综合路由到…

    4.4k GitHub stars~1.9k tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Volcengine Asr

    ysyecust/lecture-to-notes

    Transcribe local audio or video with Volcengine Doubao file ASR, including BigASR 1.0 Turbo direct upload and asynchronous 1.0 standard, 1.0 idle, or 2.0 standard jobs through TOS.

    273 GitHub stars~783 tokensUpdated 8 days ago
    AI & LLM EngineeringAuto-check passed

More from curiositech/some_claude_skills

All 95 skills in this repo
  • Crisis Detection Intervention AI

    curiositech/some_claude_skills

    Detect crisis signals in user content using NLP, mental health sentiment analysis, and safe intervention protocols.

    244 GitHub starsUsed in 2 repos~3.8k tokens
    Auto-check passed
  • Form Validation Architect

    curiositech/some_claude_skills

    End-to-end form handling with react-hook-form, Zod schemas, validation patterns, error messaging, field arrays, and multi-step wizards.

    244 GitHub stars~3.8k tokensUpdated 1 mo ago
    Auto-check passed
  • GitHub Actions Pipeline Builder

    curiositech/some_claude_skills

    Build production CI/CD pipelines with GitHub Actions. An agent skill from curiositech/some_claude_skills.

    244 GitHub stars~2.8k tokensUpdated 1 mo ago
    Auto-check: notes
  • Background Job Orchestrator

    curiositech/some_claude_skills

    Expert in background job processing with Bull/BullMQ (Redis), Celery, and cloud queues.

    244 GitHub stars~3.2k tokensUpdated 1 mo ago
    Auto-check passed
  • Competitive Cartographer

    curiositech/some_claude_skills

    Strategic analyst that maps competitive landscapes, identifies white space opportunities, and provides positioning recommendations.

    244 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Computer Vision Pipeline

    curiositech/some_claude_skills

    Build production computer vision pipelines for object detection, tracking, and video analysis.

    244 GitHub stars~4k tokensUpdated 1 mo ago
    Auto-check passed

Questions about Speech Pathology AI

What does Speech Pathology AI do?

Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication…. Speech Pathology AI is an agent skill from curiositech/some_claude_skills. Expert speech-language pathologist specializing in AI-powered speech therapy, phoneme analysis, articulation visualization, voice disorders, fluency intervention, and assistive communication technology.

When should I use Speech Pathology AI?

Speech Pathology AI fits situations like: tasks that involve Speech recognition and synthesis; tasks that involve Language learning.

How do I install Speech Pathology AI in Claude Code?

Run `npx skills add curiositech/some_claude_skills --skill speech-pathology-ai -a claude-code`. Or copy the skill folder (.claude/skills/speech-pathology-ai in curiositech/some_claude_skills) into .claude/skills/speech-pathology-ai in your project. Claude Code loads it when a task matches its description.

How do I install Speech Pathology AI in Codex?

Run `npx skills add curiositech/some_claude_skills --skill speech-pathology-ai -a codex`. Or copy the skill folder (.claude/skills/speech-pathology-ai in curiositech/some_claude_skills) into .agents/skills/speech-pathology-ai in your project. Codex loads it when a task matches its description.

Can I use Speech Pathology AI in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add curiositech/some_claude_skills --skill speech-pathology-ai -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/speech-pathology-ai, .gemini/skills/speech-pathology-ai, .github/skills/speech-pathology-ai and .opencode/skills/speech-pathology-ai in your project.

What does Speech Pathology AI need to run?

Going by SKILL.md and its folder, Speech Pathology AI needs the command-line tools its instructions call (pip). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Read, Write, Edit, Bash(python:*,pip:*), mcp__firecrawl__firecrawl_search, WebFetch, mcp__ElevenLabs__text_to_speech, mcp__ElevenLabs__speech_to_text.

Does Speech Pathology AI access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Speech Pathology AI safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Speech Pathology AI use?

Speech Pathology AI is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Speech Pathology AI use?

About 1.9k tokens (SKILL.md is roughly 7.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 9.9k tokens, read only when the agent opens those files.

What are the alternatives to Speech Pathology AI?

Skills that share tags, products or a category with Speech Pathology AI: Triage (TalAter/annyang, 6.8k stars), Yichen Asr (mcncarl/yichen-skills, 4.4k stars), Dingtalk Minutes (DingTalk-Real-AI/dingtalk-workspace-cli, 3.2k stars) and Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Speech Pathology AI?

curiositech (a GitHub organization) maintains it in curiositech/some_claude_skills, which has 244 GitHub stars. The repository holds 95 skills in this directory. The repository was last updated on September 6, 2026.

Source: curiositech/some_claude_skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.