Whisper Speech Recognition
Orchestra-Research/AI-Research-SKILLs
Transcribes audio with OpenAI's Whisper: 99 languages, translation to English, language detection, six model sizes and word-level timestamps, from Python or the CLI.
Transcribe audio to text via the OpenAI-compatible transcription endpoint.
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-python --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/api/telnyx-stt-python .claude/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .claude/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-pythonType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-python --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/api/telnyx-stt-python .agents/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .agents/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-python --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/api/telnyx-stt-python .cursor/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .cursor/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/majiayu000/claude-skill-registry.git --path skills/api/telnyx-stt-python--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-python --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/api/telnyx-stt-python .gemini/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .gemini/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-pythonInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/api/telnyx-stt-python .github/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .github/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install majiayu000/claude-skill-registry telnyx-stt-python --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/majiayu000/claude-skill-registry.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/api/telnyx-stt-python .opencode/skills/telnyx-stt-python && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "telnyx-stt-python" agent skill from https://github.com/majiayu000/claude-skill-registry/tree/main/skills/api/telnyx-stt-python into .opencode/skills/telnyx-stt-python/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "telnyx-stt-python", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
telnyx-stt-pythonTranscribe audio to text via the OpenAI-compatible transcription endpoint.
Telnyx Stt Python is an agent skill from majiayu000/claude-skill-registry. Transcribe audio to text via the OpenAI-compatible transcription endpoint. Supports multiple models, languages, and keyword biasing. Also lists available speech-to-text providers and service types.
Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `metadata.json`).
It sits in Media & Creative, covering Transcription and Speech recognition and synthesis. It works with Python and OpenAI. The repository describes itself as: Searchable Claude Code skills catalog with source-linked guides and generated registry artifacts. The licence is MIT.
Read from SKILL.md and the folder at commit 2d14a69. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pipFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
api.telnyx.comAlso links to:
platform.openai.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
TELNYX_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Telnyx Stt Python loads about 1.4k tokens when it runs. Until then it costs about 54 tokens; SKILL.md has 344 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from majiayu000/claude-skill-registry at commit 2d14a69, republished under its MIT licence (© majiayu000). 344 words, ~1,394 tokens.
.claude/skills/telnyx-stt-python/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.pip install telnyximport os
from telnyx import Telnyx
client = Telnyx(
api_key=os.environ.get("TELNYX_API_KEY"),
)All examples below assume client is already initialized as shown above.
All API calls can fail with network errors, rate limits (429), validation errors (422), or authentication errors (401). Always handle errors in production code:
import telnyx
try:
response = client.ai.audio.transcribe(
model="openai/whisper-large-v3-turbo",
url="https://example.com/audio.mp3",
)
except telnyx.APIConnectionError:
print("Network error — check connectivity and retry")
except telnyx.RateLimitError:
import time
time.sleep(1)
except telnyx.APIStatusError as e:
print(f"API error {e.status_code}: {e.message}")Common error codes: 401 invalid API key, 403 insufficient permissions,
404 resource not found, 422 validation error, 429 rate limited.
Transcribe an audio file to text. This endpoint is consistent with the OpenAI Transcription API and may be used with the OpenAI JS or Python SDK.
POST /ai/audio/transcriptions
| Parameter | Type | Required | Description |
|---|---|---|---|
url | string (URL) | Yes | URL of the audio file to transcribe. |
model | string | No | Model ID (e.g., openai/whisper-large-v3-turbo, distil-whisper/distil-large-v2). |
language | string | No | Language code (e.g., en, es, fr). |
prompt | string | No | Optional prompt to guide transcription style. |
response_format | enum | No | json, text, srt, verbose_json, vtt. Default: json. |
temperature | number | No | Sampling temperature (0-1). Default: 0. |
keywords | array[string] | No | Keyword biasing — improve accuracy for domain-specific terms. |
# Basic transcription
response = client.ai.audio.transcribe(
url="https://example.com/audio.mp3",
)
print(response.text)
# With specific model and language
response = client.ai.audio.transcribe(
url="https://example.com/audio.mp3",
model="openai/whisper-large-v3-turbo",
language="es",
)
print(response.text)
# With keyword biasing for domain-specific terms
response = client.ai.audio.transcribe(
url="https://example.com/audio.mp3",
keywords=["Telnyx", "API", "WebRTC", "SIP"],
)
print(response.text)
# Verbose JSON with segments
response = client.ai.audio.transcribe(
url="https://example.com/audio.mp3",
response_format="verbose_json",
)
for segment in response.segments:
print(f"[{segment.start:.1f}s - {segment.end:.1f}s] {segment.text}")Primary response fields:
response.text — Full transcription textresponse.duration — Audio duration in secondsresponse.segments — Array of segment objects (with start, end, text) when using verbose_json formatRetrieve a list of available speech-to-text providers and their service types.
GET /ai/audio/transcriptions/providers
response = client.ai.audio.list_providers()
for provider in response.providers:
print(f"{provider['name']} — {provider['service_type']}")
# Filter by provider name
response = client.ai.audio.list_providers(provider="telnyx")
for provider in response.providers:
print(f"{provider['name']} — {provider['service_type']}")
# Filter by service type
response = client.ai.audio.list_providers(service_type="transcription")
for provider in response.providers:
print(f"{provider['name']} — {provider['service_type']}")Primary response fields:
response.providers — Array of provider objects with name and service_typeThe Telnyx Agent CLI provides composite commands for STT:
# Transcribe audio
telnyx-agent stt --audio-url https://example.com/audio.mp3 --json
# With specific model and language
telnyx-agent stt --audio-url https://example.com/audio.mp3 --model openai/whisper-large-v3-turbo --language es --json
# List available providers
telnyx-agent stt-providers --json
# Filter by provider or service type
telnyx-agent stt-providers --provider telnyx --service-type transcription --jsonhttps://api.telnyx.com/v2/ai/openai.keywords to improve transcription accuracy for domain-specific terms, product names, or acronyms that generic models may mishear.openai/whisper-large-v3-turbo (fast, accurate) and distil-whisper/distil-large-v2 (lightweight). Check stt-providers for the full list.en, es, fr, de, ja, etc.). Omit to auto-detect.verbose_json to get timestamps and segments. Use srt or vtt for subtitle files.© majiayu000, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file in skills/api/telnyx-stt-python of majiayu000/claude-skill-registry.
Open the folder on GitHubat commit 2d14a69
We found 3 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in majiayu000/claude-skill-registry, which our catalogue first saw on October 7, 2026.
Telnyx Stt Python next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Telnyx Stt Python this skillmajiayu000/claude-skill-registry | 666 | 1 repos | ~1.4k | Automated safety check: Pass | MIT | |
| Whisper Speech RecognitionOrchestra-Research/AI-Research-SKILLs | 13k | 8 repos | ~1.9k | Automated safety check: Notes | MIT | |
| Claude Real VideoHUANGCHIHHUNGLeo/claude-real-video | 2.2k | — | ~639 | Automated safety check: Pass | MIT | |
| 9Router Speech-to-Textdecolua/9router | 30k | — | ~745 | Automated safety check: Pass | MIT | |
| Claude Real Video For AgentsHUANGCHIHHUNGLeo/claude-real-video | 2.2k | — | ~2k | Automated safety check: Notes | MIT | |
| Video To NotesKIRVO-REPORTING/video-to-notes | 105 | — | ~1.5k | Automated safety check: Pass | MIT |
Orchestra-Research/AI-Research-SKILLs
Transcribes audio with OpenAI's Whisper: 99 languages, translation to English, language detection, six model sizes and word-level timestamps, from Python or the CLI.
HUANGCHIHHUNGLeo/claude-real-video
Watch a video for the user. An agent skill from HUANGCHIHHUNGLeo/claude-real-video.
decolua/9router
Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others.
HUANGCHIHHUNGLeo/claude-real-video
Install and use crv (claude-real-video) — a tool that lets any AI agent watch videos by extracting scene-aware keyframes, deduplicating them, and transcribing audio.
KIRVO-REPORTING/video-to-notes
Use immediately for any bare YouTube or YouTube Shorts URL, youtu.be link, Bilibili or b23.tv link, or other video URL; do not ask what the user wants.
openclaw/openclaw
OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1.
majiayu000/claude-skill-registry
Multi-source deep research using firecrawl and exa MCPs. An agent skill from majiayu000/claude-skill-registry.
majiayu000/claude-skill-registry
Neural search via Exa MCP for web, code, and company research.
majiayu000/claude-skill-registry
Unified media generation via fal.ai MCP — image, video, and audio.
majiayu000/claude-skill-registry
Interact with Zotero reference management libraries using the pyzotero Python client.
majiayu000/claude-skill-registry
Search scientific papers and retrieve structured experimental data extracted from full-text studies via the BGPT MCP server.
majiayu000/claude-skill-registry
Perform pairwise sequence alignment using Biopython Bio.Align.PairwiseAligner.
Categories
Transcribe audio to text via the OpenAI-compatible transcription endpoint. Telnyx Stt Python is an agent skill from majiayu000/claude-skill-registry. Transcribe audio to text via the OpenAI-compatible transcription endpoint.
Telnyx Stt Python fits situations like: tasks that involve Transcription; tasks that involve Speech recognition and synthesis.
Run `npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a claude-code`. Or copy the skill folder (skills/api/telnyx-stt-python in majiayu000/claude-skill-registry) into .claude/skills/telnyx-stt-python in your project. Claude Code loads it when a task matches its description.
Run `npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a codex`. Or copy the skill folder (skills/api/telnyx-stt-python in majiayu000/claude-skill-registry) into .agents/skills/telnyx-stt-python in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add majiayu000/claude-skill-registry --skill telnyx-stt-python -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/telnyx-stt-python, .gemini/skills/telnyx-stt-python, .github/skills/telnyx-stt-python and .opencode/skills/telnyx-stt-python in your project.
Going by SKILL.md and its folder, Telnyx Stt Python needs the command-line tools its instructions call (pip) and credentials named TELNYX_API_KEY. Our summary lists: Python 3; A credential in TELNYX_API_KEY.
SKILL.md names 2 domains. In commands or code: api.telnyx.com; the agent is likely to contact it when it follows the instructions. As links in the text: platform.openai.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Telnyx Stt Python is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.4k tokens (SKILL.md is roughly 5.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Telnyx Stt Python: Whisper Speech Recognition (Orchestra-Research/AI-Research-SKILLs, 13k stars), Claude Real Video (HUANGCHIHHUNGLeo/claude-real-video, 2.2k stars), 9Router Speech-to-Text (decolua/9router, 30k stars) and Claude Real Video For Agents (HUANGCHIHHUNGLeo/claude-real-video, 2.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
majiayu000 (a GitHub user) maintains it in majiayu000/claude-skill-registry, which has 666 GitHub stars. The repository holds 1,273 skills in this directory. The repository was last updated on October 7, 2026.
Source: majiayu000/claude-skill-registry on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.