Search
OpenAI · Transcription
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Adds subtitles to video with the videocaptioner CLI: transcribes speech, tidies and translates the text, and burns styled subtitles into the video or exports SRT and ASS files. | WEIFENG2333/ | 16k | — | ~1.7k | Automated safety check: Pass | GPL-3.0 | 29 days ago |
| 2 | Transcribes audio with OpenAI's Whisper: 99 languages, translation to English, language detection, six model sizes and word-level timestamps, from Python or the CLI. | Orchestra-Research/ | 13k | 7 repos | ~1.9k | Automated safety check: Notes | MIT | 3 mo ago |
| 3 | Routes agents to the right KrillinAI command for subtitles, dubbing, video rendering, covers and speech, and explains how to read its JSON and manifest output. | krillinai/ | 13k | — | ~869 | Automated safety check: Pass | Apache-2.0 | today |
| 4 | Transcribe and summarize a video or podcast from a URL (YouTube, TikTok, Bilibili, Apple Podcasts, SoundCloud, 30+ platforms) or from a local media/.txt file. | wendy7756/ | 3.3k | — | ~937 | Automated safety check: Notes | Apache-2.0 | 26 days ago |
| 5 | Dub a video into another language and generate subtitles using the default Together + Cartesia stack. | shang-zhu/ | 1.1k | — | ~1k | Automated safety check: Notes | MIT | 1 mo ago |
| 6 | Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others. | decolua/ | 31k | — | ~914 | Automated safety check: Pass | MIT | 3 days ago |
| 7 | Transcribe audio files to text with optional diarization and known-speaker hints. | JetBrains/ | 366 | 4 repos | ~776 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 8 | OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1. | openclaw/ | 392k | 1 repo | ~518 | Automated safety check: Pass | MIT | today |
| 9 | Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API. | amd/ | 408 | — | ~5k | Automated safety check: Notes | MIT | yesterday |
| 10 | Transcribe audio via OpenAI Audio Transcriptions API (Whisper). | trpc-group/ | 1.9k | 12 repos | ~288 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | 11.Xsai A skill your agent uses when the user is building with xsai or any @xsai/ package, or is evaluating xsAI for a small OpenAI-compatible workflow with text generation, streaming, tool calling… | moeru-ai/ | 50k | 1 repo | ~1.3k | Automated safety check: Pass | MIT | today |
| 12 | Turns footage, audio and a storyboard plan into a finished short video with FFmpeg jump-cuts, subtitle burn-in and a final polish pass. | foryourhealth111-pixel/ | 3.6k | — | ~838 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 13 | 13.Ax Audio This skill helps an LLM generate correct audio code with @ax-llm/ax. | dosco/ | 107 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 14 | Speech-to-text via KeiRouter /v1/audio/transcriptions using OpenAI Whisper / Groq / Gemini / Deepgram / AssemblyAI models. | mydisha/ | 147 | — | ~680 | Automated safety check: Pass | MIT | 1 mo ago |
| 15 | Send inference requests through the FastLLM OpenAI-compatible gateway — chat completions, completions, embeddings, rerank, score, responses, moderations, audio speech and transcription, image… | azrtydxb/ | 108 | — | ~926 | Automated safety check: Pass | Apache-2.0 | today |
| 16 | Azure OpenAI SDK for .NET. An agent skill from microsoft/skills. | microsoft/ | 3.1k | 5 repos | ~3.4k | Automated safety check: Pass | MIT | yesterday |
| 17 | Transcribe audio files to text via POST /audio/transcriptions. | veniceai/ | 144 | — | ~1.7k | Automated safety check: Pass | MIT | 5 days ago |
| 18 | Transcribe audio and video from URLs (YouTube, direct media links) using WhisperKit locally. | nicepkg/ | 285 | — | ~1.5k | Automated safety check: Pass | MIT | 8 mo ago |
| 19 | Integrates local AI capabilities into applications using Embeddable Lemonade. | amd/ | 408 | — | ~6k | Automated safety check: Pass | MIT | yesterday |
| 20 | Send images, audio, video, or documents into an AG2 beta Agent alongside text. | ag2ai/ | 252 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 21 | Automatically fetch YouTube video transcripts, generate structured summaries, and send full transcripts to messaging platforms. | BrianRWagner/ | 441 | 1 repo | ~3.8k | Automated safety check: Pass | No licence | 6 mo ago |
| 22 | A skill your agent uses when transcribing non-realtime speech with Alibaba Cloud Model Studio Qwen ASR models (qwen3-asr-flash, qwen-audio-asr, qwen3-asr-flash-filetrans). | cinience/ | 397 | — | ~1.3k | Automated safety check: Pass | MIT | 2 mo ago |
| 23 | A skill your agent uses for Azure AI: Search, Speech, OpenAI, Document Intelligence. | microsoft/ | 255 | 1 repo | ~852 | Automated safety check: Pass | MIT | yesterday |
| 24 | Add voice message transcription to ClaudeClaw using OpenAI's Whisper API. | sbusso/ | 194 | — | ~1.1k | Automated safety check: Notes | MIT | 1 mo ago |
| 25 | Video summarization for Bilibili, Xiaohongshu, Douyin, and YouTube. | LeoYeAI/ | 2.2k | — | ~4.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 26 | 26.Ax AI This skill helps an LLM generate correct AI provider setup and configuration code using @ax-llm/ax. | dosco/ | 107 | — | ~8.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 27 | 27.Pollinations Pollinations.ai API for AI generation - text, images, videos, audio, and analysis. | sundial-org/ | 663 | — | ~1.7k | Automated safety check: Pass | No licence | 7 mo ago |
| 28 | Speech-to-text transcription via OpenAI Whisper. An agent skill from coco-research/coco. | coco-research/ | 513 | — | ~964 | Automated safety check: Pass | Unknown | yesterday |
| 29 | 29.Parakeet Stt Local speech-to-text with NVIDIA Parakeet TDT 0.6B v3 (ONNX on CPU). | sundial-org/ | 663 | — | ~771 | Automated safety check: Pass | No licence | 7 mo ago |
| 30 | Transcribe audio files via OpenRouter using audio-capable models (Gemini, GPT-4o-audio, etc). | sundial-org/ | 663 | — | ~588 | Automated safety check: Pass | No licence | 7 mo ago |