Search
Whisper · Speech recognition and synthesis
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Transcribes audio with OpenAI's Whisper: 99 languages, translation to English, language detection, six model sizes and word-level timestamps, from Python or the CLI. | Orchestra-Research/ | 13k | 7 repos | ~1.9k | Automated safety check: Notes | MIT | 3 mo ago |
| 2 | Transcribes audio files into text or subtitles through 9Router's Whisper-compatible endpoint, using models from OpenAI, Groq, Gemini, Deepgram and others. | decolua/ | 31k | — | ~914 | Automated safety check: Pass | MIT | 3 days ago |
| 3 | OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1. | openclaw/ | 392k | 1 repo | ~518 | Automated safety check: Pass | MIT | today |
| 4 | Transcribe audio via OpenAI Audio Transcriptions API (Whisper). | trpc-group/ | 1.9k | 12 repos | ~288 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 5 | A skill your agent uses when users provide YouTube, Bilibili, or X/Twitter lecture URLs and want reader-first Chinese LaTeX/PDF notes with source-faithful claims, fluent authored prose, and verified… | ysyecust/ | 273 | — | ~14k | Automated safety check: Notes | Unknown | 8 days ago |
| 6 | Convert a HuggingFace ASR fine-tune into a sherpa-onnx external model, publish it, and add it to the Anti-Vocale community catalog. | RisorseArtificiali/ | 118 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 7 | Local speech-to-text with the Whisper CLI (no API key). An agent skill from huangruiteng/CS-Notes. | huangruiteng/ | 4k | 18 repos | ~228 | Automated safety check: Pass | MIT | 3 days ago |
| 8 | Speech-to-text via KeiRouter /v1/audio/transcriptions using OpenAI Whisper / Groq / Gemini / Deepgram / AssemblyAI models. | mydisha/ | 147 | — | ~680 | Automated safety check: Pass | MIT | 1 mo ago |
| 9 | 把 TikTok、Reels、YouTube Shorts、UGC 广告、本地 MP4/MOV/WebM 等视频拆成 Codex 可读的视频上下文。 | binggandata/ | 605 | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 10 | A skill your agent uses when the user wants local voice transcription instead of OpenAI Whisper API. | sbusso/ | 194 | — | ~1.3k | Automated safety check: Notes | MIT | 1 mo ago |
| 11 | 11.Watch Video When you want to extract content from a video — YouTube, Loom, Vimeo, Riverside, Zoom recording, local MP4, X/IG video, anything yt-dlp supports. | coreyhaines31/ | 851 | — | ~3.8k | Automated safety check: Pass | MIT | 3 days ago |
| 12 | 本地录音转文字工具。当用户发送已有录音、音频或视频文件,并希望把语音转成 Markdown 文稿和 SRT 字幕时使用。Apple Silicon 优先用 MLX/Apple GPU 和 whisper-large-v3-turbo-q4,本地转写,不生成 txt/json/vtt,不用于现场临时录音,也不默认调用云端语音识别服务。 | chujianyun/ | 742 | — | ~1k | Automated safety check: Pass | Unknown | 2 days ago |
| 13 | A skill your agent uses when the user has audio or video and wants a timestamped transcript (SRT) in the source language. | jianshuo/ | 131 | — | ~4.4k | Automated safety check: Notes | MIT | 1 mo ago |
| 14 | Automated video editing skill for talk/vlog/standup videos. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~2.7k | Automated safety check: Notes | MIT | 2 mo ago |
| 15 | Transcribe audio/video to text with word-level timestamps using OpenAI Whisper. | benchflow-ai/ | 1.8k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 16 | Deep dive into migrating to Deepgram from other transcription providers. | jeremylongshore/ | 2.8k | — | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 17 | Transcribes audio and video files to text using pluggable ASR backends. | swyxio/ | 176 | — | ~8.5k | Automated safety check: Pass | MIT | 6 days ago |
| 18 | 18.Whisper OpenAI Whisper for speech recognition and transcription — local inference, multiple model sizes, language detection, and subtitle generation. | AlexAI-MCP/ | 135 | — | ~1.9k | Automated safety check: Pass | MIT | 6 mo ago |
| 19 | Speech-to-text transcription via OpenAI Whisper. An agent skill from coco-research/coco. | coco-research/ | 513 | — | ~964 | Automated safety check: Pass | Unknown | yesterday |
| 20 | Local speech-to-text using faster-whisper. An agent skill from sundial-org/awesome-openclaw-skills. | sundial-org/ | 663 | — | ~3k | Automated safety check: Pass | No licence | 7 mo ago |
| 21 | Transcribe audio via OpenAI Whisper, Atlas Cloud, or MuAPI speech-to-text APIs. | CoWork-OS/ | 477 | — | ~411 | Automated safety check: Pass | MIT | yesterday |