Topic · Media & Creative
Best transcription skills, page 9
Transcription skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 385 | Simulate transcription factor perturbation effects on cell state in silico with CellOracle and Dynamo, and predict transcriptional responses to genetic perturbations with GEARS, scGen, and CPA. | GPTomics/ | 1.2k | 1 repo | ~3.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 386 | Infer transcription factor regulons from single-cell RNA-seq with pySCENIC by combining GRNBoost2 co-expression, cisTarget motif-enrichment pruning, and AUCell per-cell activity scoring. | GPTomics/ | 1.2k | 1 repo | ~3.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 387 | Designs pegRNAs and nicking guides for prime editing (PE) -- choosing the nick/strand, tuning the primer-binding site (PBS) and reverse-transcription template (RTT) as a per-locus panel, selecting… | GPTomics/ | 1.2k | 1 repo | ~4.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 388 | 388.Bio Motif Search Find sequence motifs, degenerate IUPAC patterns, and transcription-factor binding sites in DNA/RNA using Biopython and regex, including position weight matrix (PWM/PSSM) scoring. | GPTomics/ | 1.2k | 1 repo | ~2.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 389 | Transcribe DNA to RNA and translate to protein using Biopython, with NCBI codon-table selection, CDS validation, and six-frame ORF finding. | GPTomics/ | 1.2k | 1 repo | ~3.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 390 | 390.Jiaying Tool A skill your agent uses when editing videos for Xiaohongshu, creating short video content, adding effects and transitions to videos, or needing to add subtitles and music to video clips | vivy-yi/ | 481 | — | ~1.8k | Automated safety check: Pass | No licence | 8 mo ago |
| 391 | Ingest a voice note with exact-phrasing preservation (never paraphrased). | inbrainfun/ | 142 | 1 repo | ~1.7k | Automated safety check: Pass | Unknown | 2 mo ago |
| 392 | 392.Douyin Resolver 抖音视频解析技能。当用户发送抖音链接(douyin.com / v.douyin.com)、 抖音分享口令("复制打开抖音")、或提到解析抖音、下载抖音视频、 提取视频文案、视频转文字、语音转文本、抖音内容总结时,使用此技能。 | infometa/ | 348 | — | ~659 | Automated safety check: Notes | No licence | yesterday |
| 393 | 393.Speaker Id Name anonymous diarized SRT speakers from a persistent local voiceprint library and repair speaker drift; local, CPU-only, no API key. | BlackBeltTechnology/ | 315 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 394 | Transcribe video and audio files to SRT subtitles with speaker diarization using the Soniox (default) or AssemblyAI API. | BlackBeltTechnology/ | 315 | — | ~1.6k | Automated safety check: Notes | MIT | yesterday |
| 395 | 395.Youtube Srt Download YouTube subtitles (SRT + timestamped TXT) for a channel, playlist or video via yt-dlp, no media download; optionally mine them into a categorized catalog (e.g. | BlackBeltTechnology/ | 315 | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 396 | Transcribes audio or video to speaker-labeled, timestamped text, locally with MLX on Apple Silicon or remotely. | daymade/ | 1.4k | — | ~13k | Automated safety check: Pass | MIT | yesterday |
| 397 | 397.Transcript Fixer Corrects ASR/STT transcription errors — homophones, garbled terms, person-name errors, mixed Chinese/English — with dictionary rules plus Claude's built-in AI, no external API key required. | daymade/ | 1.4k | — | ~11k | Automated safety check: Pass | MIT | yesterday |
| 398 | 398.Youtube CLI Searches YouTube and fetches video transcripts via the cli-web-youtube command-line tool — video search, video details (views, duration, description, keywords), trending by category, channel info… | ItamarZand88/ | 231 | — | ~764 | Automated safety check: Pass | MIT | 10 days ago |
| 399 | 399.Ponyflash Generate images, videos, speech audio, and music using the PonyFlash Python SDK. | aiskillstore/ | 433 | 1 repo | ~4.7k | Automated safety check: Pass | MIT | yesterday |
| 400 | 用户给了视频链接或本机音视频,要转录、加字幕、做译文字幕或双语字幕时用:建可编辑视频并导入,转写、润色,需要时翻译,把字幕层放到画面上,再按需导出字幕文件或成片。有媒体时「转录」「加字幕」「翻译成某种语言」都按这个流程。不用于与视频无关的文字翻译。 | JimLiu/ | 611 | — | ~1.1k | Automated safety check: Pass | Unknown | yesterday |
| 401 | 用户要把视频的转写或字幕翻译成另一种语言、做双语字幕,或更新原文改动后过期的译文时用:逐句翻译写成视频里的译文文档,再放到画面上、按需导出。有媒体时「翻译」默认指字幕翻译。不用于翻译与视频无关的文章、文档或一段纯文本;对话里没有视频或字幕、又看不出指的是什么时先问一次。 | JimLiu/ | 611 | — | ~1.1k | Automated safety check: Pass | Unknown | yesterday |
| 402 | Bridge a live Podium call transcript or webchat turn to an LLM by fetching relevant historical conversation context as a structured RAG bundle — vector search over embedded prior conversations +… | jeremylongshore/ | 2.8k | — | ~5.2k | Automated safety check: Pass | MIT | yesterday |
| 403 | 403.Image To Text Converts one or more images into faithful text descriptions or OCR with the local 1.3B MiniCPM-V 4.6 GGUF model through llama.cpp, automatically preferring an available Vulkan GPU and falling back… | godot-fun/ | 183 | — | ~813 | Automated safety check: Pass | MIT | yesterday |
| 404 | 404.Storyboard Tts Converts storyboard markdown (bilingual Chinese + English narration per shot) into speech audio via IndexTTS2 (same stack as ai-text-to-speech). | godot-fun/ | 183 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 405 | 405.Video Transcript Deprecated compatibility package. An agent skill from bozhouDev/video-skills-toolkit. | bozhouDev/ | 150 | — | ~208 | Automated safety check: Pass | MIT | 2 mo ago |
| 406 | 406.Minutes Search Search past meeting transcripts and voice memos for specific topics, people, decisions, or ideas. | silverstein/ | 1.5k | — | ~1.2k | Automated safety check: Pass | MIT | 2 days ago |
| 407 | 407.Minutes Recap Generate a daily digest of today's policy-authorized meetings and voice memos — key decisions, action items, and themes across available recordings. | silverstein/ | 1.5k | — | ~931 | Automated safety check: Pass | MIT | 2 days ago |
| 408 | 408.Media Production Create and process finished videos with Remotion, Motion Canvas, Manim, FFmpeg or an available AI video provider. | WrongStack/ | 371 | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 409 | 409.Omh Media Input [omh] Audio, video, or screenshot to process: user-sent media - audio, video, YouTube links, screenshots, receipts, OCR, meeting recordings, transcripts, timestamps, and clip summaries, gated for… | rlaope/ | 3.2k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 410 | 410.Omh Voice Input [omh] Dictated voice note about project work or status: terse voice and mobile-style requests - turn short spoken-style asks into clarify, plan, status, handoff, or confirmation actions. | rlaope/ | 3.2k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 411 | Turn error logs, screenshots, voice notes, and rough bug reports into crisp, developer-ready GitHub issues with repro steps, impact, and evidence. | aiskillstore/ | 433 | 3 repos | ~1k | Automated safety check: Pass | No licence | yesterday |
| 412 | 412.Video Processor Process video files with audio extraction, format conversion (mp4, webm), and Whisper transcription. | aiskillstore/ | 433 | 1 repo | ~2.2k | Automated safety check: Pass | No licence | yesterday |
| 413 | 413.Wispr Analytics This skill should be used when analyzing Wispr Flow voice dictation history for self-reflection, work patterns, mental health insights, or productivity analytics AND when managing the Wispr Flow… | glebis/ | 391 | — | ~3.7k | Automated safety check: Pass | MIT | 3 days ago |
| 414 | 414.Chea API Access ChEA3 and Harmonizome ChEA data for transcription factor enrichment analysis and metadata retrieval. | aipoch/ | 1.9k | — | ~1.7k | Automated safety check: Pass | MIT | 24 days ago |
| 415 | Convert raw notes, error logs, voice dictation, or screenshots into crisp GitHub-flavored markdown issue reports. | microsoft/ | 3.1k | — | ~910 | Automated safety check: Pass | MIT | 2 days ago |
| 416 | A skill your agent uses when the user has a video + a target-language SRT and wants the video to actually speak that language — generates a time-aligned TTS voice dub. | jianshuo/ | 131 | — | ~5.3k | Automated safety check: Notes | MIT | 1 mo ago |
| 417 | 417.Video Clipper Repurposes long-form video (podcasts, interviews, talks) into short-form vertical clips for Instagram Reels, TikTok, and YouTube Shorts. | gooseworks-ai/ | 1.2k | 1 repo | ~3.1k | Automated safety check: Notes | MIT | 2 days ago |
| 418 | 418.Video Polish Takes an existing screen recording or demo video and adds professional zoom/pan effects synchronized to the narration. | gooseworks-ai/ | 1.2k | 1 repo | ~3.5k | Automated safety check: Notes | MIT | 2 days ago |
| 419 | Configure (enable OR disable) call recording and call transcription for Native Voice (Thunderbird Voice) programmatically via the Metadata API, for headless / API-driven support where no human uses… | forcedotcom/ | 1.1k | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 420 | 420.Voice Persona 让 Agent 变成能语音对话的机器人:语音文件转文字(支持微信 silk 格式)+ 多音色人格回复(Edge TTS 免费中文音色),全本地零 API 成本。Voice persona chat: transcribe voice messages (incl. | davepoon/ | 3.6k | — | ~947 | Automated safety check: Pass | MIT | 2 days ago |
| 421 | Framework-wide voice dictation in the agent sidebar composer. | BuilderIO/ | 7.1k | — | ~4.2k | Automated safety check: Pass | No licence | yesterday |
| 422 | 422.API Media Call nodetool.media or nodetool.generations from a code action: generate or edit images, video, speech and music with a picked model, transcribe and embed, judge images with a vision model, read a… | nodetool-ai/ | 560 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 423 | 423.Suede Aso Suede-owned app-store optimization discipline for keyword fields, titles, subtitles, descriptions, screenshots, ratings context, and competitor listing audits. | JasonColapietro/ | 127 | — | ~4.3k | Automated safety check: Pass | MIT | yesterday |
| 424 | 424.Hyperframes Create HyperFrames HTML video compositions — animations, title cards, overlays, captions, GSAP timelines, registry blocks/components, voiceovers, audio-reactive visuals, and scene transitions. | OpenMinis/ | 446 | — | ~6.2k | Automated safety check: Pass | MIT | 3 days ago |
| 425 | 425.Hyperframes CLI Use the HyperFrames CLI development loop: init, add, catalog, capture, lint, check, snapshot, compare, grade-compare, preview, play, present, beats, keyframes, single or batch render, publish… | aiskillstore/ | 433 | 1 repo | ~2.7k | Automated safety check: Pass | No licence | yesterday |
| 426 | Forge a title + subtitle (or reframe) for a BOOK, article, talk, or any technical piece, then hand off to a cover so it has LIFE and honesty — not clinical/dated. | Soul-Brews-Studio/ | 123 | — | ~2k | Automated safety check: Pass | MIT | 8 days ago |
| 427 | 427.Watch Extract YouTube video transcripts via yt-dlp and pipe to /learn. | Soul-Brews-Studio/ | 123 | — | ~1.3k | Automated safety check: Pass | MIT | 8 days ago |
| 428 | Convert physician verbal dictation into structured SOAP notes. | aipoch/ | 1.9k | — | ~2.5k | Automated safety check: Pass | MIT | 24 days ago |
| 429 | Thin orchestrator for the end-to-end video localization pipeline. | jianshuo/ | 131 | — | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 430 | A skill your agent uses when the user has an SRT (or transcript text) in one language and wants it translated to another, with punctuation-bounded re-segmentation so cues end at real sentence breaks. | jianshuo/ | 131 | — | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 431 | 431.Video Subtitles Generate SRT subtitles from video/audio with translation support. | aAAaqwq/ | 105 | 1 repo | ~566 | Automated safety check: Pass | MIT | 3 days ago |
| 432 | Video post-production rules: audio mastering, color, captions, platform export. | AnastasiyaW/ | 154 | — | ~1.7k | Automated safety check: Pass | MIT | yesterday |
Explore related skills
Category
More topics in Media & Creative
- Video production890
- Image generation724
- AI video generation667
- Text to speech and voice639
- Motion graphics387
- Design review and critique240
- Logo and visual identity226
- Comics and storyboards219
- Image editing193
- Social media graphics190
- Video scripts and shorts180
- Infographics159
- Music and audio generation150
- Podcasting122
- Generative and creative coding60