Topic · Media & Creative
Best text to speech and voice skills, page 2
Text to speech and voice skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | HyperFrames CLI tool — hyperframes init, lint, preview, render, transcribe, tts, doctor, browser, info, upgrade, compositions, docs, benchmark. | nateherkai/ | 1.2k | 3 repos | ~1.2k | Automated safety check: Pass | Unknown | 10 days ago |
| 50 | Turns text into spoken audio through a 9Router server, choosing among voices from OpenAI, ElevenLabs, Deepgram, Edge TTS and other providers. | decolua/ | 30k | — | ~765 | Automated safety check: Pass | MIT | 7 days ago |
| 51 | AivisSpeech エディタの Sentry issue を調査し、修正すべき Electron / Vue 側の不具合と、ローカル環境・ブラウザ実装・外部通信由来のノイズを切り分けるためのスキルです。Sentry 側で既知ノイズを整理する作業や、src/domain/sentryEventFilter.ts と関連テストを更新して既知ノイズを送信前に破棄する作業で使用します。 | Aivis-Project/ | 483 | — | ~636 | Automated safety check: Pass | LGPL-3.0 | 2 mo ago |
| 52 | Analyze images, video, speech, motion, products, and websites; route local vision, audio, depth, tracking, matting, identity, and restoration models; auto-edit, replicate, enhance, caption, voice… | MartinDelophy/ | 896 | — | ~7k | Automated safety check: Pass | MIT | yesterday |
| 53 | Diagnose and revise defensive academic writing while preserving claim ceilings, evidence status, scope conditions, rival explanations, and conceptual hierarchy. | lensback940701/ | 255 | — | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 54 | 当用户说想做一个视频、宣传片、产品演示、动画短片、抖音/YouTube 内容,或者说要改分镜、调节奏、换镜头、调字幕、加配音、改转场时使用。通过苏格拉底式追问收集视频需求,主动激发渲染层的全部能力(TTS / 字幕 / 3D / shader / 音频反应等),输出标准化的 video-spec.md 用于渲染。 | feicaiclub/ | 1k | — | ~2.9k | Automated safety check: Pass | MIT | 4 mo ago |
| 55 | 55.Qiaomu Cut 把一句话需求转成可复现、可验收视频工程的乔木智能剪辑导演。Use when the user asks to create, plan, edit, remix, explain, narrate, subtitle, animate, composite, or render a video—including one-line requests such as “制作一个科普视频:介绍… | joeseesun/ | 369 | — | ~6.8k | Automated safety check: Notes | MIT | 10 days ago |
| 56 | 56.Adaptation This skill should be used when the user asks to "make an audiobook", "narration script", "narrator", "ACX", "Findaway", "pronunciation guide", "how long is the audiobook", "adapt to a screenplay"… | danjdewhurst/ | 279 | 2 repos | ~3.5k | Automated safety check: Notes | MIT | today |
| 57 | 57.Podcast Generate a podcast episode from content you provide. An agent skill from zarazhangrui/personalized-podcast. | zarazhangrui/ | 437 | — | ~2.3k | Automated safety check: Notes | No licence | 6 mo ago |
| 58 | 从书名或飞书多维表格中的成稿文案出发,结合微信读书资料与公开点评创作图书带货/书评短视频,并用豆包 TTS、Pexels、Codex 生图和本机 OpenChatCut 完成配音、配图、双语字幕、音效、动效、BGM、可编辑初稿与按需导出。用户提出“根据一本书做带货视频”“读取飞书文案制作图书视频”“写书评口播并自动剪成抖音视频”“仿参考样式做图书推荐短视频”时使用;仅查书、仅写普通书评或无关剪辑… | Kianzzz/ | 215 | — | ~3.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 59 | 按需把一部成片拆成可复用的制作参考:测镜头节奏与响度,标注段落与音轨分工,把原片事实与可迁移方法分开, 导出不含原片人名台词的 productionreference.json 供下次制作参考。不在默认生产路径上。 | zenstory-ai/ | 555 | 1 repo | ~1.3k | Automated safety check: Pass | MIT | 4 days ago |
| 60 | Orchestrate a configurable, debuggable local pipeline from approved narration packages to timed audio, captions, visual assets, semantic motion, rendered video, optional avatar compositing… | JayceHuang/ | 103 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 61 | Generates Chinese speech audio from text using Volcengine's Doubao TTS API, for narration, voiceovers or multi-voice podcasts. | xvirobotics/ | 991 | — | ~1.1k | Automated safety check: Notes | MIT | 23 days ago |
| 62 | 62.Ig Reel Write an Instagram Reel from a raw idea - hook options off 26 formulas, the spoken script, the on-screen text, and a timed beat sheet - in the user's own voice and scored before they shoot it. | Jakeschincariol/ | 615 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 25 days ago |
| 63 | Makes the Tiny Engineer desk robot speak a short WAV clip through its speaker while it gestures, by posting to its /play endpoint with a bundled script. | jamro/ | 565 | — | ~492 | Automated safety check: Pass | Unknown | yesterday |
| 64 | Renders landscape videos with the KrillinAI CLI, either the original footage with bilingual subtitles or a dubbed video with target-language subtitles. | krillinai/ | 13k | 1 repo | ~422 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 65 | Create Chinese HBG “模拟人生 / 人生副本” narrative videos with a consistent comic IP, a rapid multi-life opening, continuous natural-speed narration, synchronized short captions, dense static manga… | Mr-funny/ | 140 | — | ~5.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 66 | Prepare, validate, publish, and verify Megaphone releases. An agent skill from Kuberwastaken/megaphone. | Kuberwastaken/ | 169 | — | ~1.3k | Automated safety check: Pass | MIT | 2 mo ago |
| 67 | AI educational video production pipeline. An agent skill from speechlab0210/video-production-skill. | speechlab0210/ | 105 | — | ~4.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 68 | For creators, educators, and social-video editors who need a tactile paper-collage language for narration, knowledge points, opinions, or abstract topics. | tl2012tl/ | 239 | 4 repos | ~5.2k | Automated safety check: Pass | No licence | 16 days ago |
| 69 | 69.Create Video Tạo một video MỚI cho series "so sánh / phân biệt kiến thức" của repo này — clip dọc TikTok/Reels/Shorts 30-40s, layout 3-zone cố định theo DESIGN.md, voiceover tiếng Việt sinh bằng VieNeu TTS… | Cuongyd196/ | 210 | — | ~2.8k | Automated safety check: Notes | Unknown | 15 days ago |
| 70 | Designs and attaches voice samples or final narration/line audio on the filmmaking canvas via the local generatevoice.js CLI. | Utopai-Research/ | 354 | 1 repo | ~631 | Automated safety check: Pass | Unknown | 9 days ago |
| 71 | Generate smooth hand-drawn whiteboard and story videos directly inside Codex from text scripts, GPT Image 2 color storyboards, scene plans, SVGs, line art, or local images. | gnipbao/ | 324 | — | ~7.2k | Automated safety check: Notes | MIT | 1 mo ago |
| 72 | End-to-end AI video production skill for agentic frameworks. | Bomx/ | 308 | — | ~11k | Automated safety check: Notes | No licence | 2 mo ago |
| 73 | 73.Explainroo Make an explainer video (MP4) with a voice-over using explainroo. | vincentsch/ | 489 | — | ~439 | Automated safety check: Pass | MIT | 4 days ago |
| 74 | Smallest personal narrated-explainer-video pipeline (spoken narration over visuals, not an audio podcast), fully tool-agnostic and autonomous by default — topic → research ∥ asset collection →… | Agents365-ai/ | 1.7k | — | ~3.6k | Automated safety check: Pass | MIT | 7 days ago |
| 75 | 75.Video Script 对已完成分析的视频进行导演与剪辑策划,再写带时间戳的中文解说并校验;也处理已有短片的 宣发标题、花字修订和外部文案回填。普通策划输入 workdir 的 agentnarrationbrief.md 与 vlmanalysis.json;文案返修输入当前成片的工程与内容证据。策划输出 recapstoryplan.json、visualaudioboard.json、 可选… | zenstory-ai/ | 466 | — | ~2k | Automated safety check: Pass | MIT | 5 days ago |
| 76 | AivisSpeech Engine の Sentry issue を調査し、修正すべきエンジン側の不具合と、入力値・ローカル環境・外部サービス由来のノイズを切り分けるためのスキルです。Sentry 側で既知ノイズを永続アーカイブする作業や、voicevoxengine/utility/sentryutility.py と関連テストを更新して既知ノイズを送信前に破棄する作業で使用します。 | Aivis-Project/ | 181 | — | ~545 | Automated safety check: Pass | LGPL-3.0 | yesterday |
| 77 | Transcribes audio/video files using ElevenLabs Scribe v2 API. | qdhenry/ | 1.3k | — | ~1.5k | Automated safety check: Notes | No licence | 7 mo ago |
| 78 | 78.Showtime A skill your agent uses when the user wants a video made, edited or finished: a launch or promo, product demo, explainer, trailer or teaser, tutorial or walkthrough, a screen recording turned into a… | FavioVazquez/ | 178 | — | ~3k | Automated safety check: Pass | MIT | today |
| 79 | The repeatable workflow for short Thai documentary/explainer videos — one topic in, one finished MP4 out. | killernay/ | 140 | — | ~10k | Automated safety check: Notes | MIT | 2 mo ago |
| 80 | Expert guidance on iOS accessibility best practices, patterns, and implementation. | dadederk/ | 172 | — | ~2.4k | Automated safety check: Pass | MIT | 7 mo ago |
| 81 | Validates a KrillinAI multi-stage output plan with the pipeline command's dry run, then maps it to the individual stage commands that do the real work. | krillinai/ | 13k | — | ~488 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 82 | Generates images, audio and video through Leon's media tools, joins them with FFmpeg, checks the output and attaches playable files for the owner. | leon-ai/ | 18k | — | ~622 | Automated safety check: Pass | MIT | today |
| 83 | Turn ONE topic, talking-head video, or photo into a finished Vox-style paper-collage explainer / ad video on the MuAPI platform (api.muapi.ai) + local ffmpeg — script, collage keyframes, motion… | Anil-matcha/ | 243 | — | ~679 | Automated safety check: Pass | No licence | 1 mo ago |
| 84 | 84.Gemini Tts Generates spoken MP3 audio from text or Markdown with Gemini TTS. | iurysza/ | 419 | 1 repo | ~968 | Automated safety check: Pass | MIT | yesterday |
| 85 | Maintain VoiceOver/TalkBack-focused accessibility in stream-chat-react-native. | GetStream/ | 1.2k | — | ~6.5k | Automated safety check: Pass | Unknown | yesterday |
| 86 | 把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)或显式配置的通用 IndexTTS HTTP 服务逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 ttssegments 与 ttsmeta.json。 | zenstory-ai/ | 555 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 4 days ago |
| 87 | 87.Motion Video Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote… | bestagentkits/ | 113 | — | ~1.5k | Automated safety check: Pass | MIT | 10 days ago |
| 88 | 88.Runninghub Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId. | HM-RunningHub/ | 141 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 89 | 89.Kkclaw 给你的 AI Agent 一个桌面身体 — Setup Wizard、14情绪球体、语音克隆、歌词窗、Doctor 自检、跨平台支持(Windows + macOS) | kk43994/ | 174 | — | ~576 | Automated safety check: Pass | MIT | 5 mo ago |
| 90 | 90.Oma Video Create short, explainer, or recorded-demo videos through the OMA video CLI. | first-fluke/ | 1.3k | 1 repo | ~1.7k | Automated safety check: Pass | MIT | today |
| 91 | MiniMax multimodal model skill — use MiniMax Multi-Modal models for speech, music, video, and image. | poco-ai/ | 1.4k | — | ~7.6k | Automated safety check: Pass | MIT | 16 days ago |
| 92 | Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details… | mikefutia/ | 102 | — | ~747 | Automated safety check: Notes | No licence | 5 mo ago |
| 93 | Turns an idea, article, outline or audio file into a sourced, reviewable AI video, tracking whether narration uses a human, synthetic or cloned voice. | wanghui2323/ | 101 | — | ~924 | Automated safety check: Pass | MIT | 1 mo ago |
| 94 | Run wally's LLM e2e on Apple Neural Engine (NeuRT) and Snapdragon Hexagon NPU (QHexRT) devices. | RunanywhereAI/ | 1.6k | — | ~845 | Automated safety check: Pass | MIT | today |
| 95 | Turn arbitrary source content (README, article, story, slides, deck, data/report, product description, tutorial text, audio/transcript, or a bare topic) into a finished, high-quality MP4 video. | architectds/ | 117 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 96 | 96.Media Gen Generate or edit images, videos, or audio in the current task. | clacky-ai/ | 1.2k | — | ~7.3k | Automated safety check: Pass | MIT | today |
Explore related skills
Category
More topics in Media & Creative
- Video production878
- Image generation721
- AI video generation664
- Transcription582
- Motion graphics378
- Design review and critique237
- Logo and visual identity225
- Comics and storyboards217
- Image editing194
- Social media graphics188
- Video scripts and shorts182
- Infographics157
- Music and audio generation150
- Podcasting120
- Generative and creative coding59