Search
Text to speech and voice
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Diagnose and revise defensive academic writing while preserving claim ceilings, evidence status, scope conditions, rival explanations, and conceptual hierarchy. | lensback940701/ | 259 | — | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 50 | 当用户说想做一个视频、宣传片、产品演示、动画短片、抖音/YouTube 内容,或者说要改分镜、调节奏、换镜头、调字幕、加配音、改转场时使用。通过苏格拉底式追问收集视频需求,主动激发渲染层的全部能力(TTS / 字幕 / 3D / shader / 音频反应等),输出标准化的 video-spec.md 用于渲染。 | feicaiclub/ | 1k | — | ~2.9k | Automated safety check: Pass | MIT | 4 mo ago |
| 51 | 51.Qiaomu Cut 把一句话需求转成可复现、可验收视频工程的乔木智能剪辑导演。Use when the user asks to create, plan, edit, remix, explain, narrate, subtitle, animate, composite, or render a video—including one-line requests such as “制作一个科普视频:介绍… | joeseesun/ | 372 | — | ~6.8k | Automated safety check: Notes | MIT | 13 days ago |
| 52 | 52.Podcast Generate a podcast episode from content you provide. An agent skill from zarazhangrui/personalized-podcast. | zarazhangrui/ | 438 | — | ~2.3k | Automated safety check: Notes | No licence | 6 mo ago |
| 53 | 从书名或飞书多维表格中的成稿文案出发,结合微信读书资料与公开点评创作图书带货/书评短视频,并用豆包 TTS、Pexels、Codex 生图和本机 OpenChatCut 完成配音、配图、双语字幕、音效、动效、BGM、可编辑初稿与按需导出。用户提出“根据一本书做带货视频”“读取飞书文案制作图书视频”“写书评口播并自动剪成抖音视频”“仿参考样式做图书推荐短视频”时使用;仅查书、仅写普通书评或无关剪辑… | Kianzzz/ | 218 | — | ~3.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 54 | 54.Video Recap 从输入视频生成中文解说成片或原声剧情短片。用户提供 .mp4 / .mov / .mkv / .webm,并要求剪辑、添加旁白、 配音、总结、短剧/电视剧/电影/纪录片/科普解说时使用。负责编排 video- 技能链:视频理解 → Agent 制定故事与视听方案 → 剪辑 → 配音 → 合成。触发词:视频解说、视频旁白、生成解说、 视频 recap、video… | zenstory-ai/ | 561 | — | ~2.4k | Automated safety check: Pass | MIT | 7 days ago |
| 55 | Orchestrate a configurable, debuggable local pipeline from approved narration packages to timed audio, captions, visual assets, semantic motion, rendered video, optional avatar compositing… | JayceHuang/ | 105 | — | ~1.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 56 | Generates Chinese speech audio from text using Volcengine's Doubao TTS API, for narration, voiceovers or multi-voice podcasts. | xvirobotics/ | 994 | — | ~1.1k | Automated safety check: Notes | MIT | 25 days ago |
| 57 | Makes the Tiny Engineer desk robot speak a short WAV clip through its speaker while it gestures, by posting to its /play endpoint with a bundled script. | jamro/ | 612 | — | ~492 | Automated safety check: Pass | Unknown | today |
| 58 | A skill your agent uses when asked to make an explainer / walkthrough video (讲解视频、解题视频、例题精讲、微课) for a physics problem (物理题: mechanics/力学 受力分析 牛顿定律 斜面 传送带 板块 平抛 圆周 能量 动量, optics/光学 折射… | wy51ai/ | 1.4k | — | ~2.3k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 59 | Create Chinese HBG “模拟人生 / 人生副本” narrative videos with a consistent comic IP, a rapid multi-life opening, continuous natural-speed narration, synchronized short captions, dense static manga… | Mr-funny/ | 141 | — | ~5.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 60 | 60.Showtime A skill your agent uses when the user wants a video made, edited or finished: a launch or promo, product demo, explainer, trailer or teaser, tutorial or walkthrough, a screen recording turned into a… | FavioVazquez/ | 220 | — | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 61 | Prepare, validate, publish, and verify Megaphone releases. An agent skill from Kuberwastaken/megaphone. | Kuberwastaken/ | 170 | — | ~1.3k | Automated safety check: Pass | MIT | 2 mo ago |
| 62 | For creators, educators, and social-video editors who need a tactile paper-collage language for narration, knowledge points, opinions, or abstract topics. | tl2012tl/ | 241 | 4 repos | ~5.2k | Automated safety check: Pass | No licence | 19 days ago |
| 63 | 63.Create Video Tạo một video MỚI cho series "so sánh / phân biệt kiến thức" của repo này — clip dọc TikTok/Reels/Shorts 30-40s, layout 3-zone cố định theo DESIGN.md, voiceover tiếng Việt sinh bằng VieNeu TTS… | Cuongyd196/ | 211 | — | ~2.8k | Automated safety check: Notes | Unknown | 17 days ago |
| 64 | AI educational video production pipeline. An agent skill from speechlab0210/video-production-skill. | speechlab0210/ | 105 | — | ~4.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 65 | Generate smooth hand-drawn whiteboard and story videos directly inside Codex from text scripts, GPT Image 2 color storyboards, scene plans, SVGs, line art, or local images. | gnipbao/ | 327 | — | ~7.2k | Automated safety check: Notes | MIT | 1 mo ago |
| 66 | 66.Explainroo Make an explainer video (MP4) with a voice-over using explainroo. | vincentsch/ | 515 | — | ~439 | Automated safety check: Pass | MIT | 7 days ago |
| 67 | End-to-end AI video production skill for agentic frameworks. | Bomx/ | 310 | — | ~11k | Automated safety check: Notes | No licence | 2 mo ago |
| 68 | Smallest personal narrated-explainer-video pipeline (spoken narration over visuals, not an audio podcast), fully tool-agnostic and autonomous by default — topic → research ∥ asset collection →… | Agents365-ai/ | 1.7k | — | ~3.6k | Automated safety check: Pass | MIT | 9 days ago |
| 69 | AivisSpeech Engine の Sentry issue を調査し、修正すべきエンジン側の不具合と、入力値・ローカル環境・外部サービス由来のノイズを切り分けるためのスキルです。Sentry 側で既知ノイズを永続アーカイブする作業や、voicevoxengine/utility/sentryutility.py と関連テストを更新して既知ノイズを送信前に破棄する作業で使用します。 | Aivis-Project/ | 182 | — | ~545 | Automated safety check: Pass | LGPL-3.0 | today |
| 70 | Transcribes audio/video files using ElevenLabs Scribe v2 API. | qdhenry/ | 1.3k | — | ~1.5k | Automated safety check: Notes | No licence | 7 mo ago |
| 71 | The repeatable workflow for short Thai documentary/explainer videos — one topic in, one finished MP4 out. | killernay/ | 140 | — | ~10k | Automated safety check: Notes | MIT | 2 mo ago |
| 72 | Expert guidance on iOS accessibility best practices, patterns, and implementation. | dadederk/ | 174 | — | ~2.4k | Automated safety check: Pass | MIT | 7 mo ago |
| 73 | Generates target-language dubbing audio from an SRT subtitle file with the KrillinAI CLI, and optionally a dubbed video. | krillinai/ | 13k | — | ~428 | Automated safety check: Pass | Apache-2.0 | today |
| 74 | Generates images, audio and video through Leon's media tools, joins them with FFmpeg, checks the output and attaches playable files for the owner. | leon-ai/ | 18k | — | ~999 | Automated safety check: Pass | MIT | today |
| 75 | Turn ONE topic, talking-head video, or photo into a finished Vox-style paper-collage explainer / ad video on the MuAPI platform (api.muapi.ai) + local ffmpeg — script, collage keyframes, motion… | Anil-matcha/ | 246 | — | ~679 | Automated safety check: Pass | No licence | 1 mo ago |
| 76 | Maintain VoiceOver/TalkBack-focused accessibility in stream-chat-react-native. | GetStream/ | 1.2k | — | ~6.5k | Automated safety check: Pass | Unknown | 2 days ago |
| 77 | 77.Motion Video Produce beat-synced 1080p motion-graphic videos in HyperFrames (HTML + GSAP) with an AI voice-over, Vietnamese karaoke captions, SFX and generated music, in one of two proven styles (glass keynote… | bestagentkits/ | 118 | — | ~1.5k | Automated safety check: Pass | MIT | 12 days ago |
| 78 | 78.Audio Tts Generate speech audio from text using Qwen3 TTS, or clone a voice from reference audio. | second-state/ | 233 | — | ~1.6k | Automated safety check: Pass | No licence | 4 mo ago |
| 79 | 79.Runninghub Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId. | HM-RunningHub/ | 142 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 80 | 80.Kkclaw 给你的 AI Agent 一个桌面身体 — Setup Wizard、14情绪球体、语音克隆、歌词窗、Doctor 自检、跨平台支持(Windows + macOS) | kk43994/ | 174 | — | ~576 | Automated safety check: Pass | MIT | 6 mo ago |
| 81 | MiniMax multimodal model skill — use MiniMax Multi-Modal models for speech, music, video, and image. | poco-ai/ | 1.4k | — | ~7.6k | Automated safety check: Pass | MIT | 19 days ago |
| 82 | 82.Adaptation This skill should be used when the user asks to "make an audiobook", "narration script", "narrator", "ACX", "Findaway", "pronunciation guide", "how long is the audiobook", "adapt to a screenplay"… | danjdewhurst/ | 286 | 1 repo | ~3.5k | Automated safety check: Notes | MIT | 2 days ago |
| 83 | Turns an idea, article, outline or audio file into a sourced, reviewable AI video, tracking whether narration uses a human, synthetic or cloned voice. | wanghui2323/ | 101 | — | ~924 | Automated safety check: Pass | MIT | 1 mo ago |
| 84 | Analyzes a video file with Google Gemini and returns a structured markdown report covering top-level summary, scene-by-scene breakdown, audio transcript (or honest "silent" note), visual details… | mikefutia/ | 101 | — | ~747 | Automated safety check: Notes | No licence | 5 mo ago |
| 85 | Run wally's LLM e2e on Apple Neural Engine (NeuRT) and Snapdragon Hexagon NPU (QHexRT) devices. | RunanywhereAI/ | 1.6k | — | ~845 | Automated safety check: Pass | MIT | today |
| 86 | A skill your agent uses whenever the user wants speech to sound more human, companion-like, or emotionally expressive. | NoizAI/ | 526 | — | ~1.8k | Automated safety check: Pass | No licence | 13 days ago |
| 87 | Builds podcast-style audio narration from text with Azure OpenAI's GPT Realtime Mini over WebSocket, from a Python FastAPI backend to a React player. | microsoft/ | 3.1k | 1 repo | ~947 | Automated safety check: Pass | MIT | yesterday |
| 88 | 合成视频解说最终成片:把旁白音频铺到源视频上,按旁白窗口压低原声,生成 SRT / ASS 字幕并可烧录, 最后做响度标准化。作为最终合成阶段使用。输入源视频、ttsmeta.json 与旁白位置; 输出 recap 成片和字幕。触发词:视频合成、混音、字幕、压字幕、assemble video、mux、ducking、subtitles、成片。 | zenstory-ai/ | 561 | — | ~1.7k | Automated safety check: Pass | MIT | 7 days ago |
| 89 | Generate, convert, clean, and integrate audio for Three.js browser games with ElevenLabs: sound effects, looping ambience, UI sounds, impact/weapon/vehicle audio, creature and boss stingers… | valkor-ai/ | 1.2k | 1 repo | ~1.3k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 90 | Turn arbitrary source content (README, article, story, slides, deck, data/report, product description, tutorial text, audio/transcript, or a bare topic) into a finished, high-quality MP4 video. | architectds/ | 117 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 91 | 91.Media Gen Generate or edit images, videos, or audio in the current task. | clacky-ai/ | 1.2k | — | ~7.5k | Automated safety check: Pass | MIT | today |
| 92 | Create a conversation practice chat MulmoScript with speech bubble UI and character illustration (voiceover approach). | receptron/ | 475 | — | ~3k | Automated safety check: Notes | No licence | today |
| 93 | End-to-end pipeline for producing Vox-style explainer videos from a single topic prompt. | CK42BB/ | 109 | — | ~2.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 94 | Builds math and concept animations with Manim Community Edition, with optional local text-to-speech voiceover and synced subtitles, delivering script.py and video.mp4. | Yusuke710/ | 168 | — | ~1.8k | Automated safety check: Pass | MIT | 4 days ago |
| 95 | 95.Agents Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop. | tadaspetra/ | 296 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 8 days ago |
| 96 | Look up HyperTTS crash reports in Sentry — the project is language-tools/anki-hyper-tts, project ID 6170140. | Vocab-Apps/ | 284 | — | ~1.9k | Automated safety check: Pass | GPL-3.0 | 28 days ago |