Local AI Use
amd/skills
Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.
Generate narration with the public Fish Audio REST API using environment credentials and a user-selected voice.
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-api --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/fish-audio-api .claude/skills/fish-audio-api && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .claude/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-apiType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-api --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/fish-audio-api .agents/skills/fish-audio-api && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .agents/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-api --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/fish-audio-api .cursor/skills/fish-audio-api && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .cursor/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/waker240/FullVideoProductionSkill.git --path skills/fish-audio-api--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-api --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/fish-audio-api .gemini/skills/fish-audio-api && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .gemini/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-apiInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/fish-audio-api .github/skills/fish-audio-api && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .github/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install waker240/FullVideoProductionSkill fish-audio-api --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/waker240/FullVideoProductionSkill.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/fish-audio-api .opencode/skills/fish-audio-api && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "fish-audio-api" agent skill from https://github.com/waker240/FullVideoProductionSkill/tree/main/skills/fish-audio-api into .opencode/skills/fish-audio-api/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "fish-audio-api", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
fish-audio-apiGenerate narration with the public Fish Audio REST API using environment credentials and a user-selected voice.
Fish Audio API is an agent skill from waker240/FullVideoProductionSkill. Generate narration with the public Fish Audio REST API using environment credentials and a user-selected voice. Provides shared public API helpers for Fish narration, OpenAI word timestamps, and optional GPT Image generation; no browser cookies or private proxy.
Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including scripts.
It sits in Media & Creative, covering Text to speech and voice, Image generation and REST APIs. It works with OpenAI. The repository describes itself as: 20亿Claude Opus 4.6-4.8的Token. The licence is Apache-2.0.
Read from SKILL.md and the folder at commit 0223baa. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (JavaScript), which the agent can run.
Shell commands in SKILL.md call:
nodeFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
api.fish.audioAlso links to:
docs.fish.audioFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
FISH_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Fish Audio API loads about 1.1k tokens when it runs. Until then it costs about 69 tokens; SKILL.md has 468 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
env.example) to the **video project's** `.env`, ensure `.env` is ignored by Git, then fill it locally. Never request thaAutomated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from waker240/FullVideoProductionSkill at commit 0223baa, republished under its Apache-2.0 licence (© waker240). 468 words, ~1,075 tokens.
.claude/skills/fish-audio-api/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.Use this skill when a video needs Fish narration. The release uses the official POST https://api.fish.audio/v1/tts endpoint. It does not ship a web-session proxy, cookies, an account, or a voice. Node.js 22.20 or later is required. Install FFmpeg/ffprobe for the video audio pipeline.
Copy env.example to the video project's .env, ensure .env is ignored by Git, then fill it locally. Never request that users paste credentials into chat. Existing process environment values take precedence over .env. Scripts never print secret values or upstream error bodies.
| Variable | Meaning |
|---|---|
FISH_API_KEY | API key created in the user's Fish account; required for generation. |
FISH_REFERENCE_ID | A voice ID the user has permission to use; required unless a project explicitly supplies one. |
FISH_MODEL | Defaults to s2.1-pro-free; supported choices also include s1, s2-pro, s2.1-pro. Account access and quotas are provider-controlled. |
No silent model fallback is allowed. An unknown model fails locally, because the provider may otherwise choose a paid default. Availability and pricing can change; check the account before generation. Setting a key does not itself authorize a charge: run synthesis only for the user's requested scope.
For a scaffolded HyperFrames project, edit scripts/narration.json. Its voice.referenceId and voice.model may be null to use the environment. All other voice settings are explicit in the template. Preview one short paragraph before producing a whole narration.
node scripts/tts-fish.mjs --section s0
node scripts/tts-fish.mjs
node scripts/tts-fish.mjs --section s2 --para 1 --force--para requires every unselected paragraph in that section to exist. To start a new project, first generate one short section with --section s0; use --para for a later repair. Matching paragraph text and voice profiles reuse frozen audio. --force makes another API request for the selected speech.
The common helper public-media-api.cjs supports the shared audio engine as well. Its Fish configuration accepts fish.transport: "official-api", fish.reference_id, fish.model, and fish.request. Sampling fields are top-level request fields, such as temperature and top_p; a website-style backend/sampler profile is not this API contract. Never put credentials in JSON project profiles.
Every generation POST is sent once. A timeout can occur after the provider started work; inspect provider usage before deliberately retrying. Redirects are refused. Official endpoints are pinned. FISH_API_URL with HYPERFRAMES_TEST_ALLOW_FISH_LOOPBACK=1 is only for local mock tests.
The same helper supports OpenAI whisper-1 transcription with word timestamps, and GPT Image PNG generation. These are separate account/billing capabilities, not features bundled by an agent subscription. See requirements and transcription. Images require an explicit OPENAI_IMAGE_MODEL or --model; choose a GPT Image model available to the user's account. External image/video tools, including Codex image tools and video-generation plugins, are optional and are not installed or authenticated by this skill package.
node --test skills/fish-audio-api/tests/public-media-api.test.cjs
node --test skills/hyperframes/scripts/tests/fish-tts.test.mjsThese tests use fake keys and loopback HTTP, not paid services. The release verification does not establish that a user's account has model access.
Official references checked on 2026-09-26: Fish TTS API, Fish developer guide.
© waker240, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 4 other files (scripts) in skills/fish-audio-api of waker240/FullVideoProductionSkill.
Open the folder on GitHubat commit 0223baa
Fish Audio API next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Fish Audio API this skillwaker240/FullVideoProductionSkill | 197 | — | ~1.1k | Automated safety check: Notes | Apache-2.0 | |
| Local AI Useamd/skills | 398 | — | ~5k | Automated safety check: Notes | MIT | |
| Agentselevenlabs/skills | 481 | — | ~6.5k | Automated safety check: Pass | MIT | |
| Local AI App Integrationamd/skills | 398 | — | ~6k | Automated safety check: Pass | MIT | |
| Codex Imagendarkamenosa/codex-imagen | 135 | — | ~2.6k | Automated safety check: Pass | MIT | |
| Chatgpt CLIItamarZand88/CLI-Anything-WEB | 231 | — | ~781 | Automated safety check: Pass | MIT |
amd/skills
Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.
elevenlabs/skills
Build voice AI agents with ElevenLabs. An agent skill from elevenlabs/skills.
amd/skills
Integrates local AI capabilities into applications using Embeddable Lemonade.
darkamenosa/codex-imagen
Generate or edit raster images by calling the ChatGPT/Codex hosted imagegeneration flow with local Codex or OpenClaw OAuth credentials, then save decoded image files for OpenClaw and other agent…
ItamarZand88/CLI-Anything-WEB
Drives ChatGPT from the terminal via cli-web-chatgpt — ask questions, generate and download images, list and view conversations, browse models, and manage OpenAI SSO auth.
ninehills/skills
生图 / 生成图片 / 画图 — 用 OpenAI gpt-image-2 生成图像。支持文生图、参考图生图 (img2img)、蒙版修补 (inpainting)。当用户要求用 GPT 画图、OpenAI 生图、gpt-image-2、文+图生图、参考图片生成、img2img、inpainting 时必加载此技能。Auth 自动继承 OPENAIAPIKEY / Codex OAuth…
waker240/FullVideoProductionSkill
Build deterministic animation for HyperFrames using motion rules, transitions, scene blueprints, and runtime adapters.
waker240/FullVideoProductionSkill
Develop video direction, typography, color, visual mechanisms, storyboards, and asset prompts.
waker240/FullVideoProductionSkill
Resolve reviewed local BGM, sound effects, images, icons, and brand assets into frozen project files plus a manifest.
waker240/FullVideoProductionSkill
Build a HyperFrames video around a supplied music track, using measured rhythm and editable visual concepts.
waker240/FullVideoProductionSkill
Port an existing Remotion composition to HyperFrames HTML while preserving visible behavior.
waker240/FullVideoProductionSkill
Prepare Fish official-API narration, OpenAI word timestamps, and reviewed user-supplied local music/SFX for HyperFrames.
Works with
Categories
Generate narration with the public Fish Audio REST API using environment credentials and a user-selected voice. Fish Audio API is an agent skill from waker240/FullVideoProductionSkill. Generate narration with the public Fish Audio REST API using environment credentials and a user-selected voice.
Fish Audio API fits situations like: tasks that involve Text to speech and voice; tasks that involve Image generation; tasks that involve REST APIs.
Run `npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a claude-code`. Or copy the skill folder (skills/fish-audio-api in waker240/FullVideoProductionSkill) into .claude/skills/fish-audio-api in your project. Claude Code loads it when a task matches its description.
Run `npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a codex`. Or copy the skill folder (skills/fish-audio-api in waker240/FullVideoProductionSkill) into .agents/skills/fish-audio-api in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add waker240/FullVideoProductionSkill --skill fish-audio-api -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/fish-audio-api, .gemini/skills/fish-audio-api, .github/skills/fish-audio-api and .opencode/skills/fish-audio-api in your project.
Going by SKILL.md and its folder, Fish Audio API needs JavaScript for the scripts in its folder, the command-line tools its instructions call (node) and credentials named FISH_API_KEY. Our summary lists: Node.js; A credential in FISH_API_KEY.
SKILL.md names 2 domains. In commands or code: api.fish.audio; the agent is likely to contact it when it follows the instructions. As links in the text: docs.fish.audio. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Fish Audio API is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.1k tokens (SKILL.md is roughly 4.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Fish Audio API: Local AI Use (amd/skills, 398 stars), Agents (elevenlabs/skills, 481 stars), Local AI App Integration (amd/skills, 398 stars) and Codex Imagen (darkamenosa/codex-imagen, 135 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
waker240 (a GitHub user) maintains it in waker240/FullVideoProductionSkill, which has 197 GitHub stars. The repository holds 13 skills in this directory. The repository was last updated on September 26, 2026.
Source: waker240/FullVideoProductionSkill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.