Agent skill

Fal AI Media

by affaan-m in affaan-m/ECC

通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。

MITAuto-check passedMedia & Creative

Install Fal AI Media

skills CLI
$ npx skills add affaan-m/ECC --skill fal-ai-media -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install affaan-m/ECC fal-ai-media --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/affaan-m/ECC.git skills-src && mkdir -p .claude/skills && cp -r skills-src/docs/zh-CN/skills/fal-ai-media .claude/skills/fal-ai-media && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
fal-ai-media
GitHub stars
277k
Used in
2 other repos
Token cost
~1.2k tokens
SKILL.md length
168 words
Files
1
Skills in repo
683
Repo updated
First seen
Licence
MIT

At a glance

通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。

  • Tasks that involve AI video generation
  • SKILL.md covers 何时激活, MCP 要求, MCP 工具 and 图像生成, plus 6 more sections
  • Reaches api.elevenlabs.io; needs FAL_KEY and ELEVENLABS_API_KEY
  • Tasks that involve Image generation

What it does

Fal AI Media is an agent skill from affaan-m/ECC. 通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering AI video generation and Image generation. It works with fal, Model Context Protocol, Google Gemini and Google Veo. The repository describes itself as: The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond. The licence is MIT.

When your agent uses it

  • Tasks that involve AI video generation
  • Tasks that involve Image generation

Example prompts

  • “/fal-ai-media”

Requirements

  • Python 3
  • A credential in FAL_KEY
  • A credential in ELEVENLABS_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 2d515e4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are python and json).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.elevenlabs.io

    Also links to:

    • fal.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • FAL_KEY
    • ELEVENLABS_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Fal AI Media loads about 1.2k tokens when it runs. Until then it costs about 40 tokens; SKILL.md has 168 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~40
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from affaan-m/ECC at commit 2d515e4, republished under its MIT licence (© affaan-m). 168 words, ~1,242 tokens.

Download SKILL.mdSave it as .claude/skills/fal-ai-media/SKILL.md (or your agent's skills folder).
name
fal-ai-media
description
通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。
origin
ECC

fal.ai 媒体生成

通过 MCP 使用 fal.ai 模型生成图像、视频和音频。

何时激活

  • 用户希望根据文本提示生成图像
  • 根据文本或图像创建视频
  • 生成语音、音乐或音效
  • 任何媒体生成任务
  • 用户提及“生成图像”、“创建视频”、“文本转语音”、“制作缩略图”或类似表述

MCP 要求

必须配置 fal.ai MCP 服务器。添加到 ~/.claude.json:

json
"fal-ai": {
  "command": "npx",
  "args": ["-y", "fal-ai-mcp-server"],
  "env": { "FAL_KEY": "YOUR_FAL_KEY_HERE" }
}

在 fal.ai 获取 API 密钥。

MCP 工具

fal.ai MCP 提供以下工具:

  • search — 通过关键词查找可用模型
  • find — 获取模型详情和参数
  • generate — 使用参数运行模型
  • result — 检查异步生成状态
  • status — 检查作业状态
  • cancel — 取消正在运行的作业
  • estimate_cost — 估算生成成本
  • models — 列出热门模型
  • upload — 上传文件用作输入

图像生成

Nano Banana 2(快速)

最适合:快速迭代、草稿、文生图、图像编辑。

generate(
  app_id: "fal-ai/nano-banana-2",
  input_data: {
    "prompt": "未来主义日落城市景观,赛博朋克风格",
    "image_size": "landscape_16_9",
    "num_images": 1,
    "seed": 42
  }
)
Nano Banana Pro(高保真)

最适合:生产级图像、写实感、排版、详细提示。

generate(
  app_id: "fal-ai/nano-banana-pro",
  input_data: {
    "prompt": "专业产品照片,无线耳机置于大理石表面,影棚灯光",
    "image_size": "square",
    "num_images": 1,
    "guidance_scale": 7.5
  }
)
常见图像参数
参数类型选项说明
prompt字符串必需描述您想要的内容
image_size字符串square、portrait_4_3、landscape_16_9、portrait_16_9、landscape_4_3宽高比
num_images数字1-4生成数量
seed数字任意整数可重现性
guidance_scale数字1-20遵循提示的紧密程度(值越高越贴近字面)
图像编辑

使用 Nano Banana 2 并输入图像进行修复、扩展或风格迁移:

# 首先上传源图像
upload(file_path: "/path/to/image.png")

# 然后使用图像输入进行生成
generate(
  app_id: "fal-ai/nano-banana-2",
  input_data: {
    "prompt": "same scene but in watercolor style",
    "image_url": "<uploaded_url>",
    "image_size": "landscape_16_9"
  }
)

视频生成

Seedance 1.0 Pro(字节跳动)

最适合:文生视频、图生视频,具有高运动质量。

generate(
  app_id: "fal-ai/seedance-1-0-pro",
  input_data: {
    "prompt": "a drone flyover of a mountain lake at golden hour, cinematic",
    "duration": "5s",
    "aspect_ratio": "16:9",
    "seed": 42
  }
)
Kling Video v3 Pro

最适合:文生/图生视频,带原生音频生成。

generate(
  app_id: "fal-ai/kling-video/v3/pro",
  input_data: {
    "prompt": "海浪拍打着岩石海岸,乌云密布",
    "duration": "5s",
    "aspect_ratio": "16:9"
  }
)
Veo 3(Google DeepMind)

最适合:带生成声音的视频,高视觉质量。

generate(
  app_id: "fal-ai/veo-3",
  input_data: {
    "prompt": "夜晚熙熙攘攘的东京街头市场,霓虹灯招牌,人群喧嚣",
    "aspect_ratio": "16:9"
  }
)
图生视频

从现有图像开始:

generate(
  app_id: "fal-ai/seedance-1-0-pro",
  input_data: {
    "prompt": "camera slowly zooms out, gentle wind moves the trees",
    "image_url": "<uploaded_image_url>",
    "duration": "5s"
  }
)
视频参数
参数类型选项说明
prompt字符串必需描述视频内容
duration字符串"5s"、"10s"视频长度
aspect_ratio字符串"16:9"、"9:16"、"1:1"帧比例
seed数字任意整数可重现性
image_url字符串URL用于图生视频的源图像

音频生成

CSM-1B(对话语音)

文本转语音,具有自然、对话式的音质。

generate(
  app_id: "fal-ai/csm-1b",
  input_data: {
    "text": "Hello, welcome to the demo. Let me show you how this works.",
    "speaker_id": 0
  }
)
ThinkSound(视频转音频)

根据视频内容生成匹配的音频。

generate(
  app_id: "fal-ai/thinksound",
  input_data: {
    "video_url": "<video_url>",
    "prompt": "ambient forest sounds with birds chirping"
  }
)
ElevenLabs(通过 API,无 MCP)

如需专业的语音合成,直接使用 ElevenLabs:

python
import os
import requests

resp = requests.post(
    "https://api.elevenlabs.io/v1/text-to-speech/<voice_id>",
    headers={
        "xi-api-key": os.environ["ELEVENLABS_API_KEY"],
        "Content-Type": "application/json"
    },
    json={
        "text": "Your text here",
        "model_id": "eleven_turbo_v2_5",
        "voice_settings": {"stability": 0.5, "similarity_boost": 0.75}
    }
)
with open("output.mp3", "wb") as f:
    f.write(resp.content)
VideoDB 生成式音频

如果配置了 VideoDB,使用其生成式音频:

python
# Voice generation
audio = coll.generate_voice(text="Your narration here", voice="alloy")

# Music generation
music = coll.generate_music(prompt="upbeat electronic background music", duration=30)

# Sound effects
sfx = coll.generate_sound_effect(prompt="thunder crack followed by rain")

成本估算

生成前,检查估算成本:

estimate_cost(
  estimate_type: "unit_price",
  endpoints: {
    "fal-ai/nano-banana-pro": {
      "unit_quantity": 1
    }
  }
)

模型发现

查找特定任务的模型:

search(query: "text to video")
find(endpoint_ids: ["fal-ai/seedance-1-0-pro"])
models()

提示

  • 在迭代提示时,使用 seed 以获得可重现的结果
  • 先用低成本模型(Nano Banana 2)进行提示迭代,然后切换到 Pro 版进行最终生成
  • 对于视频,保持提示描述性但简洁——聚焦于运动和场景
  • 图生视频比纯文生视频能产生更可控的结果
  • 在运行昂贵的视频生成前,检查 estimate_cost

相关技能

  • videodb — 视频处理、编辑和流媒体
  • video-editing — AI 驱动的视频编辑工作流
  • content-engine — 社交媒体平台内容创作

© affaan-m, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in docs/zh-CN/skills/fal-ai-media of affaan-m/ECC.

Open the folder on GitHubat commit 2d515e4

Used in 2 other repositories

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in affaan-m/ECC, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Fal AI Media next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Fal AI Media compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Fal AI Media this skillaffaan-m/ECC277k2 repos~1.2kAutomated safety check: PassMIT
Forge Media Route Layer0x0funky/agent-sprite-forge4.4k—~2.2kAutomated safety check: PassMIT
Higgsfield ModelsOSideMedia/higgsfield-ai-prompt-skill713—~7kAutomated safety check: PassMIT
Seedance Storyboard Generatorliangdabiao/Seedance2-Storyboard-Generator2.6k—~2.2kAutomated safety check: PassNone
Imagesmixs/visual-skills494—~2.1kAutomated safety check: PassCC-BY-4.0
HiggsfieldOSideMedia/higgsfield-ai-prompt-skill713—~9.1kAutomated safety check: PassMIT

Similar skills

  • Forge Media Route Layer

    0x0funky/agent-sprite-forge

    Generates an image or an image-to-video clip through a configured provider API or a signed-in Codex or Grok CLI, and reports the route, file, hash and cost estimate.

    4.4k GitHub stars~2.2k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Higgsfield Models

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses when the user asks which model to use, wants to compare models, or needs guidance on selecting between Kling, Wan (incl.

    713 GitHub stars~7k tokensUpdated 14 days ago
    Media & CreativeAuto-check passed
  • Seedance Storyboard Generator

    liangdabiao/Seedance2-Storyboard-Generator

    专业的Seedance 2.0平台AI视频脚本和分镜生成器。当用户要求:(1) 将文章/故事转换为视频脚本,(2) 生成Seedance 2.0分镜提示词,(3) 规划多集AI视频系列,(4) 为GPT-Image-2、Seedream、Nano Banana…

    2.6k GitHub stars~2.2k tokensUpdated 20 days ago
    Media & CreativeAuto-check passed
  • Image

    smixs/visual-skills

    Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst).

    494 GitHub stars~2.1k tokensUpdated 25 days ago
    Media & CreativeAuto-check passed
  • Higgsfield

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano…

    713 GitHub stars~9.1k tokensUpdated 14 days ago
    Media & CreativeAuto-check passed
  • Nbcraft

    jieyefriic/nbcraft

    Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft.

    155 GitHub stars~2.8k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed

More from affaan-m/ECC

All 682 skills in this repo
  • Skill Stocktake

    affaan-m/ECC

    Audits your installed Claude skills and commands for quality, with a quick mode for recently changed skills and a full mode that evaluates all of them through subagents.

    277k GitHub starsUsed in 5 repos~3.1k tokens
    Auto-check passed
  • Ingests, indexes, searches, edits and monitors video, audio and live streams through the VideoDB Python SDK, returning stream links, clips and timestamps.

    277k GitHub starsUsed in 3 repos~3.5k tokens
    Auto-check: notes
  • Docs Governance

    affaan-m/ECC

    Route broad documentation-governance requests to existing ECC skills and run an opt-in, read-only audit of mapped documentation roles, links, ADR indexes, and evidence references.

    277k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Rules Distillation

    affaan-m/ECC

    Scans installed skills for principles that recur across them and proposes rule-file changes: append, revise, add a section, create a file or leave as covered.

    277k GitHub starsUsed in 2 repos~2.3k tokens
    Auto-check passed
  • Builds DRAFT counterparty agreements from one markdown template and a small JSON spec per party, with clauses picked by the party's role.

    277k GitHub stars~2.9k tokensUpdated today
    Auto-check passed
  • Set an ECC-specific frontend design direction for production UI work.

    277k GitHub starsUsed in 1 repo~2.2k tokens
    Auto-check passed

Questions about Fal AI Media

What does Fal AI Media do?

通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。. Fal AI Media is an agent skill from affaan-m/ECC.

When should I use Fal AI Media?

Fal AI Media fits situations like: tasks that involve AI video generation; tasks that involve Image generation.

How do I install Fal AI Media in Claude Code?

Run `npx skills add affaan-m/ECC --skill fal-ai-media -a claude-code`. Or copy the skill folder (docs/zh-CN/skills/fal-ai-media in affaan-m/ECC) into .claude/skills/fal-ai-media in your project. Claude Code loads it when a task matches its description.

How do I install Fal AI Media in Codex?

Run `npx skills add affaan-m/ECC --skill fal-ai-media -a codex`. Or copy the skill folder (docs/zh-CN/skills/fal-ai-media in affaan-m/ECC) into .agents/skills/fal-ai-media in your project. Codex loads it when a task matches its description.

Can I use Fal AI Media in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add affaan-m/ECC --skill fal-ai-media -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/fal-ai-media, .gemini/skills/fal-ai-media, .github/skills/fal-ai-media and .opencode/skills/fal-ai-media in your project.

What does Fal AI Media need to run?

Going by SKILL.md and its folder, Fal AI Media needs credentials named FAL_KEY and ELEVENLABS_API_KEY. Our summary lists: Python 3; A credential in FAL_KEY; A credential in ELEVENLABS_API_KEY.

Does Fal AI Media access the network?

SKILL.md names 2 domains. In commands or code: api.elevenlabs.io; the agent is likely to contact it when it follows the instructions. As links in the text: fal.ai. This is read from the text; nothing was executed.

Is Fal AI Media safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Fal AI Media use?

Fal AI Media is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Fal AI Media use?

About 1.2k tokens (SKILL.md is roughly 5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Fal AI Media?

Skills that share tags, products or a category with Fal AI Media: Forge Media Route Layer (0x0funky/agent-sprite-forge, 4.4k stars), Higgsfield Models (OSideMedia/higgsfield-ai-prompt-skill, 713 stars), Seedance Storyboard Generator (liangdabiao/Seedance2-Storyboard-Generator, 2.6k stars) and Image (smixs/visual-skills, 494 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Fal AI Media?

affaan-m (a GitHub user) maintains it in affaan-m/ECC, which has 276,673 GitHub stars. The repository holds 683 skills in this directory. The repository was last updated on October 11, 2026.

Source: affaan-m/ECC on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.