何时激活
用户希望根据文本提示生成图像 根据文本或图像创建视频 生成语音、音乐或音效 任何媒体生成任务 用户提及“生成图像”、“创建视频”、“文本转语音”、“制作缩略图”或类似表述
affaan-m/ECC
Use it for engineering tasks; the detail page covers purpose, installation, and practical steps.
npx skills add https://github.com/affaan-m/ECC --skill "docs/zh-CN/skills/fal-ai-media"Source checked Jul 28, 2026·Refresh due Oct 26, 2026
Reorganized from the pinned upstream SKILL.md
通过 MCP 使用 fal.ai 模型生成图像、视频和音频。
npx skills add https://github.com/affaan-m/ECC --skill "docs/zh-CN/skills/fal-ai-media"The pinned source supports a structured brief, but not an expanded tutorial. Only detected inputs, outputs, and sections are shown.
352 source words · 28 usable sections
Documentation workflow
Sections are extracted automatically from the pinned SKILL.md and link back to the source.
用户希望根据文本提示生成图像 根据文本或图像创建视频 生成语音、音乐或音效 任何媒体生成任务 用户提及“生成图像”、“创建视频”、“文本转语音”、“制作缩略图”或类似表述
必须配置 fal.ai MCP 服务器。添加到 /.claude.json:
search — 通过关键词查找可用模型 find — 获取模型详情和参数 generate — 使用参数运行模型 result — 检查异步生成状态 status — 检查作业状态 cancel — 取消正在运行的作业 estimatecost — 估算生成成本 models — 列出热门模型 upload — 上传文件用作输入
使用 Nano Banana 2 并输入图像进行修复、扩展或风格迁移:
Documentation checklist
The source section “何时激活” has been checked.
The source section “MCP 要求” has been checked.
The source section “MCP 工具” has been checked.
The source section “图像生成” has been checked.
Static permission evidence
These are source excerpts matched by deterministic rules, not findings of malicious behavior, safety, or actual execution.
SKILL.md · L220
resp = requests.post(The documentation includes sending, uploading, or posting data to a remote service.
SKILL.md · L220
resp = requests.post(The documentation includes network, browsing, or remote request actions.
SKILL.md · L221
"https://api.elevenlabs.io/v1/text-to-speech/<voice_id>",The documentation includes network, browsing, or remote request actions.
Choose a different workflow
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-to-speech (CSM-1B), and video-to-audio (ThinkSound). Use when the user wants to generate images, videos, or audio with AI.
A separate implementation from affaan-m/ECC; compare its source, maintenance signals, and permission requirements.
Open source detailUnified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-to-speech (CSM-1B), and video-to-audio (ThinkSound). Use when the user wants to generate images, videos, or audio with AI.
A separate implementation from affaan-m/ECC; compare its source, maintenance signals, and permission requirements.
Open source detailUse it for engineering tasks; the detail page covers purpose, installation, and practical steps.
A separate implementation from affaan-m/ECC; compare its source, maintenance signals, and permission requirements.
Open source detailFAQ
通过 MCP 使用 fal.ai 模型生成图像、视频和音频。
The source record exposes this install command: npx skills add https://github.com/affaan-m/ECC --skill "docs/zh-CN/skills/fal-ai-media". Inspect the command and pinned source before running it.
Static rules flagged send-data, network in the source; the page lists the matching lines and excerpts.
Quality breakdown
Based on traceable docs and repository signals; stars are not treated as quality.
Compare before choosing
These links are selected from shared tasks, functions, stacks, platforms, and same-name variants. Compare the source owner, documentation, permissions, and maintenance signals.
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-to-speech (CSM-1B), and video-to-audio (ThinkSound). Use when the user wants to generate images, videos, or audio with AI.
Unified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video (Seedance, Kling, Veo 3), text-to-speech (CSM-1B), and video-to-audio (ThinkSound). Use when the user wants to generate images, videos, or audio with AI.
Use it for engineering tasks; the detail page covers purpose, installation, and practical steps.
When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "should I test this," "which version is better," "test two versions," "statistical significance," "how long should I run this test," "growth experiments," "experiment velocity," "experiment backlog," "ICE score," "experimentation program
Grounded design brief from the adopted corpus — style, WCAG-checked color tokens, typography, layout pattern, anti-patterns. Use on ui-design-brief or any which-style/palette/font/chart decision.
通过 MCP 使用 fal.ai 模型生成图像、视频和音频。
必须配置 fal.ai MCP 服务器。添加到 ~/.claude.json:
"fal-ai": {
"command": "npx",
"args": ["-y", "fal-ai-mcp-server"],
"env": { "FAL_KEY": "YOUR_FAL_KEY_HERE" }
}
在 fal.ai 获取 API 密钥。
fal.ai MCP 提供以下工具:
search — 通过关键词查找可用模型find — 获取模型详情和参数generate — 使用参数运行模型result — 检查异步生成状态status — 检查作业状态cancel — 取消正在运行的作业estimate_cost — 估算生成成本models — 列出热门模型upload — 上传文件用作输入最适合:快速迭代、草稿、文生图、图像编辑。
generate(
app_id: "fal-ai/nano-banana-2",
input_data: {
"prompt": "未来主义日落城市景观,赛博朋克风格",
"image_size": "landscape_16_9",
"num_images": 1,
"seed": 42
}
)
最适合:生产级图像、写实感、排版、详细提示。
generate(
app_id: "fal-ai/nano-banana-pro",
input_data: {
"prompt": "专业产品照片,无线耳机置于大理石表面,影棚灯光",
"image_size": "square",
"num_images": 1,
"guidance_scale": 7.5
}
)
| 参数 | 类型 | 选项 | 说明 |
|---|---|---|---|
prompt | 字符串 | 必需 | 描述您想要的内容 |
image_size | 字符串 | square、portrait_4_3、landscape_16_9、portrait_16_9、landscape_4_3 | 宽高比 |
num_images | 数字 | 1-4 | 生成数量 |
seed | 数字 | 任意整数 | 可重现性 |
guidance_scale | 数字 | 1-20 | 遵循提示的紧密程度(值越高越贴近字面) |
使用 Nano Banana 2 并输入图像进行修复、扩展或风格迁移:
# 首先上传源图像
upload(file_path: "/path/to/image.png")
# 然后使用图像输入进行生成
generate(
app_id: "fal-ai/nano-banana-2",
input_data: {
"prompt": "same scene but in watercolor style",
"image_url": "<uploaded_url>",
"image_size": "landscape_16_9"
}
)
最适合:文生视频、图生视频,具有高运动质量。
generate(
app_id: "fal-ai/seedance-1-0-pro",
input_data: {
"prompt": "a drone flyover of a mountain lake at golden hour, cinematic",
"duration": "5s",
"aspect_ratio": "16:9",
"seed": 42
}
)
最适合:文生/图生视频,带原生音频生成。
generate(
app_id: "fal-ai/kling-video/v3/pro",
input_data: {
"prompt": "海浪拍打着岩石海岸,乌云密布",
"duration": "5s",
"aspect_ratio": "16:9"
}
)
最适合:带生成声音的视频,高视觉质量。
generate(
app_id: "fal-ai/veo-3",
input_data: {
"prompt": "夜晚熙熙攘攘的东京街头市场,霓虹灯招牌,人群喧嚣",
"aspect_ratio": "16:9"
}
)
从现有图像开始:
generate(
app_id: "fal-ai/seedance-1-0-pro",
input_data: {
"prompt": "camera slowly zooms out, gentle wind moves the trees",
"image_url": "<uploaded_image_url>",
"duration": "5s"
}
)
| 参数 | 类型 | 选项 | 说明 |
|---|---|---|---|
prompt | 字符串 | 必需 | 描述视频内容 |
duration | 字符串 | "5s"、"10s" | 视频长度 |
aspect_ratio | 字符串 | "16:9"、"9:16"、"1:1" | 帧比例 |
seed | 数字 | 任意整数 | 可重现性 |
image_url | 字符串 | URL | 用于图生视频的源图像 |
文本转语音,具有自然、对话式的音质。
generate(
app_id: "fal-ai/csm-1b",
input_data: {
"text": "Hello, welcome to the demo. Let me show you how this works.",
"speaker_id": 0
}
)
根据视频内容生成匹配的音频。
generate(
app_id: "fal-ai/thinksound",
input_data: {
"video_url": "<video_url>",
"prompt": "ambient forest sounds with birds chirping"
}
)
如需专业的语音合成,直接使用 ElevenLabs:
import os
import requests
resp = requests.post(
"https://api.elevenlabs.io/v1/text-to-speech/<voice_id>",
headers={
"xi-api-key": os.environ["ELEVENLABS_API_KEY"],
"Content-Type": "application/json"
},
json={
"text": "Your text here",
"model_id": "eleven_turbo_v2_5",
"voice_settings": {"stability": 0.5, "similarity_boost": 0.75}
}
)
with open("output.mp3", "wb") as f:
f.write(resp.content)
如果配置了 VideoDB,使用其生成式音频:
# Voice generation
audio = coll.generate_voice(text="Your narration here", voice="alloy")
# Music generation
music = coll.generate_music(prompt="upbeat electronic background music", duration=30)
# Sound effects
sfx = coll.generate_sound_effect(prompt="thunder crack followed by rain")
生成前,检查估算成本:
estimate_cost(
estimate_type: "unit_price",
endpoints: {
"fal-ai/nano-banana-pro": {
"unit_quantity": 1
}
}
)
查找特定任务的模型:
search(query: "text to video")
find(endpoint_ids: ["fal-ai/seedance-1-0-pro"])
models()
seed 以获得可重现的结果estimate_costvideodb — 视频处理、编辑和流媒体video-editing — AI 驱动的视频编辑工作流content-engine — 社交媒体平台内容创作