$npx skillfedfor your agent

fal-ai-media

Create images, videos, and audio content using fal.ai's suite of generative models through an MCP interface. The skill supports text-to-image with Nano Banana, video generation from text or images, and speech synthesis, with tools for model discovery, cost estimation, and job management.

fal-ai-media generates images, videos, and audio using AI models via the fal.ai platform.

AI-generated summary based on this skill's SKILL.md

234,207 35,692 MITupdated by affaan-m

Decision gist · record as of 2026-07-27

fal-ai-media generates images, videos, and audio using AI models via the fal.ai platform. Create images, videos, and audio content using fal.ai's suite of generative models through an MCP interface. The skill supports text-to-image with Nano Banana, video generation from text or images, and speech synthesis, with tools for model discovery, cost estimation, and job management.

manual: git clone https://github.com/affaan-m/ECC → cp -r ECC/skills/fal-ai-media ~/.claude/skills/fal-ai-media
skills/fal-ai-media/SKILL.md · version 5fc39802

Use it when

  • fal-ai-media supports video generation from both text prompts and existing images.
  • Yes, fal-ai-media includes speech synthesis capabilities for text-to-speech conversion.

Verify before relying

Read SKILL.md below before installing (1 file). Open directory: indexed for reading, not audited.

Same gist for agents: .md · .json

Install

affaan-m/ECC/fal-ai-media · repository language: JavaScript

Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

Can fal-ai-media generate image from text?

Yes, fal-ai-media generates images from text descriptions using AI models like Nano Banana. You provide a text prompt describing what you want, and the skill creates the corresponding image through fal.ai's generative models.

How do I create video with AI using fal-ai-media?

fal-ai-media supports video generation from both text prompts and existing images. You can describe a video scene in text or provide an image, and the skill uses fal.ai's video generation models to create the video content for you.

Does fal-ai-media support text to speech conversion?

Yes, fal-ai-media includes speech synthesis capabilities for text-to-speech conversion. You can convert written text into audio content using the skill's audio generation tools powered by fal.ai's models.

What models are available in fal-ai-media?

fal-ai-media provides access to fal.ai's suite of generative models including Nano Banana for image synthesis, video generation models, and speech synthesis tools. The skill includes model discovery features so you can explore and compare available options.

Can I estimate costs before generating media with fal-ai-media?

Yes, fal-ai-media includes cost estimation tools. You can check the estimated cost of your media generation jobs before running them, helping you manage expenses when using fal.ai's generative models.

What license does fal-ai-media use?

fal-ai-media is released under the MIT license, making it freely available for use, modification, and distribution in both open-source and commercial projects.

SKILL.md

Rendered from the published skill. Quoted content, verbatim.

fal.ai Media Generation

> Drift-prone skill. fal.ai model IDs, pricing, inputs, and MCP tool names > change quickly. Search or fetch the current model metadata before promising a > specific model, parameter, output format, or cost.

Generate images, videos, and audio using fal.ai models via MCP.

When to Activate

  • User wants to generate images from text prompts
  • Creating videos from text or images
  • Generating speech, music, or sound effects
  • Any media generation task
  • User says "generate image", "create video", "text to speech", "make a thumbnail", or similar

MCP Requirement

fal.ai MCP server must be configured. Add to ~/.claude.json:

"fal-ai": {
  "command": "npx",
  "args": ["-y", "fal-ai-mcp-server"],
  "env": { "FAL_KEY": "YOUR_FAL_KEY_HERE" }
}

Get an API key at fal.ai.

MCP

(truncated - see the full file via the links below)

File tree — 1 file
skills/fal-ai-media/SKILL.md

Let your AI agent find skills like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.

wish › “Generate images, videos, or audio using AI models”

Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →

Related skills

Fal Ai
by hoodini · hoodini/ai-agents-skills

Fal Ai lets you run machine learning inference on serverless infrastructure, supporting image generation with Flux and SDXL, video creation, audio processing, and real-time streaming. Deploy models without managing servers, with built-in support for editing, upscaling, and background removal.

no license declared → metadata onlyupdated Jul 2026
★ 257repo stars
fal-model-guide
by JosiahSiegel · JosiahSiegel/claude-plugin-marketplace

fal-model-guide helps you navigate fal.ai's model catalog by comparing FLUX, Stable Diffusion, Kling, LTX, and audio models across performance tiers. It provides side-by-side quality, speed, and pricing comparisons to match your use case—whether you need production-grade output, fast iteration, or cost optimization.

MITupdated Jun 2026
★ 49repo stars
fal-text-to-image
by JosiahSiegel · JosiahSiegel/claude-plugin-marketplace

fal-text-to-image provides access to multiple text-to-image generation models including FLUX variants, SDXL, and specialized options like Recraft for design assets. Configure parameters like guidance scale, inference steps, image size presets, and batch generation to control output quality and speed.

MITupdated Jun 2026
★ 49repo stars
Fan Cam
by fal-ai-community · fal-ai-community/skills

Fan Cam transforms a user photo into a broadcast-ready spectator video for any sport. The skill handles scene composition, crowd reactions, and overlay design, then renders the final video through genmedia's pipeline. Ideal for creating realistic fan moments that feel like genuine live-event cutaways.

no license declared → metadata onlyupdated May 2026
★ 218repo stars
fal-image-to-video
by JosiahSiegel · JosiahSiegel/claude-plugin-marketplace

fal-image-to-video brings multiple image animation engines into one skill, supporting Kling 2.5/2.6 Pro, MiniMax Hailuo, LTX, Runway Gen-3 Turbo, Luma Dream Machine, and Stable Video Diffusion. Choose the right model for your use case—from cinematic portraits to looping ambient scenes—and describe the motion you want to see.

MITupdated Jun 2026
★ 49repo stars
Storytelling
by fal-ai-community · fal-ai-community/skills

Storytelling structures multi-shot creative projects by breaking narratives into beats, planning shot sequences with continuity anchors, and executing genmedia runs across video, audio, and image models. It guides you through reference uploads, model selection, and async job management to produce coordinated sequences rather than isolated assets.

no license declared → metadata onlyupdated May 2026
★ 218repo stars

More skills fal-text-to-video (MIT)

Tags
generative-mediamultimodal-outputprompt-based-creationasync-processingcost-estimationmodel-discoverybatch-generationreproducible-seeds