Fal
Create and edit high-quality media across various formats including images, videos, audio, and 3D models.
Community: Submitted by a user or imported; check the owner before granting accessOnlineAPI key requiredGlobalFreeCan modify data
What it can do
- Generate Image: Generate images from a text prompt. Models: flux-dev (balanced), flux-schnell (fast), flux-pro (best quality), recraft-v4 (design/typography), nano-banana (ultra-fast/cheap), nano-bana
- Edit Image: Edit or transform an existing image using AI. Models: flux-kontext (default), nano-banana-2-edit (Gemini 3.1 Flash, fast), nano-banana-pro-edit (Gemini 3 Pro), gpt-image-edit (GPT-Image 1.
- Generate Video: Generate a video from text and/or an image. Models: kling-v3-pro (best), sora-2-t2v, sora-2-i2v, ltx-2 (fast), wan-v2
What data it sees
Do you need an account
An API key from the service settings is required
Create and edit high-quality media across various formats including images, videos, audio, and 3D models. Enhance creative workflows by performing advanced tasks like background removal, image upscaling, and text-to-speech generation. Access a wide range of state-of-the-art generative models to bring complex visual and auditory ideas to life instantly.
Server tool list (9)
Raw names from tools/list. Only developers need these.
| generate-image | Generate images from a text prompt. Models: flux-dev (balanced), flux-schnell (fast), flux-pro (best quality), recraft-v4 (design/typography), nano-banana (ultra-fast/cheap), nano-banana-2 (Gemini 3.1 Flash), nano-banana-pro (Gemini 3 Pro, best), gpt-image (GPT-Image 1.5), seedream-v4 (ByteDance, cheap) |
| edit-image | Edit or transform an existing image using AI. Models: flux-kontext (default), nano-banana-2-edit (Gemini 3.1 Flash, fast), nano-banana-pro-edit (Gemini 3 Pro), gpt-image-edit (GPT-Image 1.5), seedream-v4-edit (ByteDance, cheap) |
| generate-video | Generate a video from text and/or an image. Models: kling-v3-pro (best), sora-2-t2v, sora-2-i2v, ltx-2 (fast), wan-v2 |
| generate-audio | Generate audio: speech, music, or sound effects. Models: chatterbox-tts (TTS), minimax-speech (HD TTS), beatoven-music (music), beatoven-sfx (SFX) |
| generate-3d | Generate a 3D model from an image. Models: meshy-v6, trellis-2 (fast) |
| remove-background | Remove the background from an image |
| upscale-image | Upscale an image to higher resolution (4x by default) |
| run-model | Run any fal.ai model directly with custom arguments. Use for models not in the catalog or for advanced parameter control. |
| list-models | List available fal.ai models from the built-in catalog. Filter by category: "image", "image_edit", "video", "audio", "3d", "utility". Empty for all. |