image-to-imageNano Banana 2
fal-ai/nano-banana-2/editNano Banana 2 is Google's new state-of-the-art image generation and editing model
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-imagefal-ai/nano-banana-2/editNano Banana 2 is Google's new state-of-the-art image generation and editing model
image-to-imagefal-ai/nano-banana-pro/editNano Banana Pro is Google's new state-of-the-art image generation and editing model
image-to-imageopenai/gpt-image-2/editGPT Image 2, OpenAI's latest image model, is capable of making fine-grained, detailed edits to images.
text-to-imagefal-ai/flux/schnellFLUX.1 [schnell] is a 12 billion parameter flow transformer that generates high-quality images from text in 1 to 4 steps, suitable for personal and commercial use.
text-to-imagefal-ai/nano-banana-2Nano Banana 2 is Google's new state-of-the-art fast image generation and editing model
text-to-imageopenai/gpt-image-2GPT Image 2, OpenAI's latest image model, is capable of creating extremely detailed images with fine typography.
image-to-imageopenai/gpt-image-2.5/sunburst/editEditing built for the tightest control, edits scoped precisely to the instruction, with subject and composition preserved across many rounds of revision.
image-to-videominimax/h3-max/image-to-videofal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality
text-to-imagefal-ai/nano-banana-proNano Banana Pro is Google's new state-of-the-art image generation and editing model
text-to-imagefal-ai/flux/devFLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.
image-to-imagefal-ai/nano-banana/editGoogle's famous original image generation and editing model
image-to-videominimax/h3-max-turbo/image-to-videofal's H3 Max Turbo is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality
image-to-videofal-ai/kling-video/v3/pro/image-to-videoKling 3.0 Pro: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation, with custom element support.
image-to-imageopenai/gpt-image-2.5/flare/editPrecise image editing that changes only what's asked, keeping subject, composition, and background intact, with reference subjects staying recognizable across styles and successive edits.
image-to-videominimax/h3-max/reference-to-videofal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality
image-to-imagebytedance/seedream/v5/pro/editSeedream 5.0 Pro is grounded, region-precise image editing model that changes one element while keeping the rest of the frame intact with layer separation, sketch completion, and up to 10 reference images.
image-to-imagefal-ai/birefnet/v2bilateral reference framework (BiRefNet) for high-resolution dichotomous image segmentation (DIS)
text-to-imageopenai/gpt-image-2.5/sunburst/text-to-imageOpenAI's precision-focused image model, built for premium visual work, extra fidelity on intricate detail, in exchange for longer generation times.
text-to-imagefal-ai/flux-2-proImage editing with FLUX.2 [pro] from Black Forest Labs. Ideal for high-quality image manipulation, style transfer, and sequential editing workflows
text-to-imageopenai/gpt-image-2.5/flare/text-to-imageOpenAI's default image model for most applications. Fast, high-quality generation with natural lighting, rich textures, and support for complex layouts including transparent backgrounds.
image-to-videobytedance/seedance-2.5/reference-to-videoDreamina Seedance 2.5 generates video from up to 50 multimodal references images, video, audio, and style inputs, locking a character, set, and palette across a full 30-second take for production-grade consistency.
text-to-videominimax/h3-max/text-to-videofal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality
text-to-imagefal-ai/nano-bananaGoogle's famous original image generation and editing model
image-to-videobytedance/seedance-2.5/image-to-videoDreamina Seedance 2.5 animates a single still into a native 30-second clip at up to 720p, extending one frame into continuous, coherent motion without the drift or stitching of shorter multi-clip workflows.
image-to-imagefal-ai/flux-pro/kontextFLUX.1 Kontext [pro] handles both text and reference images as inputs, seamlessly enabling targeted, local edits and complex transformations of entire scenes.
fal-ai/flux-pro/v1.1FLUX1.1 [pro] is an enhanced version of FLUX.1 [pro], improved image generation capabilities, delivering superior composition, detail, and artistic fidelity compared to its predecessor.
image-to-imagefal-ai/bytedance/seedream/v4.5/editA new-generation image creation model ByteDance, Seedream 4.5 integrates image generation and image editing capabilities into a single, unified architecture.
image-to-videofal-ai/kling-video/v2.5-turbo/pro/image-to-videoKling 2.5 Turbo Pro: Top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.