image-to-videoMiniMax Hailuo 02 [Standard] (Image to Video)
fal-ai/minimax/hailuo-02/standard/image-to-videoMiniMax Hailuo-02 Image To Video API (Standard, 768p, 512p): Advanced image-to-video generation model with 768p and 512p resolutions
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-videofal-ai/minimax/hailuo-02/standard/image-to-videoMiniMax Hailuo-02 Image To Video API (Standard, 768p, 512p): Advanced image-to-video generation model with 768p and 512p resolutions
image-to-videogoogle/gemini-omni-flash/image-to-videoAnimates a still image into video with audio. Extends a single frame into coherent motion, grounded in Gemini's physical understanding of how scenes and subjects behave.
image-to-3dmeshy/v7/image-to-3dTurns a single image into a fully textured, PBR-ready 3D mesh with complete geometry, in game-ready Smart Topology at a target polygon count
image-to-imagefal-ai/flux-2-max/editFLUX.2 [max] delivers state-of-the-art image generation and advanced image editing with exceptional realism, precision, and consistency.
video-to-videotopaz/upscale/video/precisionProfessional video upscaling powered by Topaz Labs. Precision models (Proteus, Artemis, Iris, Dione, Theia, Gaia, Rhea) enhance footage up to 4x while staying faithful to the source. Best for clean, natural upscales of real-world footage.
image-to-3dfal-ai/trellis-2Generate 3D models from your images using Trellis 2. A native 3D generative model enabling versatile and high-quality 3D asset creation.
image-to-imagefal-ai/recraft/vectorizeConverts a given raster image to SVG format using Recraft model.
text-to-audiocassetteai/music-generatorCassetteAI’s model generates a 30-second sample in under 2 seconds and a full 3-minute track in under 10 seconds. At 44.1 kHz stereo audio, expect a level of professional consistency with no breaks, no squeaks, and no random interruptions in your creations.
image-to-imagefal-ai/gemini-3.1-flash-image-preview/editGemini 3.1 Flash Image (a.k.a. Nano Banana 2) is Google's new state-of-the-art fast image generation and editing model
image-to-videofal-ai/kling-video/v3/turbo/standard/image-to-videoKling 3.0 Turbo Standard animates a first and last frame reference image into 720P video with native audio, delivering quick, affordable image-driven motion for fast turnaround
image-to-imagexai/grok-imagine-image/quality/editGrok Imagine Pro is an advanced AI model from xAI that creates high-quality visuals from text prompts and allows you to edit or analyze existing images.
jsonfal-ai/ffmpeg-api/metadataGet encoding metadata from video and audio files using FFmpeg API.
image-to-imagefal-ai/flux-2/flash/editImage-to-image editing with FLUX.2 [dev] from Black Forest Labs. Precise modifications using natural language descriptions and hex color control—in a flash.
image-to-videofal-ai/wan/v2.7/image-to-videoWan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
image-to-videofal-ai/kling-video/o3/standard/reference-to-videoTransform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.
image-to-imagefal-ai/qwen-image-edit-2511Endpoint for Qwen's Image Editing 2511 model.
image-to-image
text-to-audiobytedance/seed-audio-1.0Seed Audio 1.0 is a new audio model from Bytedance that can generate high-quality, natural sounding audio using text, reference audios or an image.
image-to-videofal-ai/bytedance/seedance/v1/pro/fast/image-to-videoImage to Video endpoint for Seedance 1.0 Pro Fast, a next-generation video model designed to deliver maximum performance at minimal cost
text-to-imagefal-ai/flux-1/schnellFastest inference in the world for the 12 billion parameter FLUX.1 [schnell] text-to-image model.
audio-to-audiofal-ai/demucsSOTA stemming model for voice, drums, bass, guitar and more.
image-to-videofal-ai/kling-video/v2.5-turbo/standard/image-to-videoKling 2.5 Turbo Standard: Top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.
image-to-videobytedance/seedance-2.0/mini/reference-to-videoSeedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.
image-to-imagetopaz/upscale/image/generativeProfessional generative image upscaling powered by Topaz Labs. Wonder 3.5 leads the range, with Redefine for prompt-guided detail and Recovery for extreme low-resolution sources. Best for rebuilding sharp detail in small or blurry images.
image-to-imagebytedance/seedream/v5/lite/editImage editing endpoint for the fast Lite version of Seedream 5.0, supporting high quality intelligent image editing with multiple inputs.
image-to-3dtripo3d/h3.1/image-to-3dGenerate high-quality 3D models from a single image using Tripo H3.1.
trainingfal-ai/flux-lora-fast-trainingTrain styles, people and other subjects at blazing speeds.
image-to-videofal-ai/kling-video/v2.1/pro/image-to-videoKling 2.1 Pro is an advanced endpoint for the Kling 2.1 model, offering professional-grade videos with enhanced visual fidelity, precise camera movements, and dynamic motion control, perfect for cinematic storytelling.