video-to-videoWorkflow Utilities Reverse Video
fal-ai/workflow-utilities/reverse-videoFFMPEG Utility to Reverse Videos
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
video-to-videofal-ai/workflow-utilities/reverse-videoFFMPEG Utility to Reverse Videos
image-to-imagefal-ai/florence-2-large/referring-expression-segmentationFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
image-to-imagefal-ai/fast-sdxl/inpaintingRun SDXL at the speed of light
video-to-videofal-ai/wan-vace-apps/video-editEdit videos using plain language and Wan VACE
image-to-imagefal-ai/image-editing/plushie-styleTransform your photos into cool plushies while keeping the original characters likeness
image-to-videoblackforestlabs/flux-3/keyframes-to-video/draftFLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews pinned to your keyframe images, with a reusable draft cache for full-quality enhancement.
image-to-imagefal-ai/kolors/image-to-imagePhotorealistic Image-to-Image
image-to-imagefal-ai/image-editing/youtube-thumbnailsGenerate YouTube thumbnails with custom text
text-to-imagefal-ai/sdxl-controlnet-unionAn efficent SDXL multi-controlnet text-to-image model.
image-to-image
image-to-jsonfal-ai/bagel/understandBagel is a 7B parameter multimodal model from Bytedance-Seed that can generate both text and images.
image-to-videofal-ai/amt-interpolation/frame-interpolationInterpolate between image frames
image-to-imagefal-ai/leffa/pose-transferLeffa Pose Transfer is an endpoint for changing pose of an image with a reference image.
text-to-imagefal-ai/flux-1/srpoFLUX.1 SRPO [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.
image-to-imagefal-ai/flux-2-lora-gallery/add-backgroundAdd a background to images with white/clean background
text-to-videofal-ai/longcat-video/distilled/text-to-video/720pGenerate long videos in 720p/30fps from text using LongCat Video Distilled
text-to-audiofal-ai/kokoro/brazilian-portugueseA natural and expressive Brazilian Portuguese text-to-speech model optimized for clarity and fluency.
image-to-videofal-ai/fast-svd-lcmGenerate short video clips from your images using SVD v1.1 at Lightning Speed
image-to-imagefal-ai/chrono-editNVIDIA's Logically Consistent and Physics-Aware Image Editing Model
image-to-imagefal-ai/rifeInterpolate images with RIFE - Real-Time Intermediate Flow Estimation
text-to-imagefal-ai/fast-sdxl-controlnet-cannyGenerate Images with ControlNet.
text-to-videofal-ai/kandinsky5-pro/text-to-videoKandinsky 5.0 Pro is a diffusion model for fast, high-quality text-to-video generation.
image-to-imagefal-ai/ideogram/character/remixTransform your consistent character into different art styles, settings, or scenarios while maintaining their distinctive appearance and identity
text-to-videofal-ai/cogvideox-5bGenerate videos from prompts using CogVideoX-5B
image-to-videofal-ai/vidu/start-end-to-videoVidu Start-End to Video generates smooth transition videos between specified start and end images.
audio-to-videofal-ai/ltx-2.3-22b/audio-to-videoGenerate video with audio from audio, text and images using LTX-2
image-to-imagefal-ai/image-editing/age-progressionSee how you or others might look at different ages, from younger to older, while preserving core facial features.
audio-to-audiofal-ai/qwen-3-tts/clone-voice/0.6bClone your voices using Qwen3-TTS Clone-Voice model with zero shot cloning capabilities and use it on text-to-speech models to create speeches of yours!