text-to-videoMiniMax (Hailuo AI) Video 01 Director
fal-ai/minimax/video-01-directorGenerate video clips more accurately with respect to natural language descriptions and using camera movement instructions for shot control.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-videofal-ai/minimax/video-01-directorGenerate video clips more accurately with respect to natural language descriptions and using camera movement instructions for shot control.
visionfal-ai/nemotron-diffusion-vlmNemotron-Labs-Diffusion-VLM-8B is the vision-language extension of the Nemotron-Labs-Diffusion family.
trainingfal-ai/flux-2-trainer-v2/editFine-tune FLUX.2 [dev] from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.
image-to-imagefal-ai/playground-v25/inpaintingState-of-the-art open-source model in aesthetic quality
text-to-imagefal-ai/flux-2-lora-gallery/sepia-vintageApplies sepia vintage effect to images
text-to-videofal-ai/kling-video/v1.6/pro/effectsGenerate video clips from your prompts using Kling 1.6 (pro)
video-to-videotopaz/denoise/videoProfessional video denoising powered by Topaz Labs. Nyx models remove noise at source resolution, with Nyx Fast as a lighter, cheaper pass. Best for low-light and high-ISO footage.
text-to-speechfal-ai/vibevoice/0.5bGenerate long speech snippets fast using Microsoft's powerful TTS.
image-to-imagefal-ai/qwen-image-edit-plus-lora-gallery/next-sceneCreate cinematic transitions and scene progressions (camera movements, framing changes)
audio-to-videofal-ai/ltx-2-19b/audio-to-videoGenerate video with audio from audio, text and images using LTX-2
trainingfal-ai/flux-2-trainer/editFine-tune FLUX.2 [dev] from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.
trainingfal-ai/wan-22-image-trainerWan 2.2 text to image LoRA trainer. Fine-tune Wan 2.2 for subjects and styles with unprecedented detail.
text-to-jsonbria/fibo-lite/generate/structured_promptConvert plain text into Fibo-Lite's transparent JSON-structured prompts - Bria's unique controllability layer that no closed model offers. Built for agentic and enterprise workflows.
video-to-videomirelo-ai/sfx-v1/video-to-videoGenerate synced sounds for any video, and return it with its new sound track (like MMAudio)
trainingfal-ai/wan-22-trainer/t2v-a14bTrain custom LoRAs for Wan-2.2 T2V/I2V 480P
audio-to-audiofal-ai/stable-audio-3/medium/base/audio-inpaintingStable Audio 3 Medium Base audio inpainting is the foundational 1.4 billion parameter checkpoint for editing or filling selected stereo audio segments guided by text prompts.
text-to-video
image-to-imagefal-ai/image-preprocessors/hedHolistically-Nested Edge Detection (HED) preprocessor.
audio-to-audiofal-ai/stable-audio-3/small/sfx/audio-to-audioStable Audio 3 Small SFX audio-to-audio is a 459 million parameter latent diffusion model that transforms input audio into new sound-effect variations guided by text prompts.
text-to-imagefal-ai/flux-2-lora-gallery/ballpoint-pen-sketchBallpoint pen sketch drawing style
text-to-imagefal-ai/stable-cascade/sote-diffusionAnime finetune of Würstchen V3.
audio-to-audiofal-ai/tada/3b/text-to-speechA unified speech-language model that synchronizes speech and text into a single, cohesive stream via 1:1 alignment.
text-to-videofal-ai/fast-svd/text-to-videoGenerate short video clips from your prompts using SVD v1.1
text-to-imagefal-ai/fast-lcm-diffusionRun SDXL at the speed of light
trainingfal-ai/wan-22-trainer/i2v-a14bTrain custom LoRAs for Wan-2.2 T2V/I2V 480P
image-to-imagefal-ai/sdxl-controlnet-union/inpaintingAn efficent SDXL multi-controlnet inpainting model.
text-to-videofal-ai/krea-wan-14b/text-to-videoFast Text-to-Video endpoint for Krea's Wan 14b model.
video-to-videofal-ai/pixverse/extend/fastPixVerse Extend model is a video extending tool for your videos using with high-quality video extending techniques