trainingQwen Image 2512 Trainer
fal-ai/qwen-image-2512-trainerQwen Image 2512 LoRA training
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
trainingfal-ai/qwen-image-2512-trainerQwen Image 2512 LoRA training
image-to-videofal-ai/ltx-2-19b/distilled/image-to-videoGenerate video with audio from images using LTX-2 Distilled
image-to-imagefal-ai/qwen-image-edit-2509Endpoint for Qwen's Image Editing Plus model also known as Qwen-Image-Edit-2509. Has superior text editing capabilities and multi-image support.
image-to-imagefal-ai/live-portrait/imageTransfer expression from a video to a portrait.
image-to-imagebria/upscale/creativeProfessional-grade creative upscaler that doubles resolution up to 10MP, regenerating sharper textures, refined details, and cleaner faces. Trained exclusively on licensed data for risk-free commercial use.
image-to-imagefal-ai/lora/image-to-imageRun Any Stable Diffusion model with customizable LoRA weights.
video-to-videofal-ai/ltx-2.3-quality/render-to-realTransform your 3D video render into realistic using first frame with Ltx 2.3
text-to-videofal-ai/wan/v2.2-5b/text-to-video/fast-wanWan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
video-to-audiomirelo-ai/sfx-v1.5/video-to-audioGenerate synced sounds for any video, and return the new sound track (like MMAudio)
text-to-imagefal-ai/ernie-imageHigh-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.
text-to-videofal-ai/heygen/v3/video-agentGenerate videos with a single prompt. Describe what you want in plain text, and the agent handles avatar selection, scripting, scene composition - all in one.
image-to-imagefal-ai/bria/genfillBria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us
text-to-audiofal-ai/stable-audio-3/medium/base/text-to-audioStable Audio 3 Medium Base is the foundational 1.4 billion parameter text-to-audio checkpoint generating stereo music up to 6 minutes, intended as the unmodified base for custom fine-tuning workflows.
text-to-imagefal-ai/playground-v25State-of-the-art open-source model in aesthetic quality
image-to-imagefal-ai/flux/krea/image-to-imageFLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.
image-to-videofal-ai/vidu/q3/reference-to-video/mixVidu's latest Q3 Reference to Video Mix model
image-to-imagefal-ai/fast-lightning-sdxl/image-to-imageRun SDXL at the speed of light
text-to-imagerecraft/v4/style/text-to-vectorGenerates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.
text-to-imagefal-ai/pony-v7Pony V7 is a finetuned text to image for superior aesthetics and prompt following.
text-to-speechfal-ai/kling-video/v1/ttsGenerate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-imagefal-ai/recraft-20bRecraft 20b is a new and affordable text-to-image model.
text-to-videofal-ai/minimax/video-01-liveGenerate video clips from your prompts using MiniMax model
video-to-videofal-ai/wan-22-vace-fun-a14b/inpaintingVACE Fun for Wan 2.2 A14B from Alibaba-PAI
image-to-videofal-ai/ltx-2.3-22b/image-to-videoGenerate video with audio from images using LTX-2.3
image-to-videofal-ai/wan-effectsWan Effects generates high-quality videos with popular effects from images
image-to-image
text-to-imagefal-ai/recraft/v4/pro/text-to-vectorRecraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.
video-to-textopenrouter/router/video/enterpriseRun any VLM (Video Language Model) with fal, powered by OpenRouter.