image-to-imageIC-Light-v2 for Image Relighting
fal-ai/iclight-v2An endpoint for re-lighting photos and changing their backgrounds per a given description
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-imagefal-ai/iclight-v2An endpoint for re-lighting photos and changing their backgrounds per a given description
text-to-videoalibaba/happy-horse/text-to-videoGenerate 1080p video with synchronized native audio from a text prompt. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.
image-to-imageideogram/v4/image-to-imageIdeogram V4.0q Image-to-Image transforms an input image with a text prompt, restyling and reworking the composition while preserving its core structure for prompt-faithful, high-fidelity edits.
video-to-videofal-ai/kling-video/o3/4k/video-to-video/editKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
video-to-audiosonilo/v1.1/video-to-sound-effectsAnalyzes a video and generates synchronized, royalty-free sound effects timed to visible actions. Returns the generated sound-effects audio track for commercial use.
video-to-videofal-ai/bytedance/dreamactor/v2Transfer motion from a video to characters in an image using Dreamactor v2. Great performance for non-human and multiple characters
text-to-videofal-ai/wan/v2.2-a14b/text-to-video/turboWan-2.2 turbo text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.
image-to-imagefal-ai/flux-2/klein/9b/base/edit/loraImage-to-image editing with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Specialized style transfer and domain-specific modifications.
image-to-3dfal-ai/hunyuan-3d/v3.1/rapid/image-to-3dRapidly generate 3D models from images using Hunyuan 3D.
image-to-3dfal-ai/hunyuan3d/v2Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.
image-to-videoblackforestlabs/flux-3/image-to-video/draftFLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews that animate a still image, with a reusable draft cache for full-quality enhancement.
image-to-imagefal-ai/flux-lora-fillFLUX.1 [dev] Fill is a high-performance endpoint for the FLUX.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.
image-to-imagefal-ai/bria/background/replaceBria Background Replace allows for efficient swapping of backgrounds in images via text prompts or reference image, delivering realistic and polished results. Trained exclusively on licensed data for safe and risk-free commercial use
image-to-imagefal-ai/qwen-image-edit/inpaintInpainting Endpoint for the Qwen Edit Image editing model.
image-to-imagefal-ai/ideogram/v3/reframeExtend existing images with Ideogram V3's reframe feature. Create expanded versions and adaptations while preserving main image and adding new creative directions through prompt guidance.
image-to-3dfal-ai/sam-3/3d-bodySAM 3D allows for accurate 3D reconstruction of human body shape and position from a single image.
visionfal-ai/moondream3-preview/detectMoondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.
image-to-videofal-ai/minimax/hailuo-02-fast/image-to-videoCreate blazing fast and economical videos with MiniMax Hailuo-02 Image To Video API at 512p resolution
text-to-audio
text-to-imagefal-ai/kling-image/v3/text-to-imageKling V3: Latest Kling Image model
text-to-imagefal-ai/hunyuan-image/v3/text-to-imageLeverage the state-of-the-art capabilities of Hunyuan Image 3.0 to generate visual content that effectively conveys the messaging of your written material.
text-to-imageideogram/v4/instantGenerate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs FRACTION OF A SECOND.
image-to-imagefal-ai/florence-2-large/caption-to-phrase-groundingFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
image-to-3dtripo3d/triposplatTripoSplat is an open-source model from TripoAI / VAST AI Research that converts a single 2D image into high-quality 3D Gaussians using a novel learned density-control approach
image-to-imagefal-ai/sam2/auto-segmentSAM 2 is a model for segmenting images automatically. It can return individual masks or a single mask for the entire image.
image-to-videofal-ai/kling-video/o1/standard/image-to-videoGenerate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.
audio-to-audiofal-ai/kling-video/create-voiceCreate Voices to be used with Kling Models Voice Control
text-to-audiofal-ai/stable-audio-3/small/sfx/text-to-audioStable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.