llmOpenRouter [Enterprise]
openrouter/router/enterpriseRun any LLM (Large Language Model) with fal, powered by OpenRouter.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
llmopenrouter/router/enterpriseRun any LLM (Large Language Model) with fal, powered by OpenRouter.
visionfal-ai/moondream-nextMoonDreamNext is a multimodal vision-language model for captioning, gaze detection, bbox detection, point detection, and more.
image-to-imagefal-ai/image2svgImage2SVG transforms raster images into clean vector graphics, preserving visual quality while enabling scalable, customizable SVG outputs with precise control over detail levels.
text-to-imagerecraft/v4/style/pro/text-to-imageGenerates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.
text-to-imagefal-ai/recraft/v4.1/pro/text-to-vectorRecraft V4.1 Pro Vector generates large-format, fully editable SVGs with the structural clarity professional illustrators expect. Built for poster art, complex brand assets, and detailed scene illustration, it scales without losing geometric integrity.
image-to-imagefal-ai/hidream-o1-image/editUnified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.
image-to-videofal-ai/luma-dream-machine/ray-2-flash/image-to-videoRay2 Flash is a fast video generative model capable of creating realistic visuals with natural, coherent motion.
text-to-videofal-ai/minimax/hailuo-2.3/pro/text-to-videoMiniMax Hailuo-2.3 Text To Video API (Pro, 1080p): Advanced text-to-video generation model with 1080p resolution
image-to-imagefal-ai/flux-2/klein/realtimeRealtime generation with FLUX.2 [klein] from Black Forest Labs.
text-to-audiomirelo-ai/sfx1.6/text-to-audioGenerate ambient sounds for any text prompt. Now you can turn any SFX into a natural loop for ambient soundscapes.
image-to-imageluma/agent/uni-1/v1/max/editLuma Uni-1 Max Edit applies text-guided edits to a source image at maximum fidelity, holding the original structure while honoring reference images for precise, high-detail revisions.
text-to-audiofal-ai/minimax-music/v2.5MiniMax Music 2.5 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.
video-to-videofal-ai/kling-video/o1/video-to-video/referenceKling O1 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.
video-to-videofal-ai/wan-motionWan Motion is a streamlined character animation model that transfers motion from a driving video onto a reference character image. Based on Wan-Animate which preserves the original character's proportions, Simple uses pose retargeting to adapt the driving video's skeleton to match the reference character's body shape, producing more natural results when the two have different builds. It outputs at 720p with optimized defaults for fast, high-quality generation — just provide a video, an image, and an optional prompt.
text-to-imagefal-ai/flux-1/kreaFLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.
visionfal-ai/moondream2Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.
text-to-imagefal-ai/flux-2/klein/9b/baseText-to-image generation with FLUX.2 [klein] 9B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.
visionfal-ai/moondream2/object-detectionMoondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.
text-to-imagefal-ai/ernie-image/turboHigh-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.
image-to-videofal-ai/pixverse/c1/reference-to-videoGenerate character-consistent videos from reference images using PixVerse C1, with subject and background references.
video-to-video
image-to-imagefal-ai/fashn/tryon/v1.5FASHN v1.5 delivers precise virtual try-on capabilities, accurately rendering garment details like text and patterns at 576x864 resolution from both on-model and flat-lay photo references.
text-to-imagefal-ai/hidream-i1-fullHiDream-I1 full is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.
image-to-imagefal-ai/imageutils/marigold-depthCreate depth maps using Marigold depth estimation.
image-to-imagefal-ai/flux-pro/v1/vtoGenerate virtual try-on results from a person image plus one or more garment references.
text-to-imagefal-ai/flux-2-lora-gallery/realismMakes images more photorealistic and natural
text-to-imagebria/fibo-gen-1.5/text-to-imageText-to-image model with high-fidelity outputs, accurate typography, and style preset, strong in photorealism, textures, and beyond. JSON-structured prompts give enterprise and agentic workflows production-ready control. Trained on licensed data.
image-to-imagefal-ai/z-image/turbo/image-to-image/loraGenerate images from text and images using custom LoRA and Z-Image Turbo, Tongyi-MAI's super-fast 6B model.