EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

llm

Page 1 · 28 per page
llm
OpenAIREVIEW REQUIRED

OpenRouter Chat Completions [OpenAI Compatible]

openrouter/router/openai/v1/chat/completions

OpenAI-compatible chat completions API. Drop-in replacement for the OpenAI API — use any OpenAI SDK or client to access Claude, Gemini, Grok, DeepSeek, Llama, Qwen, Mistral, and all OpenAI models (GPT-5, GPT-4o, o3) through fal. Powered by OpenRouter.

llm
openrouterREVIEW REQUIRED

OpenRouter

openrouter/router

Run any LLM with fal. Access Claude (Anthropic), ChatGPT / GPT-5 / GPT-4o (OpenAI), Gemini (Google), Grok (xAI), DeepSeek, Llama (Meta), Qwen (Alibaba), Mistral, and 200+ more models through a single API. Supports reasoning, structured output, and streaming. Powered by OpenRouter.

llm
OpenAIREVIEW REQUIRED

OpenRouter Embeddings [OpenAI Compatible]

openrouter/router/openai/v1/embeddings

Generate text embeddings using OpenAI-compatible API. Access embedding models like text-embedding-3-small, text-embedding-3-large (OpenAI), and other embedding models available through OpenRouter. Drop-in replacement for the OpenAI embeddings API. Powered by OpenRouter.

llm
OpenAIREVIEW REQUIRED

OpenRouter Responses [OpenAI Compatible]

openrouter/router/openai/v1/responses

The OpenRouter Responses API with fal, powered by OpenRouter, provides unified access to a wide range of large language models - including GPT, Claude, Gemini, and many others through a single API interface.

llm
openrouterREVIEW REQUIRED

OpenRouter [Enterprise]

openrouter/router/enterprise

Run any LLM (Large Language Model) with fal, powered by OpenRouter.

llm
ByteDanceREVIEW REQUIRED

Bytedance Seed V2 Mini

fal-ai/bytedance/seed/v2/mini

Seed 2.0 Mini is a high-performance multimodal model optimized for low latency and high concurrency. It supports text, image, and video input with 256K context and configurable thinking/reasoning modes.

llm
falREVIEW REQUIRED

Video Prompt Generator

fal-ai/video-prompt-generator

Generate video prompts using a variety of techniques including camera direction, style, pacing, special effects and more.

motiontransformationchatclaude
llm
nvidiaREVIEW REQUIRED

Nemotron 3 Nano Omni

nvidia/nemotron-3-nano-omni

Open, efficient reasoning model from NVIDIA. 30B A3B hybrid Transformer-Mamba MoE, built for enterprise agentic workflows.

nemotronnvidiareasoningllm