Fal
Low-latency generative media APIs for Flux, video models and custom pipelines.
8.6AIAI Models
Usage-based
Run community models from a URL — image, video, voice and custom containers.
Pricing
Pay as you go
Usage-based
Low-latency generative media APIs for Flux, video models and custom pipelines.
8.6AIAI Models
Usage-based
Train, fine-tune and serve open models on a dedicated inference cloud.
8.4AIAI Models
Usage-based
The public hub for models, datasets, Spaces and inference APIs.
9.1AIAI Models
$9/mo
General-purpose AI assistant for writing, research, coding and image generation.
9.3AIAI Chat
$20/mo
Run Llama, Qwen, DeepSeek, Gemma and other local models with one command.
9.2AIAI Models
One API to route across hundreds of models with spend controls and fallbacks.
8.8AIAI Models
Usage-based
LPU inference for open models — very fast tokens, simple OpenAI-compatible API.
8.7AIAI Models
Usage-based