fal.ai
fal
An inference platform serving image and video models by the API call, priced per output.

Platforms
Overview
fal runs image and video generation models fast. The site calls it "fast, reliable, and cost-efficient" and says you only pay for the computing power you consume. Per-model rates are published: Seedream V4 at $0.03 per image, Flux Kontext Pro at $0.04 per image, Wan 2.5 at $0.05 per second of video and Veo 3 at $0.4 per second. Dedicated GPUs are also available by the hour, from $1.89/hr for an H100 at discounted rates.
Main features
- Image and video models over an API
- Published per-model rates
- On-demand GPUs
- Pay only for what you generate
- Custom enterprise arrangements
Pricing
Image models (usage-based)
$0.03 / image
Published price (USD) · Source
- Seedream V4: $0.03/image
- Flux Kontext Pro: $0.04/image
- Nanobanana: $0.0398/image
- Qwen: $0.02/megapixel
Last verified: September 2026
Video models (usage-based)
$0.05 / second of video
Published price (USD) · Source
- Wan 2.5: $0.05/second
- Kling 2.5 Turbo Pro: $0.07/second
- Veo 3: $0.4/second
- Ovi: $0.2/video
Last verified: September 2026
On-demand GPU
$1.89 / GPU hour
Published price (USD) · Source
- H100: $1.89/hr (discounted)
- B300: $8.50/hr (list price)
- Contact sales for discounted rates
Last verified: September 2026
Enterprise
Contact sales
- Custom arrangements
- Enterprise contact form and support email
fal.ai publishes one price worldwide and bills in USD. There is no separate price in your local currency.
Screenshots

Related AI tools
Replicate
Replicate is an AI infrastructure and API service useful for AI development and prototyping.
Together AI
Together AI is an AI infrastructure and API service useful for AI development and inference.
Fireworks AI
Inference, fine-tuning and dedicated deployments for open models, billed by tokens and GPU hours.
Hugging Face
Hugging Face is an AI development platform & API service that can be used for AI development and model discovery.