Skip to main content
When you deploy your own applications on fal Serverless, you are billed for the total time your runners are alive, measured per-second by machine type.

Billing by runner state

Every runner transitions through the states below during its lifecycle. You are billed for the states marked Yes at the per-second rate for your machine type. 5xx errors (HTTP 500+) are also not charged. See Runners for full details on each state and transitions.

GPU count multiplier

Multi-GPU instances are billed as gpu_count x duration. For example, a runner using 2x A100 GPUs for 60 seconds is billed as 120 GPU-seconds.

Monitoring your usage

Dashboard Billing

View your overall spend, invoices, and payment methods.

App Analytics

See per-app cost breakdown, request counts, and runner utilization.