Flux1 AIDevelopers
Flux1 AI API docs

Rate limits

Per-key request limits, a per-account cap on in-flight generations, and the headers that tell you where you stand.

LimitValueScope
Submissions30 / minuteper key
Reads (poll, list, models, credits)300 / minuteper key
In-flight generations10 queued at onceper account (all queued tasks)
Failed authentications20 / minuteper IP

Every authenticated response carries x-ratelimit-limit, x-ratelimit-remaining and x-ratelimit-reset (Unix seconds). A 429 adds retry-after in seconds.

Two kinds of 429

  • rate_limited means the per-key request rate for this window is used up. Wait for retry-after and continue.
  • concurrency_limited means 10 of your generations are still queued. Poll and let them finish before submitting more. All queued tasks count, including older tasks. Concurrent submissions reserve capacity before billable work starts.

Neither is charged. Need more? Contact us with your use case and expected volume.

On this page