loopling.ai

Guides

Rate limits & concurrency

Stay within each key’s request rate and each account’s concurrency.

Limits

LimitApplies toOver it
60 requests a minuteEach key, across every authenticated /v1 route. The public /v1/models catalogue isn't counted.429 rate_limit_exceeded
2 generations at onceYour account, across keys and the Playground429 concurrency_limit
Daily generation capYour account, set by your plan429 daily_cap_reached
Ink cap, optionalEach key402 key_ink_limit_reached

Headers

Every 429 and 503 carries Retry-After: the seconds to wait before you retry.

codeRetry-After
rate_limit_exceededUntil the minute frees a slot
concurrency_limit15
daily_cap_reached3600
provider_unavailable5 or 60
rate_limiter_unavailable5

Responses carry no rate counters. Track your own request rate.

Concurrency

A submit over the concurrent limit is refused: nothing is held and nothing runs. Queue on your side, and submit the next generation when a poll returns succeeded or failed.

Back off

  • Wait at least Retry-After, then retry.
  • If it repeats, double the wait each time, up to a minute.
  • Poll each generation every 5 seconds, no faster. Polls count toward the per-minute limit.

Explore models

loopling.ai · grows with you, grows itself

hello@loopling.ai · for agents · community · pricing · Enterprise · API · privacy · terms