Guides

Limits & quotas

Codiv is a free public experiment. Limits keep the shared GPU fast for everyone.

Token quota

Every account gets 100M free input tokens. Output tokens are always 0 and are not counted. Usage is shown on the dashboard. When the quota is used up, requests return 429 with error_type: "quota_exceeded_error". Requests already in flight at that moment still complete. Paid plans with larger quotas are coming soon.

Rate limits

LimitValueWhen exceeded
Requests per key1,200 per minute (default)429 rate_limit_error
Active keys per account10Key creation is refused
Server capacityShared529 overloaded_error with retry-after

Request limits

  • Context window: 65,536 tokens. The state and questions are read together, so keep the state a little under that (about 60,000 tokens leaves room for a large question set).
  • Request body: 512 KB.
  • choice: up to 128 options. score: 2 to 10 levels.
  • Questions per request: no fixed limit. Large sets are answered in parallel chunks.

Retries

Retry 429 rate_limit_error and 529 with exponential backoff, and honour retry-after when it is present. The TypeSafe SDKs do this for you. Do not retry quota_exceeded_error.