SpyBara
Go Premium

Documentation 2026-09-11 21:00 UTC to 2026-09-13 05:00 UTC

3 files changed +13 −6. View all changes and history on the product overview
2026
Wed 30 23:57 Tue 29 23:59 Mon 28 23:57 Sun 27 22:59 Sat 26 23:59 Fri 25 23:01 Thu 24 23:59 Wed 23 23:59 Tue 22 23:58 Mon 21 23:00 Sun 20 23:01 Sat 19 23:59 Fri 18 23:59 Thu 17 10:04 Wed 16 19:01 Tue 15 17:00 Mon 14 06:00 Sun 13 05:00 Fri 11 21:00 Tue 8 21:00 Mon 7 22:57 Thu 3 16:59 Wed 2 22:03
Details

135|-------|-------------|135|-------|-------------|

136| `grok-voice-latest` | Alias for `grok-voice-think-fast-2.0` |136| `grok-voice-latest` | Alias for `grok-voice-think-fast-2.0` |

137| `grok-voice-think-fast-2.0` | Flagship voice model |137| `grok-voice-think-fast-2.0` | Flagship voice model |

138| `grok-voice-think-fast-1.0` | Previous-generation voice model |

139 138 

140## Session Parameters139## Session Parameters

141 140 

rate-limits.md +12 −4

Details

4 4 

5Every xAI API team has per-model rate limits on two dimensions: **requests per second (RPS)** and **tokens per minute (TPM)**. Your per-second limit is derived from your per-minute request budget (RPM / 60): you cannot spend a full minute's requests in a single second, which protects the API from sudden bursts. These limits scale with your team's **tier**, which is determined by cumulative spend on the API.5Every xAI API team has per-model rate limits on two dimensions: **requests per second (RPS)** and **tokens per minute (TPM)**. Your per-second limit is derived from your per-minute request budget (RPM / 60): you cannot spend a full minute's requests in a single second, which protects the API from sudden bursts. These limits scale with your team's **tier**, which is determined by cumulative spend on the API.

6 6 

7You can view your team's current tier and per-model limits on the [Rate Limits](https://console.x.ai/team/default/rate-limits?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=rate-limits) page in the xAI Console.7You can view your team's current tier and per-model limits on the [Models](https://console.x.ai/team/default/models?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=models) page in the xAI Console.

8 8 

9## Rate limit tiers9## Rate limit tiers

10 10 


23 23 

24> [!NOTE]24> [!NOTE]

25>25>

26> Rate limit tiers apply to text and embedding models. For increases to Voice and Imagine API limits, contact [sales@x.ai](mailto:sales@x.ai).26> Rate limit tiers apply to text, embedding, and voice models. For increases to Imagine API limits, contact [sales@x.ai](mailto:sales@x.ai).

27 27 

28## Per-model limits28## Per-model limits

29 29 

30Each tier sets hard RPS and TPM caps per model. Limits scale exponentially with tier. Exceeding any limit returns a `429 Too Many Requests` error.30Each tier sets hard RPS and TPM caps per model. Limits scale exponentially with tier. Exceeding any limit returns a `429 Too Many Requests` error.

31 31 

32The table below lists RPS and TPM limits at each tier for every model. You can also view your team's personalized limits on the [Rate Limits](https://console.x.ai/team/default/rate-limits?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=rate-limits) page in the xAI Console.32The table below lists RPS and TPM limits at each tier for every model. Voice & Audio endpoints follow, limited by requests per second and concurrent sessions (CST) rather than tokens; Speech to Speech is limited by concurrent sessions alone. You can also view your team's personalized limits on the [Models](https://console.x.ai/team/default/models?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=models) page in the xAI Console.

33 33 

34| Model | RPS | TPM |34| Model | RPS | TPM |

35| --- | --- | --- |35| --- | --- | --- |


46| grok-imagine-video-1.5 | T0: 10, T1: 20, T2: 39, T3: 79, T4: 158 | — |46| grok-imagine-video-1.5 | T0: 10, T1: 20, T2: 39, T3: 79, T4: 158 | — |

47| grok-imagine-video | T0: 10, T1: 20, T2: 39, T3: 79, T4: 158 | — |47| grok-imagine-video | T0: 10, T1: 20, T2: 39, T3: 79, T4: 158 | — |

48 48 

49**Voice & Audio**

50 

51| Model | RPS | Concurrent sessions |

52| --- | --- | --- |

53| grok-voice-think-fast-2.0 | — | T0: 10, T1: 20, T2: 50, T3: 100, T4: 200 |

54| Text to Speech | T0: 50, T1: 50, T2: 100, T3: 250, T4: 500 | T0: 100, T1: 200, T2: 200, T3: 300, T4: 500 |

55| Speech to Text | T0: 10, T1: 10, T2: 20, T3: 30, T4: 40 | T0: 100, T1: 200, T2: 200, T3: 300, T4: 500 |

56 

49### What counts toward TPM57### What counts toward TPM

50 58 

51All tokens consumed by a request count toward the TPM limit for that model:59All tokens consumed by a request count toward the TPM limit for that model:


105## Increasing your limits113## Increasing your limits

106 114 

107* **Spend more.** Tiers upgrade automatically based on cumulative spend. No action required on your part.115* **Spend more.** Tiers upgrade automatically based on cumulative spend. No action required on your part.

108* **Request an increase.** Submit a request through the [xAI Console](https://console.x.ai/team/default/rate-limits?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=rate-limits) if you need higher limits without additional spend, or limits beyond Tier 4.116* **Request an increase.** Submit a request through the [xAI Console](https://console.x.ai/team/default/models?utm_source=docs\&utm_medium=referral\&utm_campaign=developers-rate-limits\&utm_content=models) if you need higher limits without additional spend, or limits beyond Tier 4.

109* **Contact sales.** For enterprise-grade capacity, please email [sales@x.ai](mailto:sales@x.ai).117* **Contact sales.** For enterprise-grade capacity, please email [sales@x.ai](mailto:sales@x.ai).

Details

14 14 

15* `session` (object | null) — Optional initial session configuration to bind to the client secret. This JSON value is stored alongside the secret and applied when the WebSocket connection opens.15* `session` (object | null) — Optional initial session configuration to bind to the client secret. This JSON value is stored alongside the secret and applied when the WebSocket connection opens.

16 16 

17 * `model` ("grok-voice-latest" | "grok-voice-think-fast-2.0" | "grok-voice-think-fast-1.0") — Model to use for the session. Use grok-voice-latest for the best experience.17 * `model` ("grok-voice-latest" | "grok-voice-think-fast-2.0") — Model to use for the session. Use grok-voice-latest for the best experience.

18 18 

19 * `reasoning` (object) — Reasoning settings for models that support them.19 * `reasoning` (object) — Reasoning settings for models that support them.

20 20