Groq models on the LLM gateway
3M tokens
/ 30 days
6 models routed through the gateway over the last 30 days.
Groq runs open models like Llama and Qwen on its LPU engine for ultra-low latency. Route Groq through the gateway with a spending cap and fallback.
Groq tokens processed by the gateway
Models in use
# Model |
Auth type | Tokens / last 30 days |
|---|---|---|
1.Gpt Oss 120b | API Key | 1 M |
2.Gpt Oss 20b | API Key | 1 M |
3.Qwen3.8 27b | API Key | 62 K |
4.Qwen3.6 27b | API Key | 36 K |
5.Gpt Oss Safeguard 20b | API Key | 14 K |
6.Allam 2 7b | API Key | 38 |