Logo SVG copied to clipboard

Groq models on the LLM gateway

3M tokens / 30 days

6 models routed through the gateway over the last 30 days.

Groq runs open models like Llama and Qwen on its LPU engine for ultra-low latency. Route Groq through the gateway with a spending cap and fallback.

Groq tokens processed by the gateway

Models in use

# Model
Auth type Tokens / last 30 days
1.Gpt Oss 120b
API Key1 M
2.Gpt Oss 20b
API Key1 M
3.Qwen3.8 27b
API Key62 K
4.Qwen3.6 27b
API Key36 K
5.Gpt Oss Safeguard 20b
API Key14 K
6.Allam 2 7b
API Key38
View Groq errors 1 documented error