inference cloud
Groq
Models, pricing and feature compatibility available through Groq.
Deployments
14
Models
14
Labs
6
With pricing
14
Models on Groq
Follow any model to compare this deployment with the same model through other providers.
| Model | Provider model ID | Mode | Region | Input | Output | Context | Tools |
|---|---|---|---|---|---|---|---|
| Gemma 7B IT | gemma-7b-it | Chat | global | $0.05 / 1M | $0.08 / 1M | 8.2K | supported |
| Llama 3.1 8B Instant | llama-3.1-8b-instant | Chat | global | $0.05 / 1M | $0.08 / 1M | 128K | supported |
| Llama 3.3 70B Versatile | llama-3.3-70b-versatile | Chat | global | $0.59 / 1M | $0.79 / 1M | 128K | supported |
| Meta Llama Llama 4 Maverick 17B 128E Instruct | meta-llama/llama-4-maverick-17b-128e-instruct | Chat | global | $0.20 / 1M | $0.60 / 1M | 131.1K | supported |
| Meta Llama Llama 4 Scout 17B 16E Instruct | meta-llama/llama-4-scout-17b-16e-instruct | Chat | global | $0.11 / 1M | $0.34 / 1M | 131.1K | supported |
| Meta Llama Llama Guard 4 12B | meta-llama/llama-guard-4-12b | Chat | global | $0.20 / 1M | $0.20 / 1M | 8.2K | unknown |
| Moonshotai Kimi K2 Instruct 0905 | moonshotai/kimi-k2-instruct-0905 | Chat | global | $1.00 / 1M | $3.00 / 1M | 262.1K | supported |
| gpt-oss-120b | openai/gpt-oss-120b | Chat | global | $0.15 / 1M | $0.60 / 1M | 131.1K | supported |
| gpt-oss-20b | openai/gpt-oss-20b | Chat | global | $0.07 / 1M | $0.30 / 1M | 131.1K | supported |
| Gpt Oss Safeguard 20B | openai/gpt-oss-safeguard-20b | Chat | global | $0.07 / 1M | $0.30 / 1M | 131.1K | supported |
| Playai TTS | playai-tts | Audio Speech | global | — | — | 10K | not_applicable |
| Qwen Qwen3 32B | qwen/qwen3-32b | Chat | global | $0.29 / 1M | $0.59 / 1M | 131K | supported |
| Whisper Large | whisper-large-v3 | Audio Transcription | global | — | — | — | not_applicable |
| Whisper Large V3 Turbo | whisper-large-v3-turbo | Audio Transcription | global | — | — | — | not_applicable |