GroqCloud logo

GroqCloud

LLM 5.3/10

Very fast hosted inference for open models, with free account-level limits.

GroqCloud scores 5.3/10 for free value in LLM. The main constraint: Free plan limits are shown per organization/model in Groq Console; measured as RPM/RPD/TPM/TPD.

Category LLM Kind Fast inference Caveats Exact current limits are account/model-specific; read response headers and console limits. Renewal daily Card required No Eligibility anyone

Free value score

5.3 / 10
Show the math
Model Quality 4.3/10
value = 24 artificial_analysis_score
weighted score = 0.33 * 4.3/10
Throughput 2.6/10
value = 6000 tok/min
weighted score = 0.17 * 2.6/10
Request Rate 4/10
value = 30 req/min
weighted score = 0.17 * 4/10
Requests 6.3/10
value = 432000 req/mo
weighted score = 0.17 * 6.3/10
Access 10/10
value = 10/10
weighted score = 0.17 * 10/10

Key metrics

openai/gpt-oss-120b 24 Best model
6.0K T/min Tokens/min
14K R/day Requests/day
30 R/min Requests/min
Yes Commercial use

Models

16 free · 19 total
Model Family AA Access
openai/gpt-oss-120b GPT 23.8 free
openai/gpt-oss-20b GPT 14.9 free
meta-llama/llama-4-scout-17b-16e-instruct Llama 10 free
allam-2-7b ALLaM free
canopylabs/orpheus-arabic-saudi Orpheus (Canopy Labs) free
canopylabs/orpheus-v1-english Orpheus (Canopy Labs) free
groq/compound Groq (system) free
groq/compound-mini Groq (system) free
llama-3.1-8b-instant Llama free
llama-3.3-70b-versatile Llama free
meta-llama/llama-prompt-guard-2-22m Llama free
meta-llama/llama-prompt-guard-2-86m Llama free
openai/gpt-oss-safeguard-20b GPT free
qwen/qwen3-32b Qwen free
whisper-large-v3 Whisper free
whisper-large-v3-turbo Whisper free
Minimax M2.5 MiniMax paid
Qwen3-VL 32B Qwen paid

What you get

Free plan exists

Units: RPM / RPD / TPM / TPD

Limits per organization/model

Very low latency inference

Sources

fast free-tier open-models