Hugging Face
87.5Excellent
The GitHub of machine learning: models, datasets, Spaces and inference.
ML PlatformsFreemiumMature
From Free / Pro $9 per month / Enterprise from $20 per user
by Groq · founded 2016 · updated Jul 2026
Groq serves open models like Llama, DeepSeek and Qwen at hundreds of tokens per second on its custom Language Processing Units. With an OpenAI-compatible API and aggressive pricing, it is the go-to inference platform when latency and cost per token matter most.
Fixed weights, sources attached, the formula
Each score is built from the claims below. This is a snapshot, so check the current terms with the vendor before you commit to anything.
| Claim | Source | Checked |
|---|---|---|
| LPU inference serves open models at several hundred tokens per second | Groq benchmark publications | Jun 2026 |
Hugging Face
The GitHub of machine learning: models, datasets, Spaces and inference.
Ollama, Inc.
Run open LLMs locally with one command.
Together AI
Fast inference and fine-tuning for open-source models.