Our take
Groq is an inference provider that uses its own custom inference chips instead of standard graphics chips, designed for speed. It offers a small, curated menu of downloadable models at aggressive prices.
Pick this for latency-critical products such as voice agents or interactive interfaces, on supported models. Use it for cheap small-model bulk inference. Skip it if you need a broad model catalogue, an EU data-residency endpoint, or a single accountable host.
- Purpose-built inference hardware designed for high tokens per second.
- Among the lowest small-model prices around.
- Curated, limited catalogue; you adapt to its menu, not vice versa.
- EU endpoint is unverified in our data.
Point your tools here
We have not recorded what a router needs for Groq yet — no base URL and no docs link on file. That is our gap, not a sign Groq has no API; its own documentation is the place to look until we close it.
Privacy & data handling
These answers cover requests sent to Groq through OpenRouter, as OpenRouter records them, last read 58 min ago. For Groq's own API we hold no answer: check its terms. SOC 2 and the data processing agreement below are hand-checked, not part of that sync.
We have not dated this attestation or linked the report yet.
Prompt retention: none.