Providers /Groq
Inference provider · HQ US

Groq

Models
7
Data residency
No answer on record
00

Our take

Editorial by LLMap · updated Aug 4, 2026

Groq is an inference provider that uses its own custom inference chips instead of standard graphics chips, designed for speed. It offers a small, curated menu of downloadable models at aggressive prices.

Pick this for latency-critical products such as voice agents or interactive interfaces, on supported models. Use it for cheap small-model bulk inference. Skip it if you need a broad model catalogue, an EU data-residency endpoint, or a single accountable host.

Strengths
  • Purpose-built inference hardware designed for high tokens per second.
  • Among the lowest small-model prices around.
Trade-offs
  • Curated, limited catalogue; you adapt to its menu, not vice versa.
  • EU endpoint is unverified in our data.
01

Point your tools here

We have not recorded what a router needs for Groq yet — no base URL and no docs link on file. That is our gap, not a sign Groq has no API; its own documentation is the place to look until we close it.

02

Privacy & data handling

These answers cover requests sent to Groq through OpenRouter, as OpenRouter records them, last read 58 min ago. For Groq's own API we hold no answer: check its terms. SOC 2 and the data processing agreement below are hand-checked, not part of that sync.

Trains on prompts
✗No
Logs prompts
✗No
✓Offered
✓Attested

We have not dated this attestation or linked the report yet.

No answer on record

Prompt retention: none.

03

Models & pricing

Something wrong on this page? Tell us