Models · Hardware · Providers
Ask anything about running AI —
then see the data behind the answer.
Every answer is grounded in our live database of 455 models, 71 devices from data-centre GPUs to phones, and 1,853 provider price listings. We re-read each listing on its own schedule — the newest 6 hours ago, the oldest 2 months ago.
no sign-in · answers grounded in live catalogue data
01455 entries
Models
Every open model, fully specced.
VRAM per quant, estimated tokens/sec, benchmarks, licenses — synced from primary sources daily.
0271 entries
Hardware
GPUs, Macs and phones, ranked for AI.
What fits, how fast it runs, and what to buy next — from an RTX 5090 down to the phone in your pocket.
0385 providers
Providers
Hosted inference, priced daily.
1853 live offers with per-token pricing, served quants, context limits and privacy terms.
Latest releases & changes
All news →Pareto 26.10 Preview listed6 hours agoGLM 5.3 Flash repriced across 2 hosts: InferenceNet input down 58%, Wafer output up 50%18 hours agoHost Relace raised GLM 5.3 input and cache-read pricing by 12%18 hours agoGLM 5.2 repriced across 4 hosts: Wafer input down 71%, Inceptron input up 85%18 hours agoHost Ionstream cut Qwen3.8 27B input pricing by 53%18 hours agoKimi K3 repriced across 5 hosts: Relace input down 69%, Sail Research input up 781%18 hours ago