Models · Hardware · Providers
Ask anything about running AI —
then see the data behind the answer.
Every answer is grounded in our live database of 455 models, 71 devices from data-centre GPUs to phones, and 1,854 provider price listings. We re-read each listing on its own schedule — the newest 2 hours ago, the oldest 2 months ago.
no sign-in · answers grounded in live catalogue data
01455 entries
Models
Every open model, fully specced.
VRAM per quant, estimated tokens/sec, benchmarks, licenses — synced from primary sources daily.
0271 entries
Hardware
GPUs, Macs and phones, ranked for AI.
What fits, how fast it runs, and what to buy next — from an RTX 5090 down to the phone in your pocket.
0385 providers
Providers
Hosted inference, priced daily.
1854 live offers with per-token pricing, served quants, context limits and privacy terms.
Latest releases & changes
All news →Host Mancer 2 cut gpt-oss-120b input pricing by 10%2 hours agoPareto 26.10 Preview listed8 hours agoGLM 5.3 Flash repriced across 2 hosts: InferenceNet input down 50%, Wafer output up 50%20 hours agoHost Relace raised GLM 5.3 input and cache-read pricing by 12%20 hours agoGLM 5.2 repriced across 5 hosts: Wafer input down 71%, Inceptron input up 85%20 hours agoHost Ionstream cut Qwen3.8 27B input pricing by 55%20 hours ago