Models · Hardware · Providers
Ask anything about running AI —
then see the data behind the answer.
Every answer is grounded in our live database of 376 models, 71 devices from data-centre GPUs to phones, and provider pricing synced 2h ago.
01376 entries
Models
Every open model, fully specced.
VRAM per quant, estimated tokens/sec, benchmarks, licenses — synced from primary sources daily.
0271 entries
Hardware
GPUs, Macs and phones, ranked for AI.
What fits, how fast it runs, and what to buy next — from an RTX 5090 down to the phone in your pocket.
0379 providers
Providers
Hosted inference, priced daily.
1544 live offers with per-token pricing, served quants, context limits and privacy terms.
Latest releases & changes
All news →Chutes cut GLM 5.1 pricing by 80%2h agoChutes cut Qwen3.5 397B A17B pricing by 80%2h agoDeepInfra raised DeepSeek V3.1 Terminus pricing by 8%2h agoChutes cut Kimi K2.6 pricing by 80%2h agoQwen3.6 35B A3B cut across 2 hosts, by up to 5% at Venice4h agoStreamLake raised DeepSeek V4 Pro pricing by 29%4h ago