Catalogue
Models
376 models, cross-linked to hardware verdicts and live provider pricing.
287 models
Qwen3.73 sizes here
Qwen3.7 FlashQwenQwen3.7 Flash is a hosted-only multimodal model from Alibaba that can handle up to one million tokens in a single request.Jul 2026—1M$0.030 / $0.13per 1Mvia OpenRouter—Qwen3.7 PlusQwenQwen 3.7 is a proprietary text-and-image model from Alibaba with a one-million-token request limit and measured coding strength.Jun 2026—1M$0.32 / $1.28per 1Mvia OpenRouter—Qwen3.7 MaxQwenQwen 3.7 Max is Alibaba's proprietary flagship text model, released in May 2026 with a one-million-token request limit and strong measured coding performance.May 2026—1M$1.25 / $3.75per 1Mvia Novita AI—Inkling2 sizes here
Inkling SmallThinking MachinesInkling Small is a large downloadable model from Thinking Machines that accepts text, images and audio, and can handle up to 524,288 tokens in a single request.Jul 2026266B524K$0.50 / $1.20per 1Mvia OpenRouter—InklingThinking MachinesInkling is a 952.4-billion-parameter downloadable model from Thinking Machines with a one-million-token request limit and broad multimodal input.Jul 2026952B1M$1.00 / $4.05per 1Mvia OpenRouter—Claude15 sizes here
Claude Opus 5AnthropicAnthropic's largest Claude model is built for demanding reasoning and coding tasks.Jul 2026—1M$5.00 / $25.00per 1Mvia OpenRouter—Claude Opus 5 (Fast)AnthropicClaude Opus 5 (Fast) is Anthropic's premium hosted model for deep document analysis and complex multimodal tasks, with a one-million-token request limit that exceeds most frontier alternatives. It accepts text, images and files, but carries a steep per-token cost and no measured quality data in our records.Jul 2026—1M$10.00 / $50.00per 1Mvia OpenRouter—Claude Sonnet 5AnthropicAnthropic's Claude Sonnet 5 is the middle model in the Claude line.Jun 2026—1M$2.00 / $10.00per 1Mvia OpenRouter—Claude Opus 4.8AnthropicClaude Opus 4.8 is Anthropic's flagship reasoning model that can handle up to one million tokens in a single request, with measured top-quartile scores across coding, hard prompts and instruction following. It is available only through hosted channels and sits at a premium price point with no budget tier.May 2026—1M$5.00 / $25.00per 1Mvia OpenRouter—Claude Opus 4.8 (Fast)AnthropicClaude Opus 4.8 (Fast) is Anthropic's premium reasoning model that can handle up to one million tokens in a single request.May 2026—1M$10.00 / $50.00per 1Mvia OpenRouter—Claude Opus 4.7 (Fast)AnthropicClaude Opus 4.7 (Fast) is Anthropic's premium hosted model that can handle up to one million tokens in a single request, accepting text, images and files.May 2026—1M$30.00 / $150.00per 1Mvia OpenRouter—Claude Opus 4.7AnthropicClaude Opus 4.7 is Anthropic's flagship hosted-only model that can handle up to one million tokens in a single request.Apr 2026—1M$5.00 / $25.00per 1Mvia OpenRouter—Claude Sonnet 4.6AnthropicClaude Sonnet 4.6 is a proprietary text-and-image model from Anthropic that can handle up to one million tokens in a single request.Feb 2026—1M$3.00 / $15.00per 1Mvia OpenRouter—Claude Opus 4.6AnthropicClaude Opus 4.6 is Anthropic's flagship hosted-only reasoning model with a one-million-token request limit and top scores on the independent coding and hard-prompt leaderboards we track. It is built for demanding long-document and coding work, but costs noticeably more than most alternatives and cannot be downloaded or fine-tuned.Feb 2026—1M$5.00 / $25.00per 1Mvia OpenRouter—Claude Opus 4.5AnthropicClaude Opus 4.5 is Anthropic's flagship reasoning model, available only through hosted APIs with a 200,000-token request limit.Nov 2025—200K$5.00 / $25.00per 1Mvia OpenRouter—Claude Haiku 4.5AnthropicClaude Haiku 4.5 is Anthropic's fast, lightweight entry point that can handle up to 200,000 tokens in a single request and accepts text, images and files.Oct 2025—200K$1.00 / $5.00per 1Mvia OpenRouter—Claude Sonnet 4.5AnthropicClaude Sonnet 4.5 is Anthropic's proprietary text-and-image model that can handle up to one million tokens in a single request.Sept 2025—1M$3.00 / $15.00per 1Mvia OpenRouter—Claude Opus 4.1AnthropicClaude Opus 4.1 is Anthropic's flagship hosted-only reasoning model with a 200,000-token working memory and measured strengths in coding and instruction following.Aug 2025—200K$15.00 / $75.00per 1Mvia OpenRouter—Claude Opus 4AnthropicClaude Opus 4 is Anthropic's flagship reasoning model, available only through hosted APIs with a 200,000-token request limit.May 2025—200K$15.00 / $75.00per 1Mvia OpenRouter—Claude 3 HaikuAnthropicClaude 3 Haiku is Anthropic's fastest, most affordable text-and-image model, built for high-volume work with a 200,000-token request limit.Mar 2024—200K$0.25 / $1.25per 1Mvia OpenRouter—Gemini11 sizes here
Gemini 3.6 FlashGoogleGoogle's fast top-tier chat model currently tops the independent chat leaderboard we track.Jul 2026—1M$1.50 / $7.50per 1Mvia OpenRouter—Gemini 3.5 Flash LiteGoogleGemini 3.5 Flash Lite is a lightweight multimodal model from Google that can handle up to one million tokens in a single request and accepts text, images, files, audio and video.Jul 2026—1M$0.30 / $2.50per 1Mvia OpenRouter—Gemini 3.5 FlashGoogleGemini 3.5 Flash is a general-purpose multimodal model from Google that can handle up to one million tokens in a single request and accepts text, images, files, audio and video.May 2026—1M$1.50 / $9.00per 1Mvia OpenRouter—Gemini 3.1 Flash LiteGoogleGemini 3.1 Flash Lite is a lightweight multimodal model from Google that can handle up to one million tokens in a single request and accepts text, images, files, audio and video.May 2026—1M$0.25 / $1.50per 1Mvia OpenRouter—Gemini 3.1 Flash Lite PreviewGoogleGemini 3.1 Flash Lite Preview is a lightweight hosted model from Google that handles up to one million tokens in a single request and accepts text, images, files, audio and video.Mar 2026—1M$0.25 / $1.50per 1Mvia OpenRouter—Gemini 3.1 Pro Preview Custom ToolsGoogleGemini 3.1 Pro Preview Custom Tools is Google's flagship multimodal model with a one-million-token request limit and built-in tool support.Feb 2026—1M$2.00 / $12.00per 1Mvia OpenRouter—Gemini 3 Flash PreviewGoogleGemini 3 Flash Preview is a hosted-only multimodal model from Google that accepts text, images, files, audio and video in a single request of up to one million tokens.Dec 2025—1M$0.50 / $3.00per 1Mvia OpenRouter—Gemini 2.5 Flash LiteGoogleGemini 2.5 Flash Lite is a lightweight multimodal model from Google that can handle up to one million tokens in a single request and accepts text, images, files, audio and video.Jul 2025—1M$0.10 / $0.40per 1Mvia OpenRouter—Gemini 2.5 FlashGoogleGemini 2.5 Flash is a general-purpose model from Google that handles text, images, files, audio and video in a single request of up to one million tokens.Jun 2025—1M$0.30 / $2.50per 1Mvia OpenRouter—Gemini 2.5 ProGoogleGemini 2.5 Pro is Google's flagship reasoning model that accepts text, images, files, audio and video in a single request, with a one-million-token limit.Jun 2025—1M$1.25 / $10.00per 1Mvia OpenRouter—Gemini 2.5 Pro Preview 05-06GoogleGemini 2.5 Pro Preview 05-06 is Google's flagship multimodal model that can handle up to one million tokens in a single request, accepting text, images, files, audio and video.May 2025—1M$0.63 / $5.00per 1Mvia Google Vertex AI—Muse Spark 1.1MetaMuse Spark is Meta's hosted-only flagship that accepts text, images, files, audio and video in a single request of up to one million tokens.Jul 2026—1M$1.25 / $4.25per 1Mvia OpenRouter—Laguna S 2.1poolsideLaguna S 2.1 is a 118-billion-parameter text model from poolside with a one-million-token request limit and a restricted open licence.Jul 2026118B1M$0.090 / $0.18per 1Mvia OpenRouter—KAT-Coder-Air V2.5KwaipilotKAT-Coder-Air V2.5 is a hosted-only coding model from Kwaipilot that can handle up to 256,000 tokens in a single request.Jul 2026—256K$0.15 / $0.60per 1Mvia OpenRouter—KAT-Coder-Pro2 sizes here
KAT-Coder-Pro V2.5KwaipilotKAT-Coder-Pro is a proprietary coding model from Kwaipilot that handles up to 256,000 tokens in a single request.Jul 2026—256K$0.74 / $2.96per 1Mvia OpenRouter—KAT-Coder-Pro V2KwaipilotKAT-Coder-Pro V2 is Kwaipilot's hosted-only coding model built for long-context code tasks, handling up to 262,144 tokens in a single request.Mar 2026—262K$0.30 / $1.20per 1Mvia OpenRouter—GPT-5.6 Terra2 sizes here
GPT-5.6 TerraOpenAIGPT-5.6 Terra is OpenAI's flagship hosted model that can handle up to 1.05 million tokens in a single request, accepting text, images and files.Jul 2026—1.1M$1.00 / $6.00per 1Mvia OpenRouter—GPT-5.6 Terra ProOpenAIGPT-5.6 Terra Pro is OpenAI's proprietary flagship that can handle up to 1.05 million tokens in a single request, accepting text, images and files.Jul 2026—1.1M$1.00 / $6.00per 1Mvia OpenRouter—GPT-5.6 Luna2 sizes here
GPT-5.6 Luna ProOpenAIGPT-5.6 Luna Pro is a proprietary multimodal model from OpenAI that handles text, images and files across a one-million-token request limit.Jul 2026—1.1M$0.50 / $3.00per 1Mvia OpenRouter—GPT-5.6 LunaOpenAIGPT-5.6 Luna is OpenAI's flagship text model that can handle over one million tokens in a single request, released in July 2026.Jul 2026—1.1M$0.22 / $1.32per 1Mvia Amazon Bedrock43 tok/sGPT-5.6 Sol2 sizes here
GPT-5.6 Sol ProOpenAIGPT-5.6 Sol Pro is OpenAI's hosted-only flagship that can handle over one million tokens in a single request, accepting text, images and files.Jul 2026—1.1M$5.00 / $30.00per 1Mvia OpenRouter—GPT-5.6 SolOpenAIOpenAI's top-tier reasoning and coding model can handle a 1,050,000-token request.Jul 2026—1.1M$5.00 / $30.00per 1Mvia OpenRouter—Grok4 sizes here
Grok 4.5xAIxAI's top-tier chat model placed second on Arena Text, the independent human-preference board, in July 2026.Jul 2026—500K$2.00 / $6.00per 1Mvia OpenRouter—Grok 4.3xAIGrok 4.3 is xAI's proprietary flagship that handles text, images and files with a one-million-token request limit.Apr 2026—1M$1.25 / $2.50per 1Mvia OpenRouter—Grok 4.20xAIGrok 4.20 is xAI's flagship hosted model that accepts text, images and files across a two-million-token request limit.Mar 2026—2M$1.25 / $2.50per 1Mvia OpenRouter—Grok 4.20 Multi-AgentxAIGrok 4.20 Multi-Agent is xAI's proprietary flagship that handles text, images and files with a two-million-token request limit.Mar 2026—2M$1.25 / $2.50per 1Mvia OpenRouter—Aion-3.0AionLabsAion-3.0 is a proprietary text-only model from AionLabs with a 131,072-token request limit, released in July 2026.Jul 2026—131K$3.00 / $6.00per 1Mvia OpenRouter—Aion-3.0-MiniAionLabsAion-3.0-Mini is a proprietary text-only model from AionLabs that handles up to 131,072 tokens in a single request.Jul 2026—131K$0.70 / $1.40per 1Mvia OpenRouter—LongCat 2.0MeituanLongCat 2.0 is a 1.8-trillion-parameter text model from Meituan with a permissive MIT licence and a one-million-token request limit.Jul 20261.8T1M$0.30 / $1.20per 1Mvia OpenRouter—Hy31 sizes here
Hy3TencentHy3 is a large downloadable text model from Tencent with a permissive Apache licence and a 262,144-token request limit.Jul 2026299B262K$0.13 / $0.53per 1Mvia OpenRouter—1 / 6