Qwen3.5-Flash
Qwen · released Feb 25, 2026
- Type
- Proprietary
- Input
- None held
- Output
- None held
- Cached
- None held
We don't hold a list price for this model yet · hosted only — we hold no downloadable copy
Our take
Written Aug 3, 2026Qwen 3.5 Flash is a proprietary multimodal model from Alibaba that handles up to one million tokens in a single request across text, images and video. It carries a flat rate at every provider we track and scores strongest on hard prompts and coding relative to its own profile.
Choose this for long-context multimodal tasks at one million tokens, coding workloads where its Arena Coding score applies, or hard-prompt reasoning tasks. It suits budget-sensitive inference with zero provider price spread. Skip it if you need web-development coding specifically, creative writing, or a model whose size and licence terms are disclosed.
The case for it
- One-million-token request limit — ten times the threshold common in current offerings.
- Identical pricing across all three tracked providers, with no variation to shop around.
- Hard prompts and coding exceed its own overall chat score by 19 and 40 points respectively.
The case against it
- Web development coding lags its general coding score by 199 points — its widest internal gap.
- Creative writing is its weakest measured category, 59 points below its own overall score.
- Parameter count undisclosed and no licence terms published, so scale and usage rights are opaque.
How good is it?
IntelligencePuzzles, maths, exam questions
Arena Text (overall)77th of 143 · 1396.8
CodingWriting and fixing code on its own
Arena Coding80th of 143 · 1437.2
AgenticPlanning, calling tools, staying on task
Nobody we watch has scored Qwen3.5-Flash for this. We would take the rating from Arena Agent (IPS).
WritingWe do not rate this
Two boards come close and neither tests writing: Arena Creative Writing asks people which of two replies they prefer, and LiveBench Language tests whether a model understood a passage. So we show where Qwen3.5-Flash placed and give it no mark out of five.
These tests check whether a model follows instructions — a precondition for all the work above, but not a measure of how well that work is done, which is why they get no rating.
Every published score for this model7 scoresEvery figure we hold, from 7 boards, with who ran it and a link to the source — including the boards no rating above is built on.
Or rent it from someone else
Cheapest of 3 live listings. Picked at the widest standard context we hold, within one quantisation slice, so the numbers beside it are a price one host actually charges.
- per 1M tokens
- $0.065 in / $0.26 out
- Context served
- 1M
- Throughput
- Not measured
| Provider | In / out per 1M tokens | Context | Throughput | Trains on prompts | Logs prompts | Zero retention |
|---|---|---|---|---|---|---|
| OpenRouter | $0.065 / $0.26 | 1M | not measured | Unknown | Unknown | Unknown |
| Alibaba Cloud | $0.065 / $0.26 | 1M | 53 tok/s | No | Yesunknown period | Unknown |
| Alibaba Cloudfp8 | $0.065 / $0.26 | 1M | 54 tok/s | No | Yesunknown period | Unknown |
Across the 3 listings we hold: 2 say they do not train on prompts, 0 say they do and 1 do not say. 0 appear in the zero-retention registry we check; the rest are unknown to us rather than confirmed either way.
What each host's API supports
From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.
- ✓
- Supported
- ✗
- Not supported
- Not published
- host gave no parameter list
| Provider | Tool calling | JSON output | Strict schema |
|---|---|---|---|
| OpenRouter | ✓ | ✓ | ✓ |
| Alibaba Cloud | ✓ | ✓ | ✓ |
| Alibaba Cloudfp8 | ✓ | ✓ | ✓ |
Tool calling: 3 of 3 listings say yes. JSON output: 3 of 3 listings say yes. Strict schema: 3 of 3 listings say yes.
Models people weigh against Qwen3.5-Flash
When we formed this view
Dates behind this page
Prices last checked 9d ago
What we do not know about this model yet
- Nothing we hold says whether an endpoint streams, so we do not show it either way.
- 1 of 3 listings do not say whether they train on prompts.
- We don't hold a list price for this model yet — the gap is ours, not the lab's.
- We hold no cached-input rate for any of its listings.
Licence and identifiers
What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.
Licence
Commercial API terms. We hold no licence record for this model, so there is nothing to summarise here.
Identifiers
- Modality record
- text+image+video->text
- Catalogue slug
- qwen-qwen3-5-flash