Host Relace cut GLM 5.3 Flash input pricing by 30%
Published 24 September 2026
Price move, per 1M tokens
- Relace · input↓ 30%$0.100 → $0.070
- Relace · output↓ 22%$0.36 → $0.28
- Relace · cache readunchanged$0.020 → $0.020
| host | rate | was | now | change |
|---|---|---|---|---|
| Relace | input | $0.100 | $0.070 | ↓ 30% |
| Relace | output | $0.36 | $0.28 | ↓ 22% |
| Relace | cache read | $0.020 | $0.020 | — |
input: ↓ 30%
If you buy
a cheaper floor for anyone running high-volume prompts through this host, so re-check your spend cap and whether your workload still fits the smaller output budget.
The model
GLM 5.3 Flashopen the model →
Open weights321.3B parameters1311K context35 hosts
intelligence26th of 168 writing43rd of 168 coding22nd of 168 agents27th of 55
GLM 5.3 Flash is a downloadable model you can run yourself, and it is at its best when the job is agentic: picking the right tool and finishing the task. It is at its worst when the job is holding to a format under instruction, and we list no licence for it, so the terms need checking at the source.
Source
Read on openrouter.ai · we published this on 24 September 2026.
read the original at openrouter.ai ↗