Host Relace raised DeepSeek V4.1 Flash cache-read pricing by 285%
Published 21 September 2026
Price move, per 1M tokens
- Relace · inputunchanged$0.13 → $0.13
- Relace · outputunchanged$0.52 → $0.52
- Relace · cache read↑ 285%$0.0026 → $0.0100
| host | rate | was | now | change |
|---|---|---|---|---|
| Relace | input | $0.13 | $0.13 | — |
| Relace | output | $0.52 | $0.52 | — |
| Relace | cache read | $0.0026 | $0.0100 | ↑ 285% |
cache read: ↑ 285%
If you buy
If you lean on cached context for long chats or repeated prompts, this raises what that habit costs here, so re-check your usage pattern before your next top-up.
The model
DeepSeek V4.1 Flashopen the model →
Open weights763.2B parameters1049K context35 hosts
intelligence21st of 168 writing37th of 168 coding11th of 168 agents13th of 55
DeepSeek V4.1 Flash is a downloadable model with a permissive licence, and it places 11th of 168 on Arena Coding via Max as of 25 Sep 2026. At 763.2 billion parameters it is a data-centre job to run yourself, so hosted use is the practical route for most readers.
Source
Read on openrouter.ai · we published this on 21 September 2026.
read the original at openrouter.ai ↗