Host Wafer cut DeepSeek V4 Flash input pricing by 30%
Published 27 September 2026
Price move, per 1M tokens
- Wafer · input↓ 30%$0.0590 → $0.0413
- Wafer · outputunchanged$0.35 → $0.35
- Wafer · cache read↓ 36%$0.058 → $0.037
| host | rate | was | now | change |
|---|---|---|---|---|
| Wafer | input | $0.0590 | $0.0413 | ↓ 30% |
| Wafer | output | $0.35 | $0.35 | — |
| Wafer | cache read | $0.058 | $0.037 | ↓ 36% |
input: ↓ 30%
If you buy
input and cache reads now cost less on Wafer, so this model suits steady, prompt-heavy work; re-check your cache-hit rate before your next top-up.
The model
DeepSeek V4 Flashopen the model →
Open weights290.9B parameters1049K context33 hosts
intelligence65th of 168 writing57th of 168 coding64th of 168 agents20th of 55
DeepSeek V4 Flash is a downloadable text model with a permissive licence, built for long-document work where the bill matters more than peak quality. Its measured quality is weak on the LiveBench boards and mid-pack on the Arena boards, so treat it as a cheap workhorse rather than a quality leader.
Source
Read on openrouter.ai · we published this on 27 September 2026.
read the original at openrouter.ai ↗