News / PRICE

DeepSeek V4 Flash cut across 3 hosts, by up to 53% at Inceptron (input)

Published 26 September 2026
Price move, per 1M tokens
  • Inceptron · input↓ 53%
    $0.083 → $0.039
  • Inceptron · output↓ 44%
    $0.48 → $0.27
  • Inceptron · cache read↓ 40%
    $0.050 → $0.030
  • Sail Research · input↓ 28%
    $0.0300 → $0.0215
  • Sail Research · output↓ 45%
    $0.55 → $0.30
  • Sail Research · cache read↓ 13%
    $0.016 → $0.014
  • Relace · input↓ 30%
    $0.030 → $0.021
  • Relace · outputunchanged
    $0.32 → $0.32
  • Relace · cache readunchanged
    $0.016 → $0.016
  • Sail Research (US region) · input↓ 25%
    $0.0300 → $0.0225
  • Sail Research (US region) · output↓ 24%
    $0.55 → $0.42
  • Sail Research (US region) · cache read↓ 25%
    $0.016 → $0.012
Inceptron input: ↓ 53%Sail Research output: ↓ 45%Relace input: ↓ 30%Sail Research (US region) input and cache read: ↓ 25%
If you buy

Existing users at this host should re-check their usage and caching setup, since the lower rates now suit steady, high-volume work that was previously harder to justify.

The model
DeepSeek V4 Flashopen the model →
Open weights290.9B parameters1049K context33 hosts
intelligence65th of 168 writing57th of 168 coding64th of 168 agents20th of 55

DeepSeek V4 Flash is a downloadable text model with a permissive licence, built for long-document work where the bill matters more than peak quality. Its measured quality is weak on the LiveBench boards and mid-pack on the Arena boards, so treat it as a cheap workhorse rather than a quality leader.

Source

Read on openrouter.ai · we published this on 26 September 2026.

read the original at openrouter.ai ↗