Feed
News
Scanned every 15 minutes from vendor feeds, Hugging Face and provider APIs. New releases land here within 24 hours of announcement.
You're reading the changes — get them weekly.
Every Thursday: releases, price moves and new hardware, straight from this feed.
3 August 2026
PRICEChutes cut GLM 5.1 pricing by 80% $3.08 → $3.08cache read −80% ($0.49 → $0.098 per 1M tokens)PRICEChutes cut Qwen3.5 397B A17B pricing by 80% $3.00 → $3.00cache read −80% ($0.23 → $0.045 per 1M tokens)PRICEDeepInfra raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $0.95input +8% ($0.25 → $0.27 per 1M tokens)PRICEChutes cut Kimi K2.6 pricing by 80% $3.50 → $3.50cache read −80% ($0.33 → $0.066 per 1M tokens)PRICEQwen3.6 35B A3B cut across 2 hosts, by up to 5% at VeniceVenice: $1.00 → $0.95Io Net: $2.09 → $1.89Qwen3.6 35B A3B moved on 2 hosts: Venice: input −2% ($0.10 → $0.098 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); Io Net: input −11% ($0.32 → $0.29 per 1M tokens); output −10% ($2.09 → $1.89 per 1M tokens); cache read −32% ($0.16 → $0.11 per 1M tokens)PRICEStreamLake raised DeepSeek V4 Pro pricing by 29% $0.95 → $1.22input +29% ($0.47 → $0.61 per 1M tokens); output +29% ($0.95 → $1.22 per 1M tokens); cache read +29% ($0.039 → $0.051 per 1M tokens)PRICEOpenRouter cut Qwen3 Coder 30B A3B Instruct pricing by 4% $0.28 → $0.27output −4% ($0.28 → $0.27 per 1M tokens)PRICEQwen3.6 27B repriced across 2 hosts, from a 20% rise at OpenRouter to a 7% cut at Io NetOpenRouter: $2.00 → $2.40Io Net: $1.99 → $1.89Qwen3.6 27B moved on 2 hosts: OpenRouter: input −4% ($0.30 → $0.29 per 1M tokens); output +20% ($2.00 → $2.40 per 1M tokens); Io Net: input −4% ($0.28 → $0.27 per 1M tokens); output −5% ($1.99 → $1.89 per 1M tokens); cache read −7% ($0.14 → $0.13 per 1M tokens)PRICEDeepSeek V4 Flash repriced across 2 hosts, from an 8% rise at SiliconFlow to a 10% cut at Io NetSiliconFlow: $0.28 → $0.28Io Net: $0.34 → $0.33DeepSeek V4 Flash moved on 2 hosts: SiliconFlow: input +8% ($0.13 → $0.14 per 1M tokens); Io Net: input −6% ($0.18 → $0.17 per 1M tokens); output −3% ($0.34 → $0.33 per 1M tokens); cache read −10% ($0.080 → $0.072 per 1M tokens)PRICEDeepInfra cut DeepSeek V3.1 Terminus pricing by 7% $0.95 → $0.95input −7% ($0.27 → $0.25 per 1M tokens)PRICENextBit cut Gemma 4 26B A4B pricing by 8% $0.40 → $0.40input −8% ($0.13 → $0.12 per 1M tokens)ANNOUNCEMENTGroq: Groq Raises 750 Million As Inference Demand SurgesAnnounced by GroqANNOUNCEMENTGroq: Groq Names Simon Edwards Chief Financial OfficerAnnounced by GroqPRICEGLM 5.2 cut across 6 hosts, by up to 10% at PhalaPhala: $4.40 → $3.96StreamLake: $2.23 → $1.98OpenRouter: $2.22 → $1.98Novita: $2.22 → $1.98Decart: $1.80 → $1.50Chutes: $3.95 → $3.95GLM 5.2 moved on 6 hosts: Phala: input −10% ($1.40 → $1.26 per 1M tokens); output −10% ($4.40 → $3.96 per 1M tokens); cache read −10% ($0.26 → $0.23 per 1M tokens); StreamLake: input −11% ($0.71 → $0.63 per 1M tokens); output −11% ($2.23 → $1.98 per 1M tokens); cache read −11% ($0.13 → $0.12 per 1M tokens); OpenRouter: input −11% ($0.71 → $0.63 per 1M tokens); output −11% ($2.22 → $1.98 per 1M tokens); cache read −11% ($0.13 → $0.12 per 1M tokens); Novita: input −11% ($0.71 → $0.63 per 1M tokens); output −11% ($2.22 → $1.98 per 1M tokens); cache read −11% ($0.13 → $0.12 per 1M tokens); Decart: input −17% ($0.72 → $0.60 per 1M tokens); output −17% ($1.80 → $1.50 per 1M tokens); cache read −17% ($0.12 → $0.10 per 1M tokens); Chutes: cache read −80% ($0.63 → $0.13 per 1M tokens)PRICEOpenRouter raised Qwen3 235B A22B Instruct 2507 pricing by 66% $0.55 → $0.60input +66% ($0.090 → $0.15 per 1M tokens); output +9% ($0.55 → $0.60 per 1M tokens)PRICEOpenRouter raised Qwen3.5-122B-A10B pricing by 54% $2.08 → $3.20input +54% ($0.26 → $0.40 per 1M tokens); output +54% ($2.08 → $3.20 per 1M tokens)PRICEStreamLake cut DeepSeek V4 Flash pricing by 6% $0.19 → $0.18input −6% ($0.094 → $0.088 per 1M tokens); output −6% ($0.19 → $0.18 per 1M tokens); cache read −6% ($0.019 → $0.018 per 1M tokens)PRICEMancer 2 raised gpt-oss-120b pricing by 67% $0.50 → $0.50input +67% ($0.060 → $0.10 per 1M tokens)PRICEOpenRouter cut DeepSeek V3.1 pricing by 7% $1.00 → $0.95input −7% ($0.27 → $0.25 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); cache read −4% ($0.14 → $0.13 per 1M tokens)PRICEOpenRouter raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $1.00input +8% ($0.25 → $0.27 per 1M tokens); output +5% ($0.95 → $1.00 per 1M tokens); cache read +4% ($0.13 → $0.14 per 1M tokens)
2 August 2026
PRICEDeepInfra cut Inkling pricing by 6% $4.05 → $4.05input −5% ($1.00 → $0.95 per 1M tokens); cache read −6% ($0.17 → $0.16 per 1M tokens)ANNOUNCEMENTNVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical RoboticsAnnounced by Hugging FaceANNOUNCEMENTLFM2.5-Encoders for Fast Long-Context Inference on CPUAnnounced by Hugging FacePRICEDeepInfra raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $0.95input +8% ($0.25 → $0.27 per 1M tokens)PRICEDeepInfra cut Inkling pricing by 5% $4.05 → $4.05input −5% ($1.00 → $0.95 per 1M tokens)PRICEDeepInfra cut Inkling Small pricing by 10% $1.20 → $1.20input −10% ($0.50 → $0.45 per 1M tokens)PRICEIo Net raised DeepSeek V4 Flash pricing by 29% $0.28 → $0.34input +29% ($0.14 → $0.18 per 1M tokens); output +22% ($0.28 → $0.34 per 1M tokens); cache read +15% ($0.070 → $0.080 per 1M tokens)PRICEDeepInfra cut Inkling Small pricing by 10% $1.20 → $1.20input −10% ($0.50 → $0.45 per 1M tokens)PRICEDeepSeek V4 Pro cut across 2 hosts, by up to 16% at BaiduBaidu: $1.28 → $1.08StreamLake: $1.15 → $0.95DeepSeek V4 Pro moved on 2 hosts: Baidu: input −16% ($0.64 → $0.54 per 1M tokens); output −16% ($1.28 → $1.08 per 1M tokens); cache read −16% ($0.053 → $0.045 per 1M tokens); StreamLake: input −18% ($0.57 → $0.47 per 1M tokens); output −18% ($1.15 → $0.95 per 1M tokens); cache read −18% ($0.048 → $0.039 per 1M tokens)PRICEDeepSeek V4 Flash cut across 2 hosts, by up to 2% at BaiduBaidu: $0.18 → $0.18StreamLake: $0.18 → $0.17DeepSeek V4 Flash moved on 2 hosts: Baidu: input −2% ($0.090 → $0.088 per 1M tokens); output −2% ($0.18 → $0.18 per 1M tokens); cache read −2% ($0.018 → $0.018 per 1M tokens); StreamLake: input −3% ($0.090 → $0.087 per 1M tokens); output −3% ($0.18 → $0.17 per 1M tokens); cache read −3% ($0.018 → $0.017 per 1M tokens)PRICEOpenRouter cut DeepSeek V3.1 pricing by 7% $1.00 → $0.95input −7% ($0.27 → $0.25 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); cache read −4% ($0.14 → $0.13 per 1M tokens)PRICEOpenRouter raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $1.00input +8% ($0.25 → $0.27 per 1M tokens); output +5% ($0.95 → $1.00 per 1M tokens); cache read +4% ($0.13 → $0.14 per 1M tokens)
1 August 2026
PRICEDeepSeek V4 Pro repriced across 2 hosts, from a 3% rise at Baidu to a 7% cut at StreamLakeBaidu: $1.25 → $1.28StreamLake: $1.34 → $1.25DeepSeek V4 Pro moved on 2 hosts: Baidu: input +3% ($0.63 → $0.64 per 1M tokens); output +3% ($1.25 → $1.28 per 1M tokens); cache read +3% ($0.052 → $0.053 per 1M tokens); StreamLake: input −7% ($0.67 → $0.62 per 1M tokens); output −7% ($1.34 → $1.25 per 1M tokens); cache read −7% ($0.056 → $0.052 per 1M tokens)PRICEVenice cut GLM 4.7 Flash pricing by 52% $0.50 → $0.40input −52% ($0.13 → $0.060 per 1M tokens); output −20% ($0.50 → $0.40 per 1M tokens)PRICEIo Net raised MiMo-V2.5 pricing by 61% $0.26 → $0.32input +61% ($0.13 → $0.21 per 1M tokens); output +23% ($0.26 → $0.32 per 1M tokens); cache read +61% ($0.065 → $0.10 per 1M tokens)PRICEIo Net raised Qwen3.6 35B A3B pricing by 72% $1.22 → $2.09input +48% ($0.22 → $0.32 per 1M tokens); output +72% ($1.22 → $2.09 per 1M tokens)PRICEDeepInfra cut DeepSeek V3.1 Terminus pricing by 7% $0.95 → $0.95input −7% ($0.27 → $0.25 per 1M tokens)PRICEDeepInfra raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $0.95input +8% ($0.25 → $0.27 per 1M tokens)PRICEOpenRouter cut Mistral Small 3.2 24B pricing by 33% $0.30 → $0.20input −25% ($0.10 → $0.075 per 1M tokens); output −33% ($0.30 → $0.20 per 1M tokens)PRICEOpenRouter raised Qwen3 Coder 30B A3B Instruct pricing by 4% $0.27 → $0.28output +4% ($0.27 → $0.28 per 1M tokens)PRICEOpenRouter cut DeepSeek V3.1 pricing by 7% $1.00 → $0.95input −7% ($0.27 → $0.25 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); cache read −4% ($0.14 → $0.13 per 1M tokens)PRICEOpenRouter raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $1.00input +8% ($0.25 → $0.27 per 1M tokens); output +5% ($0.95 → $1.00 per 1M tokens); cache read +4% ($0.13 → $0.14 per 1M tokens)PRICEOpenRouter cut Qwen3 VL 30B A3B Instruct pricing by 13% $0.60 → $0.52input −13% ($0.15 → $0.13 per 1M tokens); output −13% ($0.60 → $0.52 per 1M tokens)PRICEOpenRouter cut GPT-5.1 pricing by 4% $10.00 → $10.00cache read −4% ($0.13 → $0.13 per 1M tokens)PRICEKimi K2.6 cut across 2 hosts, by up to 9% at BaiduBaidu: $2.72 → $2.48OpenRouter: $4.00 → $3.41Kimi K2.6 moved on 2 hosts: Baidu: input −9% ($0.65 → $0.59 per 1M tokens); output −9% ($2.72 → $2.48 per 1M tokens); cache read −9% ($0.11 → $0.099 per 1M tokens); OpenRouter: input −37% ($0.95 → $0.60 per 1M tokens); output −15% ($4.00 → $3.41 per 1M tokens); cache read +25% ($0.16 → $0.20 per 1M tokens)PRICEDeepSeek V4 Flash repriced across 4 hosts, from a 900% rise at OpenRouter to a 3% cut at StreamLakeOpenRouter: $0.28 → $0.28Mancer 2: $0.50 → $0.50Baidu: $0.18 → $0.18StreamLake: $0.18 → $0.18DeepSeek V4 Flash moved on 4 hosts: OpenRouter: cache read +900% ($0.003 → $0.028 per 1M tokens); Mancer 2: input +33% ($0.15 → $0.20 per 1M tokens); Baidu: input −2% ($0.091 → $0.090 per 1M tokens); output −2% ($0.18 → $0.18 per 1M tokens); cache read −2% ($0.018 → $0.018 per 1M tokens); StreamLake: input −3% ($0.092 → $0.090 per 1M tokens); output −3% ($0.18 → $0.18 per 1M tokens); cache read −3% ($0.018 → $0.018 per 1M tokens)PRICEGLM 5.2 repriced across 5 hosts, from a 24% rise at StreamLake to a 40% cut at DecartStreamLake: $1.93 → $2.39Baidu: $2.20 → $2.38Novita: $2.38 → $2.25OpenRouter: $2.39 → $2.25Decart: $2.50 → $1.80GLM 5.2 moved on 5 hosts: StreamLake: input +24% ($0.61 → $0.76 per 1M tokens); output +24% ($1.93 → $2.39 per 1M tokens); cache read +24% ($0.11 → $0.14 per 1M tokens); Baidu: input +8% ($0.70 → $0.76 per 1M tokens); output +8% ($2.20 → $2.38 per 1M tokens); cache read +8% ($0.13 → $0.14 per 1M tokens); Novita: input −6% ($0.76 → $0.72 per 1M tokens); output −6% ($2.38 → $2.25 per 1M tokens); cache read −6% ($0.14 → $0.13 per 1M tokens); OpenRouter: input −6% ($0.76 → $0.72 per 1M tokens); output −6% ($2.39 → $2.25 per 1M tokens); cache read −6% ($0.14 → $0.13 per 1M tokens); Decart: input −40% ($1.20 → $0.72 per 1M tokens); output −28% ($2.50 → $1.80 per 1M tokens); cache read −40% ($0.20 → $0.12 per 1M tokens)RELEASEInkling Small listedNew model detected on OpenRouter: thinkingmachines/inkling-smallPRICEOpenRouter cut DeepSeek V4 Flash 0731 pricing by 90% $0.28 → $0.28cache read −90% ($0.028 → $0.003 per 1M tokens)
31 July 2026
PRICEOpenAI cut GPT-5.6 Terra pricing by 20% $15.00 → $12.00input −20% ($2.50 → $2.00 per 1M tokens); output −20% ($15.00 → $12.00 per 1M tokens)PRICEDeepInfra cut DeepSeek V3.1 Terminus pricing by 7% $0.95 → $0.95input −7% ($0.27 → $0.25 per 1M tokens)PRICEDeepInfra raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $0.95input +8% ($0.25 → $0.27 per 1M tokens)PRICEOpenRouter raised Kimi K2.6 pricing by 47% $2.72 → $4.00input +47% ($0.65 → $0.95 per 1M tokens); output +47% ($2.72 → $4.00 per 1M tokens); cache read +47% ($0.11 → $0.16 per 1M tokens)PRICEOpenRouter cut DeepSeek V4 Flash 0731 pricing by 90% $0.28 → $0.28cache read −90% ($0.028 → $0.003 per 1M tokens)PRICEOpenRouter cut Qwen3 235B A22B Thinking 2507 pricing by 23% $3.00 → $2.30input −23% ($0.30 → $0.23 per 1M tokens); output −23% ($3.00 → $2.30 per 1M tokens)PRICEStreamLake raised MiMo-V2.5-Pro pricing by 20% $0.87 → $1.04input +20% ($0.43 → $0.52 per 1M tokens); output +20% ($0.87 → $1.04 per 1M tokens); cache read +20% ($0.004 → $0.004 per 1M tokens)PRICENextBit cut Gemma 4 26B A4B pricing by 17% $0.50 → $0.45input −17% ($0.18 → $0.15 per 1M tokens); output −10% ($0.50 → $0.45 per 1M tokens)PRICEDeepSeek V4 Flash repriced across 2 hosts, from a 900% rise at OpenRouter to a 25% cut at Mancer 2OpenRouter: $0.28 → $0.28Mancer 2: $0.50 → $0.50DeepSeek V4 Flash moved on 2 hosts: OpenRouter: cache read +900% ($0.003 → $0.028 per 1M tokens); Mancer 2: input −25% ($0.20 → $0.15 per 1M tokens)PRICEOpenRouter cut DeepSeek V3.1 pricing by 7% $1.00 → $0.95input −7% ($0.27 → $0.25 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); cache read −4% ($0.14 → $0.13 per 1M tokens)PRICEOpenRouter raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $1.00input +8% ($0.25 → $0.27 per 1M tokens); output +5% ($0.95 → $1.00 per 1M tokens); cache read +4% ($0.13 → $0.14 per 1M tokens)PRICEOpenRouter raised GPT-5.1 Chat pricing by 4% $10.00 → $10.00cache read +4% ($0.13 → $0.13 per 1M tokens)PRICEOpenRouter cut GPT-5.1 pricing by 4% $10.00 → $10.00cache read −4% ($0.13 → $0.13 per 1M tokens)PRICEOpenRouter raised Nemotron 3 Nano 30B A3B pricing by 20% $0.20 → $0.20cache read +20% ($0.025 → $0.030 per 1M tokens)PRICEAmazon Bedrock raised GPT-5.6 Terra pricing by 10% $12.00 → $13.20input +10% ($2.00 → $2.20 per 1M tokens); output +10% ($12.00 → $13.20 per 1M tokens); cache read +10% ($0.20 → $0.22 per 1M tokens); cache write +10% ($2.50 → $2.75 per 1M tokens)PRICEAmazon Bedrock raised GPT-5.6 Sol pricing by 10% $30.00 → $33.00input +10% ($5.00 → $5.50 per 1M tokens); output +10% ($30.00 → $33.00 per 1M tokens); cache read +10% ($0.50 → $0.55 per 1M tokens); cache write +10% ($6.25 → $6.88 per 1M tokens)PRICEAmazon Bedrock raised GPT-5.6 Luna pricing by 10% $1.20 → $1.32input +10% ($0.20 → $0.22 per 1M tokens); output +10% ($1.20 → $1.32 per 1M tokens); cache read +10% ($0.020 → $0.022 per 1M tokens); cache write +10% ($0.25 → $0.28 per 1M tokens)PRICEAmazon Bedrock raised GPT-5.4 pricing by 10% $15.00 → $16.50input +10% ($2.50 → $2.75 per 1M tokens); output +10% ($15.00 → $16.50 per 1M tokens); cache read +10% ($0.25 → $0.28 per 1M tokens)PRICEAmazon Bedrock raised GPT-5.5 pricing by 10% $30.00 → $33.00input +10% ($5.00 → $5.50 per 1M tokens); output +10% ($30.00 → $33.00 per 1M tokens); cache read +10% ($0.50 → $0.55 per 1M tokens)PRICEIo Net raised Mistral Nemo pricing by 10% $0.15 → $0.17input +10% ($0.039 → $0.043 per 1M tokens); output +10% ($0.15 → $0.17 per 1M tokens); cache read +10% ($0.019 → $0.021 per 1M tokens)
30 July 2026
ANNOUNCEMENTAnthropic: Investigating Incidents Cybersecurity EvalsAnnounced by AnthropicPRICEMancer 2 raised gpt-oss-120b pricing by 9% $0.50 → $0.50input +9% ($0.055 → $0.060 per 1M tokens)PRICEMorph cut Kimi K3 pricing by 7% $15.00 → $14.00input −3% ($3.00 → $2.90 per 1M tokens); output −7% ($15.00 → $14.00 per 1M tokens); cache read −3% ($0.30 → $0.29 per 1M tokens)PRICEDeepSeek V4 Flash repriced across 2 hosts, from a 2% rise at StreamLake to a 50% cut at Mancer 2StreamLake: $0.18 → $0.18Mancer 2: $1.00 → $0.50DeepSeek V4 Flash moved on 2 hosts: StreamLake: input +2% ($0.091 → $0.092 per 1M tokens); output +2% ($0.18 → $0.18 per 1M tokens); cache read +2% ($0.018 → $0.018 per 1M tokens); Mancer 2: output −50% ($1.00 → $0.50 per 1M tokens)PRICEGPT-5.6 Terra cut across 2 hosts, by up to 20% at OpenRouterOpenRouter: $7.50 → $6.00OpenAI: $3.75 → $3.00GPT-5.6 Terra moved on 2 hosts: OpenRouter: input −20% ($1.25 → $1.00 per 1M tokens); output −20% ($7.50 → $6.00 per 1M tokens); cache read −20% ($0.13 → $0.10 per 1M tokens); cache write −20% ($1.56 → $1.25 per 1M tokens); OpenAI: input −20% ($0.63 → $0.50 per 1M tokens); output −20% ($3.75 → $3.00 per 1M tokens); cache read −20% ($0.063 → $0.050 per 1M tokens); cache write −20% ($0.78 → $0.63 per 1M tokens)PRICEGPT-5.6 Terra Pro cut across 2 hosts, by up to 20% at OpenRouterOpenRouter: $7.50 → $6.00OpenAI: $3.75 → $3.00GPT-5.6 Terra Pro moved on 2 hosts: OpenRouter: input −20% ($1.25 → $1.00 per 1M tokens); output −20% ($7.50 → $6.00 per 1M tokens); cache read −20% ($0.13 → $0.10 per 1M tokens); cache write −20% ($1.56 → $1.25 per 1M tokens); OpenAI: input −20% ($0.63 → $0.50 per 1M tokens); output −20% ($3.75 → $3.00 per 1M tokens); cache read −20% ($0.063 → $0.050 per 1M tokens); cache write −20% ($0.78 → $0.63 per 1M tokens)ANNOUNCEMENTAdvancing the price-performance frontier with GPT-5.6Announced by OpenAIPRICEOpenRouter cut Laguna S 2.1 pricing by 10% $0.20 → $0.18input −10% ($0.10 → $0.090 per 1M tokens); output −10% ($0.20 → $0.18 per 1M tokens); cache read −10% ($0.010 → $0.009 per 1M tokens)ANNOUNCEMENTGemini Robotics 2 brings whole body intelligence to robotsAnnounced by Google DeepMindANNOUNCEMENTGemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaborationAnnounced by Google DeepMindPRICEIo Net raised Mistral Nemo pricing by 15% $0.13 → $0.15input +11% ($0.035 → $0.039 per 1M tokens); output +15% ($0.13 → $0.15 per 1M tokens); cache read +11% ($0.018 → $0.019 per 1M tokens)PRICESambaNova raised MiniMax M2.7 pricing by 150% $2.40 → $1.50input −50% ($0.60 → $0.30 per 1M tokens); output −38% ($2.40 → $1.50 per 1M tokens); cache read +150% ($0.060 → $0.15 per 1M tokens)PRICENextBit raised Gemma 4 26B A4B pricing by 12% $0.48 → $0.50input +12% ($0.16 → $0.18 per 1M tokens); output +4% ($0.48 → $0.50 per 1M tokens)PRICEDeepInfra cut DeepSeek V3.1 Terminus pricing by 7% $0.95 → $0.95input −7% ($0.27 → $0.25 per 1M tokens)PRICEDeepInfra raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $0.95input +8% ($0.25 → $0.27 per 1M tokens)PRICEDeepInfra cut Qwen3.6 35B A3B pricing by 33% $0.95 → $0.95input −33% ($0.15 → $0.10 per 1M tokens)PRICEDeepInfra cut GLM 5.2 pricing by 20% $3.00 → $2.40input −19% ($0.93 → $0.75 per 1M tokens); output −20% ($3.00 → $2.40 per 1M tokens)PRICEOpenRouter cut MoonshotAI Kimi Latest pricing by 7% $15.00 → $14.00output −7% ($15.00 → $14.00 per 1M tokens)PRICEDeepSeek V3 rose across 2 hosts, by up to 29% at StreamLakeStreamLake: $0.80 → $1.03OpenRouter: $0.80 → $1.03DeepSeek V3 moved on 2 hosts: StreamLake: input +29% ($0.20 → $0.26 per 1M tokens); output +29% ($0.80 → $1.03 per 1M tokens); OpenRouter: input +29% ($0.20 → $0.26 per 1M tokens); output +29% ($0.80 → $1.03 per 1M tokens)PRICEOpenRouter cut DeepSeek V3.1 pricing by 7% $1.00 → $0.95input −7% ($0.27 → $0.25 per 1M tokens); output −5% ($1.00 → $0.95 per 1M tokens); cache read −4% ($0.14 → $0.13 per 1M tokens)PRICEOpenRouter raised DeepSeek V3.1 Terminus pricing by 8% $0.95 → $1.00input +8% ($0.25 → $0.27 per 1M tokens); output +5% ($0.95 → $1.00 per 1M tokens); cache read +4% ($0.13 → $0.14 per 1M tokens)PRICEOpenRouter raised Qwen3 VL 235B A22B Thinking pricing by 54% $2.60 → $4.00input +54% ($0.26 → $0.40 per 1M tokens); output +54% ($2.60 → $4.00 per 1M tokens)PRICEOpenRouter raised Qwen3 VL 30B A3B Instruct pricing by 15% $0.52 → $0.60input +15% ($0.13 → $0.15 per 1M tokens); output +15% ($0.52 → $0.60 per 1M tokens)PRICEOpenRouter raised GPT-5.1 Chat pricing by 4% $10.00 → $10.00cache read +4% ($0.13 → $0.13 per 1M tokens)PRICEOpenRouter cut GPT-5.1 pricing by 4% $10.00 → $10.00cache read −4% ($0.13 → $0.13 per 1M tokens)PRICEOpenRouter cut Gemma 4 31B pricing by 29% $0.40 → $0.34input −29% ($0.14 → $0.10 per 1M tokens); output −15% ($0.40 → $0.34 per 1M tokens)PRICEGLM 5.2 repriced across 6 hosts, from a 10% rise at Io Net to a 63% cut at PhalaIo Net: $7.28 → $8.01Novita: $2.20 → $1.93StreamLake: $2.20 → $1.93OpenRouter: $2.20 → $1.93DeepInfra: $3.00 → $2.40Phala: $4.40 → $4.40GLM 5.2 moved on 6 hosts: Io Net: input +10% ($3.33 → $3.66 per 1M tokens); output +10% ($7.28 → $8.01 per 1M tokens); cache read +10% ($1.66 → $1.83 per 1M tokens); Novita: input −12% ($0.70 → $0.61 per 1M tokens); output −12% ($2.20 → $1.93 per 1M tokens); cache read −12% ($0.13 → $0.11 per 1M tokens); StreamLake: input −12% ($0.70 → $0.61 per 1M tokens); output −12% ($2.20 → $1.93 per 1M tokens); cache read −12% ($0.13 → $0.11 per 1M tokens); OpenRouter: input −12% ($0.70 → $0.61 per 1M tokens); output −12% ($2.20 → $1.93 per 1M tokens); cache read −12% ($0.13 → $0.11 per 1M tokens); DeepInfra: input −19% ($0.93 → $0.75 per 1M tokens); output −20% ($3.00 → $2.40 per 1M tokens); cache read −22% ($0.18 → $0.14 per 1M tokens); Phala: cache read −63% ($0.70 → $0.26 per 1M tokens)
29 July 2026
ANNOUNCEMENTHow GPT-5.6 fuses frontier intelligence with frontier efficiencyAnnounced by OpenAIPRICEOpenRouter cut Qwen3 VL 30B A3B Instruct pricing by 13% $0.60 → $0.52input −13% ($0.15 → $0.13 per 1M tokens); output −13% ($0.60 → $0.52 per 1M tokens)PRICEDeepInfra cut gpt-oss-120b pricing by 75% $0.60 → $0.17input −75% ($0.15 → $0.037 per 1M tokens); output −72% ($0.60 → $0.17 per 1M tokens)PRICEOpenAI cut GPT-5.6 Luna Pro pricing by 50% $6.00 → $3.00input −50% ($1.00 → $0.50 per 1M tokens); output −50% ($6.00 → $3.00 per 1M tokens)