Major, public vendor updates
OpenAI — GPT‑5.6 Sol Pro: our price feed recorded large reductions to Sol Pro per‑token rates (example: input, output and cached‑input reductions visible in Oct 2026 snapshots). OpenAI has published pricing and product posts about GPT‑5.6 and earlier Sol promotions; however we did not find a single OpenAI blog post that explicitly states Oct 5 as the cause for the specific 50% Pro price moves recorded in the feed. See OpenAI’s GPT‑5.6 documentation for the vendor’s published pricing guidance. (openai.com)
DeepSeek — V4 family: DeepSeek published a new V4 rate card and release notes (DeepSeek‑V4.1 / V4‑Pro changes, peak/off‑peak tiers and a shift in Flash/Pro routing). DeepSeek’s public docs and blog posts explain the company’s move to a peak/off‑peak two‑tier rate card and the V4.1 Flash / V4‑Pro positioning, which account for many of the V4‑series price moves tracked around early October. (deepseek.ai)
Tracker‑reported moves with no vendor rationale published
MoonshotAI — Kimi K3: trackers show per‑route changes (including cached‑input and output reprice on Oct 5) but Moonshot’s public launch/pricing posts we found predate October and do not include an Oct 5 explanation. We could not find a MoonshotAI statement that explains the Oct 5 delta. (forum.moonshot.ai)
Meta — Llama 3.3 70B Instruct: price increases were recorded in multiple price trackers on Oct 5–6; Meta does not publish a public line‑item justification for that repricing in its public model pages, and we found only tracker snapshots documenting the change. (weaveprism.com)
NVIDIA — Nemotron 3.5 Lightning: small per‑token adjustments appear in tracker snapshots (input +1%, output −6%, tiny cached‑input change). NVIDIA’s product press material announces the Nemotron family but does not include a public Oct 5 pricing rationale for the small moves recorded by trackers. (nvidianews.nvidia.com)
Qwen — Qwen3.8 / Qwen3.6 (27B): OpenRouter and price trackers show small in/out adjustments in early October; we did not find a direct Qwen (Alibaba Qwen team) public changelog that explains the Oct 5 shifts. (writeworks.ai)
Google — Gemma 4 26B A4B: trackers recorded ~+33% host pricing moves in early Oct; Google publishes Gemma model pages and cloud/Vertex listings, but we did not find a Google statement specifically attributing the Oct 5 price deltas captured by trackers. (theknowngood.com)
Smaller adjustments and context
Tencent — Hy3 and Hy4 preview: Tencent’s published Hy3/Hy4 product pages and TokenHub region pages list vendor pricing for Hy3 and Hy4 preview. Trackers show the Hy3/Hy4 per‑token drops recorded on Oct 5, but we did not find a Tencent press release that calls out Oct 5 as the cause. Tencent’s model release pages and TokenHub docs remain the authoritative vendor references. (tencent.com)
Z.ai — GLM‑5.3: several trackers and aggregator pages show GLM‑5.3 route price increases around Oct 5. Z.ai’s published pricing (docs/pricing) documents the GLM family’s list card — we found tracker evidence of the Oct 5 move but no Z.ai blog post explicitly tying a product decision to the Oct 5 price deltas. (glm5.app)
Summary of the pattern we see: public vendor posts (DeepSeek, Tencent, OpenAI product pages) explain some of the larger structural changes (new rate cards, promotions, peak/off‑peak rules). For many per‑route changes captured on Oct 5 the only public evidence is real‑time price registries (OpenRouter/aggregators/price‑trackers) rather than a vendor press release; where a vendor has published an explanation we cite it, and where they have not we note that explicitly and point to the tracker snapshot.
Why providers changed prices
MoonshotAI (Kimi K3): price snapshots on Oct 5 show an increase to the model’s output token rate and a decrease in the cached‑input rate on some routes; Moonshot’s public launch/pricing pages list Kimi K3 pricing but we could not find a MoonshotAI announcement that explains the Oct 5 route‑level changes. For the vendor’s published baseline pricing and K3 launch materials see the Moonshot/Kimi K3 announcement and pricing pages. (forum.moonshot.ai)
OpenAI (GPT‑5.6 Sol Pro): trackers show large per‑token reductions recorded in early October for Sol Pro (input/output/cached input). OpenAI has published product and pricing pages for GPT‑5.6 and prior promotional pricing, but we did not find a single OpenAI blog post explicitly saying the Oct 5 snapshot was caused by a new OpenAI policy change; OpenAI’s public GPT‑5.6 pages explain prior Sol pricing and promotional adjustments. (Official OpenAI model/pricing pages and changelog referenced.) (openai.com)
Meta (Llama 3.3 70B Instruct): price trackers recorded substantial increases for many hosts on Oct 5–6. Meta publishes model cards and Llama documentation, but we did not find a Meta press release or pricing note attributing these Oct 5 tracker moves to a Meta decision. The changes are visible in multiple price‑tracking snapshots. (weaveprism.com)
NVIDIA (Nemotron 3.5 Lightning): tracker snapshots show small input/output/cached‑input adjustments on Oct 5. NVIDIA’s Nemotron family launch materials describe the model family but we found no NVIDIA statement explicitly explaining the recorded Oct 5 micro‑adjustments. (nvidianews.nvidia.com)
Tencent (Hy3 / Hy4 preview): Tencent’s product pages and TokenHub region docs publish Hy3 and Hy4 preview list prices. Trackers recorded the Oct 5 per‑token reductions for Hy3/Hy4 routes in our dataset, but we did not find a Tencent press note directly saying 'this is why' for the Oct 5 feed moves. See Tencent’s Hy3/Hy4 release and TokenHub pricing pages for the vendor’s published rates. (tencent.com)
DeepSeek (V4 series): DeepSeek published an updated V4 rate card, the DeepSeek‑V4.1 GA post and an explanation of peak/off‑peak tiers and the new Flash/Pro positioning — those vendor posts explain most of the V4 family’s repricing and the move to a two‑tier (peak/off‑peak) structure that affected the V4‑series rates tracked around Oct 5. (deepseek.ai)
Z.ai (GLM‑5.3): GLM‑5.3 list prices are documented in Z.ai’s published pricing table; multiple trackers show host‑level price increases around Oct 5. We could not find a Z.ai blog post explicitly attributing the Oct 5 host‑level price moves to a published vendor decision. (glm5.app)
Qwen (Qwen3.8 27B / Qwen3.6 27B): OpenRouter and price trackers show small early‑October adjustments (tiny input upticks, output cuts on some routes). We did not find an Alibaba/Qwen press release explaining the Oct 5 feed changes; the only public evidence for the Oct 5 moves is in route snapshots and aggregator logs. (writeworks.ai)
Google (Gemma 4 26B A4B): host pricing snapshots on early October show higher per‑token host rates for Gemma 4 26B on some routes. Google’s Gemma model pages and cloud docs list vendor rates in their published model documentation, but we did not find a Google announcement explicitly explaining the Oct 5 host price changes captured by trackers. (tokencost.app)
Sources
- OpenAI — GPT‑5.6 landing / pricing pages
- OpenAI API changelog / developer notices
- DeepSeek — pricing and V4.1 release notes
- DeepSeek V4.1 GA blog post
- Tencent — Hy4 preview release and TokenHub (product/pricing pages)
- Tencent Cloud / TokenHub Hy4 pricing guidance (region page)
- Z.ai / GLM‑5.3 pricing pages and trackers
- WeavePrism — model price change log (Llama 3.3 entries)
- MoonshotAI Kimi K3 announcement / forum
- OpenRouter / pricing trackers and daily snapshots (aggregators used for Oct 5 observed deltas)
- Tokenando / NVIDIA Nemotron price snapshots
- Qwen / Qwen3.8 price tracker summary
