Moonshot sets an August 31 sunset for the context-tiered moonshot-v1 API
Moonshot's pricing docs now flag the legacy moonshot-v1 family — the models that priced the same weights by context window — for full platform sunset on August 31, retiring context length as a billing axis.
moonshot-v1-8k / -32k / -128k live at $0.20/$2.00, $1.00/$3.00 and $2.00/$5.00 per 1M tokens, with matching -vision-preview variants
Same rates still published, but the model overview now reads 'full platform sunset expected on August 31'; legacy Kimi K2 0711 / 0905 / Turbo already removed from the card
Moonshot’s model-pricing overview now describes the Moonshot V1 family as the “classic generation model series; full platform sunset expected on August 31”. The six SKUs are still listed at their long-standing rates — moonshot-v1-8k at $0.20 in / $2.00 out, -32k at $1.00 / $3.00, -128k at $2.00 / $5.00 per 1M tokens, plus identically-priced -vision-preview variants — but the family is on a clock.
That retires the company’s most distinctive billing mechanic: moonshot-v1 charged different per-token rates for the same model depending on the context window you requested, making context length literally the meter. The Kimi K2 and K3 models that replace it fold long context into a single flat per-model rate (262,144 tokens on the K2 family, 1,048,576 on K3) and put the cost lever on automatic context caching and model choice instead. The legacy Kimi K2 0711 / 0905 / Turbo models have already vanished from the card following their May 2026 end-of-life.
Moonshot's model-pricing overview now describes the Moonshot V1 family as the 'classic generation model series; full platform sunset expected on August 31', ending the mechanic that charged different per-token rates for the same model by context window ($0.20/$2.00 at 8k up to $2.00/$5.00 at 128k). The legacy Kimi K2 0711/0905/Turbo and K2 Thinking models have also disappeared from the card. Migration is not price-neutral: moonshot-v1-128k users move down to Kimi K2.5 at $0.60/$3.00 with double the context, while moonshot-v1-8k users lose the cheapest input rate on the platform and land on a 3× higher one.