
Moonshot pauses new Kimi K3 subscriptions after GPU demand maxes out in 48 hours
Quick Answer
Moonshot has paused new subscriptions for its Kimi K3 model after demand maxed out GPU capacity within 48 hours.
Quick Take
Current subscribers remain unaffected as the company introduces a two-tier subscription model: 'Kimi Membership' for general features and 'Kimi Code Membership' for programming workflows. Meanwhile, Alibaba is promoting its Qwen 3.8 model with a discounted preview version available.
Key Points
- Moonshot's Kimi K3 model subscription halted due to maxed GPU demand.
- Current subscribers will not be affected by the pause.
- New subscription tiers include 'Kimi Membership' and 'Kimi Code Membership'.
- Alibaba's Qwen 3.8 model is now available with a discounted preview.
- The changes aim to stabilize user experience and distribute computing power.
📖 Reader Mode
~1 min readSo much for the idea that open source cuts computing needs. Moonshot, the AI startup behind the Kimi K3 model, has temporarily stopped selling new subscriptions. Over the past 48 hours, "demand has pushed close to the limits of our current capacity," the company announced on X about its recently released Kimi K3 model.
Current subscribers aren't affected, and new slots will open up again gradually. Moonshot also announced that it's splitting its subscription model into two tiers. Going forward, a "Kimi Membership" will cover web, app, and work features, while a separate "Kimi Code Membership" will handle programming workflows. The company says it wants to spread computing power more evenly and keep the user experience stable.
Meanwhile, competitor Alibaba is already pushing its rival model, Qwen 3.8, which will be open weight for the first time in a while. A paid but heavily discounted preview version is already available.
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

