
Microsoft's Copilot Cowork moves to usage-based billing and may tap DeepSeek
Quick Answer
Microsoft's Copilot Cowork is transitioning to usage-based billing, moving away from flat-rate pricing deemed unsustainable.
Quick Take
The company is also considering a refined version of DeepSeek V4 as a cost-effective model option, aligning with industry trends.
Key Points
- Copilot Cowork shifts to usage-based billing model.
- Flat-rate pricing deemed unsustainable by Copilot head Charles Lamanna.
- Microsoft may utilize a refined version of DeepSeek V4.
- This change reflects broader industry trends in pricing models.
- Cost-effective options are being prioritized for users.
📖 Reader Mode
~1 min readMicrosoft is weighing a self-hosted, fine-tuned version of Deepseek V4 as a cheaper model option for Copilot Cowork, Axios reports. The company is also shifting Cowork to usage-based pricing. Cowork adapts Anthropic's Claude technology, which leans heavily on agentic reasoning and burns through tokens fast.
Copilot EVP Charles Lamanna told Axios that flat-rate pricing isn't sustainable because of "users who do hundreds of tasks a week," driving costs up quickly. Microsoft already made a similar move with GitHub Copilot, switching it to usage-based billing.
A Chinese AI model could draw criticism, especially in the US. Microsoft stresses that Deepseek would be optional and fully hosted on Azure, keeping customer data in Microsoft's cloud. The model has been customized with safeguards against bias. A final decision is expected in the coming weeks.
The move fits a blog post CEO Satya Nadella published this week, arguing for an ecosystem of AI models companies can pick and tune for specific use cases and costs. Nadella previously called AI a consumption business, saying he wants "intense users and intense usage."
— Originally published at the-decoder.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

