
Microsoft launches its own cybersecurity model MAI-Cyber-1-Flash but still depends on OpenAI for the toughest tasks
Quick Answer
Microsoft has introduced MAI-Cyber-1-Flash, a new AI cybersecurity model that scores 96% on CyberGym, outperforming competitors like Mythos and Gemini.
Quick Take
This model handles 90% of tasks, reducing costs by 50%, while still relying on OpenAI's GPT-5.4 for complex reasoning.
Key Points
- MAI-Cyber-1-Flash is integrated into Microsoft's MDASH .
- The model achieves a 96% score on CyberGym, 12 points higher than Mythos.
- Cost reductions of 50% are expected due to the model's efficiency.
- Microsoft's Perception system monitors threats in real-time, leveraging vast data.
- The company is shifting towards open-weights advocacy in AI development.
Source Excerpt
Microsoft introduces MAI-Cyber-1-Flash, a compact security model that scores 96 percent on the CyberGym benchmark when embedded in its MDASH . Microsoft says costs should drop by 50 percent compared to pure frontier models, since only tough cases get passed to GPT-5. 4. For complex reasoning, Microsoft still relies on OpenAI.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

