
Qwen3.7-Plus is Alibaba's bid to turn multimodal AI into a full-blown autonomous agent
Quick Answer
Alibaba's Qwen3.7-Plus is a multimodal AI agent that autonomously created a vocabulary learning app, generating over 10,000 lines of code in 11 hours.
Quick Take
While it excels in visual understanding, its overall performance remains mixed. This proprietary model is priced lower than Western counterparts and lacks open weights.
Key Points
- Qwen3.7-Plus autonomously developed an app with 10,000+ lines of code.
- The model combines visual perception, GUI operation, and coding in one loop.
- Performance is strong in visual understanding but mixed overall.
- It is priced significantly lower than Western frontier models.
- No open weights are available for Qwen3.7-Plus.
Source Excerpt
Alibaba's Qwen team has released Qwen3. 7-Plus, a multimodal agent model that combines visual perception, GUI operation, and coding in a single agent loop. In a demo, an agent built on the model autonomously developed a vocabulary learning app, producing over 10,000 lines of code across 1,000 agent calls over eleven hours. The model leads on-screen understanding in Qwen's own benchmarks, but overall performance is mixed. Qwen3. 7-Plus is a proprietary offering with no open weights, priced well bel
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

