Z.ai Launches GLM-5.2 With a Usable 1M-Token Context, Two Thinking-Effort Levels, and No Benchmarks at Launch
Quick Answer
Z.ai has launched GLM-5.2, featuring a 1-million-token context window and two levels of thinking effort (High and Max).
Quick Take
The model integrates with Claude Code, Cline, and OpenClaw via an Anthropic-compatible endpoint, but no benchmarks were provided at launch, with MIT open weights expected next week.
Key Points
- GLM-5.2 offers a 1M-token context window for enhanced usability.
- Two thinking effort levels: High and Max, for varied processing needs.
- Integrates with Claude Code, Cline, and OpenClaw via Anthropic endpoint.
- No benchmarks available at launch; MIT open weights promised soon.
- Launch date was June 13, 2026, across all GLM Coding Plan tiers.
Source Excerpt
Z. ai released GLM-5. 2 with a usable 1-million-token context window across all GLM Coding Plan tiers; MIT weights pending.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from MarkTechPost
See more →Meet Flash-KMeans: An IO-Aware, Exact K-Means That Runs Over 200× Faster Than FAISS on GPUs
Flash-KMeans is an open-source, IO-aware k-means implementation that operates over 200× faster than FAISS on NVIDIA H200 GPUs. It achieves 17.9× end-to-end and 33× speedup over cuML by optimizing distance calculations and updating mechanisms without approximating results. This advancement significantly enhances performance for data scientists and machine learning practitioners.