Zhipu AIOpen Source

GLM-5.2

Compare this model

A high-performance foundation model developed by Zhipu AI. It excels in Chinese language support and handles diverse tasks.

Parameters

Undisclosed

Context Window

License

MIT

Release Date

2026-06-13

Japanese Language Capability

🌐Multilingual

General multilingual model. Basic Japanese processing is possible, but inferior to specialized models.

API Pricing

Input Price (per 1M tokens)

$1.4

Output Price (per 1M tokens)

$

Billing Mode: standard

Strengths

    Weaknesses

      Use Cases

        Deep Analysis

        BenchLM Overall Score

        81/100

        #10 of 79 models (provisional)

        SWE-bench Pro

        62.1%

        vs Claude Opus 4.8: 69.2%

        FrontierSWE

        74.4%

        vs GPT-5.5: 72.6%, vs Opus 4.8: 75.1%

        Terminal-Bench 2.1

        81.0%

        vs Claude Opus 4.8: 85.0%

        Input/Output Price

        $1.40/$4.40 per 1M tokens

        6.8x cheaper output than GPT-5.5

        Context Window

        1M tokens

        MIT open weights license

        Strengths

        • Best open-weights model for long-horizon agentic coding, rivaling closed frontier models at a fraction of the cost
        • Solid 1M-token context with engineering-grade reliability for sustained coding-agent trajectories
        • MIT license with full self-hosting support across vLLM, SGLang, transformers, and other frameworks

        Weaknesses

        • 753B parameters requires substantial GPU infrastructure (>1TB VRAM for unquantized serving)
        • Trail behind Claude Opus 4.8 on hardest frontier tasks (SWE-bench Verified gap, SWE-Marathon gap)
        • Slower time-to-first-token than closed competitors (14.4s median vs ~9s in independent tests)

        Competitor Comparison

        ModelArenaSWEGPQAPrice
        Claude Opus 4.8N/A69.2%93.6%$5.00/$25.00
        GPT-5.5N/A58.6%93.6%$5.00/$30.00
        Gemini 3.1 ProN/A54.2%94.3%$2.00/$12.00

        GLM-5.2, released by Z.ai (formerly Zhipu AI) on June 16, 2026, is a 753B-parameter mixture-of-experts (40B active) foundation model purpose-built for long-horizon agentic coding and engineering tasks. It represents a generational leap over its predecessor GLM-5.1, expanding the context window from 200K to a solid 1M tokens while dramatically improving coding capabilities across every major benchmark. On Terminal-Bench 2.1 it scored 81.0 versus 63.5 for GLM-5.1, and on FrontierSWE it jumped from 30.5 to 74.4—within 1% of Claude Opus 4.8.

        The model's positioning is deliberate: it targets the gap between expensive closed frontier models and less capable open alternatives. At $1.40/$4.40 per million tokens, it delivers near-frontier performance at roughly one-sixth the output cost of GPT-5.5 and Claude Opus 4.8. The MIT license enables full self-hosting and commercial use without regional restrictions. Key architectural innovations include IndexShare (reducing per-token FLOPs by 2.9× at 1M context) and improved multi-token prediction layers achieving 20% longer acceptance lengths for speculative decoding.

        Community reception has been strong. Lambda Labs called it a "DeepSeek moment for agents," noting that experienced labs began replacing closed-source workloads with GLM-5.2 within weeks of release. DeepLearning.ai highlighted it as the top open-weights model for post-training benchmarks and web-development coding. The model's agentic RL training with anti-hack mechanisms—a system to detect and block reward-hacking behaviors during training—sets it apart in its ability to reliably complete extended autonomous tasks without shortcutting.

        Analysis generated: 2026-07-17