Tencent AI LabOpen Source

Tencent Hy3

Compare this model

A high-performance foundation model developed by Tencent. It possesses general-purpose capabilities to handle diverse tasks.

Parameters

Undisclosed

Context Window

License

Apache 2.0

Release Date

2026-07-06

Japanese Language Capability

🌐Multilingual

General multilingual model. Basic Japanese processing is possible, but inferior to specialized models.

API Pricing

Input Price (per 1M tokens)

$1.2

Output Price (per 1M tokens)

$

Billing Mode: standard

Strengths

    Weaknesses

      Use Cases

        Deep Analysis

        SWE-Bench Verified

        78.0%

        vs GLM-5.2: 84.2%

        GPQA Diamond

        90.4%

        STEM reasoning benchmark

        Input Price

        $0.14/1M

        Tencent Cloud pricing

        Hallucination Rate

        5.4%

        Down from 12.5% in preview

        Context Window

        256K

        tokens

        Active Parameters

        21B

        of 295B total

        Strengths

        • Exceptional agentic capabilities (BrowseComp: 84.2, MCP-Atlas: 79.1) leading open-weight models
        • Dramatically reduced hallucination rate (5.4%) and commonsense errors (12.7%) for production reliability
        • Highly cost-effective ($0.14/$0.56 per million tokens) with permissive Apache 2.0 license

        Weaknesses

        • Significantly trails GLM-5.2 on all coding benchmarks (SWE-Bench: 78% vs 84.2%, DeepSWE: 28% vs 46.2%)
        • Requires 8 GPUs with substantial memory for full-precision serving (H20-3e or equivalent recommended)
        • Community skepticism about benchmark validity and some reports of suboptimal KV cache management

        Competitor Comparison

        ModelArenaSWEGPQAPrice
        GLM-5.2N/A84.2%88.0%$1.40/$4.40
        DeepSeek V4 FlashN/A76.0%85.0%$0.14/$0.28
        GPT-5.5N/AN/AN/A$5.00/$30.00

        Tencent Hy3 is a 295B-parameter Mixture-of-Experts (MoE) model with only 21B active parameters per token, released under the commercially permissive Apache 2.0 license on July 6, 2026. Positioned as a cost-effective, production-ready alternative to larger frontier models, Hy3 excels in agentic tasks, tool orchestration, and long-context reasoning while featuring dramatically reduced hallucination rates (5.4%) for enhanced reliability. The model represents a significant improvement over its April preview, incorporating feedback from over 50 internal Tencent product teams and showing substantial gains in reasoning, agent capabilities, and real-world deployment stability.

        Hy3's strategic positioning focuses on delivering frontier-adjacent performance at a fraction of the cost of larger models like GLM-5.2 (744B parameters), making it particularly attractive for organizations that prioritize cost efficiency, licensing flexibility, and production reliability over absolute peak performance in coding tasks. The model's design is optimized for deployment on export-compliant hardware (like Nvidia's H20-3e) while also running efficiently on standard Western data center GPUs, and its 256K context window supports complex, multi-step workflows that are essential for modern agent applications.

        Analysis generated: 2026-07-17