MistralAIProprietary

Mistral Medium 3.5

Compare this model

A high-performance model developed by Mistral AI. It excels in balancing efficiency and performance.

Parameters

Undisclosed

Context Window

License

Proprietary

Release Date

2026-05-01

Japanese Language Capability

High-Quality JP

Multilingual model with strong Japanese language processing capabilities.

API Pricing

API pricing for this model is not yet available

Strengths

    Weaknesses

      Use Cases

        Deep Analysis

        SWE-Bench Verified

        77.6%

        Top open-weight model, trails Claude Sonnet 4.6 (79.6%)

        τ³-Telecom

        91.4%

        Best-in-class agentic tool use

        Parameters

        128B (dense)

        All parameters active per token

        Context Window

        256K tokens

        Larger than Claude Sonnet 4.6 (200K)

        Input Price

        $1.50/1M

        Cheaper than GPT-4o ($2.50)

        Open Weights

        Yes (Modified MIT)

        Revenue cap: $20M/month for commercial use

        Strengths

        • Single unified model replaces three specialized predecessors (instruction, reasoning, coding).
        • State-of-the-art open-weight coding agent performance (77.6% SWE-Bench).
        • Configurable reasoning effort per request (fast reply vs. deep thinking).

        Weaknesses

        • High API output pricing ($7.50/1M) relative to some open-weight competitors.
        • Self-hosting requires significant hardware (minimum 4x H100 80GB GPUs).
        • Incomplete official benchmark reporting (missing MMLU, GPQA, etc.).

        Competitor Comparison

        ModelArenaSWEGPQAPrice
        Claude Sonnet 4.6N/A79.6%N/A$3.00/$15.00
        DeepSeek V4 ProN/A80.6%N/A$1.74/$3.48
        Qwen 3.6 27BN/A72.4%N/A$0.20/$0.60

        Mistral Medium 3.5 is Mistral AI's April 2026 flagship model, representing a strategic consolidation of their previous specialized models (Magistral for reasoning, Devstral 2 for coding, and Medium 3.1 for instruction-following) into a single, dense 128B parameter architecture. This 'merged model' approach aims to simplify deployment and eliminate the need for model routing in complex agentic workflows. It features a massive 256K token context window, configurable reasoning effort, and a custom-trained vision encoder for handling variable image sizes.

        Positioned as the premier open-weight coding agent model, Medium 3.5 achieves 77.6% on SWE-Bench Verified, placing it in elite company just behind closed-source leaders like Claude Sonnet 4.6. Its standout agentic capability is demonstrated by a 91.4% score on the τ³-Telecom benchmark. The release is tightly integrated with Mistral's product ecosystem, powering the new asynchronous cloud coding agents in Vibe and Le Chat's Work Mode. While its API pricing is competitive for its capability tier, the Modified MIT license and substantial self-hosting requirements create a specific value proposition for teams needing frontier performance with data sovereignty or high-volume agentic use cases.

        Analysis generated: 2026-07-17