Meta Superintelligence Labsが2026年8月10日にApache 2.0で公開した30Bパラメータのオープンウェイト・エージェンティックモデル。28Bのテキストデコーダと約2BのViT風視覚エンコーダからなり、テキストと画像を受け取りテキストを出力。Muse Spark 1.2からの知識蒸留で生まれ、4bit量子化により24GB VRAMの消費者向けGPU1枚で動作。131,072トークンのコンテキスト、100以上の言語、低・中・高・xhighの推論設定を持ち、llama.cpp・MLX・Ollama・vLLM・SGLangで動作。MCP Atlas 75.5、SWE-Bench Verified 76.0、Terminal-Bench 2.1 51.7、AIME 2026 94.7、GPQA Diamond 83.5を記録。
パラメータ
30B
コンテキスト長
131K
ライセンス
Apache 2.0
リリース日
2026-08-10
API料金
このモデルのAPI料金情報は現在未公開です
強み
- •30Bで消費者向けGPU1枚(24GB)に収まるオープンウェイト
- •Apache 2.0でLlama以来最も寛容な公開ライセンス
- •MCP Atlas 75.5・SWE-Bench Verified 76.0とエージェント性能が高い
- •テキスト+画像入力で画面・図表・文書を解釈しツールを呼び出せる
弱み
- •ベンチマーク数値は全てMeta自己申告(第三者再現待ち)
- •音声・動画は一次モダリティではなく画像のみ
- •Muse Spark 1.2の教師サイズ・蒸留レシピは非公開
- •フル精度は55GB超で量子化なしでは実用困難
活用例
- •ローカル・プライバシー重視のコーディングエージェント
- •スクリーンショット・PDFを読む社内文書・ヘルプデスク自動化
- •消費者GPUで動くオフラインAIアシスタント
- •ファインチューニング済みエージェント・スキャフォールドのベース
深度分析
Parameters
30B (28B text + 2B vision)
Dense causal transformer, 52 layers; ~1.8B ViT-G/14 perception encoder
Context Window
131,072 tokens
Knowledge cutoff January 4, 2026; 100+ languages
License
Apache 2.0
Meta's most permissive open release since Llama
VRAM to Run
24 GB (4-bit quantized)
Full precision >55GB; fits RTX 3090/4090 or Apple Silicon
MCP Atlas
75.5
Beats Gemma4-31B (54.2) and Qwen3.6-27B (62.5)
SWE-Bench Verified
76.0
Terminal-Bench 2.1 51.7; AIME 2026 94.7; GPQA Diamond 83.5
強み
- ・Open weights that run on a single 24GB consumer GPU after 4-bit quantization — local, private agentic inference without a hosted API.
- ・Apache 2.0, Meta's most permissive open license, enabling unrestricted commercial use, modification, and redistribution.
- ・Strong agentic scores (MCP Atlas 75.5, SWE-Bench Verified 76.0) and multimodal text+image input for screenshot, chart, and document understanding.
- ・Day-0 ecosystem support: llama.cpp, MLX, ExecuTorch, Ollama, LM Studio, vLLM, SGLang, plus a DFlash speculative-decoding drafter (3.1x on RTX 5090).
弱み
- ・Every benchmark figure is Meta self-reported and still pending independent replication.
- ・Image input only — no audio or video as primary modalities.
- ・The Muse Spark 1.2 teacher's size and the distillation recipe are undisclosed.
- ・Full precision exceeds 55GB, so it is impractical without quantization on consumer hardware.
競合比較
| Model | Arena | SWE | GPQA | Price |
|---|---|---|---|---|
| Muse Glimmer 30B | 75.5 (MCP Atlas) | 76.0 (SWE-Bench Verified) | 83.5 (GPQA Diamond) | Apache 2.0; 24GB |
| Gemma4-31B | 54.2 (MCP Atlas) | 66.6 (SWE-Bench Verified) | 85.7 (GPQA Diamond) | Gemma license |
| Qwen3.6-27B | 62.5 (MCP Atlas) | 77.2 (SWE-Bench Verified) | 84.2 (GPQA Diamond) | Qwen license |
| Muse Spark 1.2 | 54 (AA Intelligence) | N/A | N/A | Closed; weights promised |
Muse Glimmer 30B is Meta Superintelligence Labs' open-weight agentic model, released August 10, 2026 under Apache 2.0. Distilled from Muse Spark 1.2, it runs locally on a single 24GB GPU, accepts interleaved text and images, and is tuned around the agent loop — plan, call tools, check results, recover from failure.
出典
分析生成日: 2026-09-02