Opencode Go Bench · opencode.ai/go × Artificial Analysis

Every Go model, scored.

The full opencode.ai/go lineup joined to independent Artificial Analysis benchmarks — intelligence, skills, cost, and speed in one place, updated automatically.

35/43 modelsupdated 2026-10-05T19:56:22.654Z

Smartest

muse-spark-1.3-contributor

Index 48.1

Fastest output

deepseek-v4.1-flash

214.2 tok/s

Cheapest @ 3:1

hy3-preview

$0.10/1M blended

Lowest latency

deepseek-v4.1-flash

0.80s TTFT

Intelligence

How smart is each model?

Artificial Analysis Intelligence Index · higher is better · Go models only · hover a bar for details, click to open on AA

Benchmarks

Where does each model shine?

Independent AA skill evaluations · % scores · higher is better

Chart
ModelGPQAHLESciCodeLCRTerminal-Benchτ-BankingMMLU-ProLiveCodeBench
muse-spark-1.3-contributor93.5%48.7%58.8%83.0%84.3%50.5%——
grok-4.7—43.1%57.4%76.7%————
mimo-v2.6-pro—49.4%60.9%86.3%————
qwen3.8-max92.8%43.1%52.1%80.3%88.8%47.8%——
glm-5.391.7%42.3%59.0%79.7%83.9%50.3%——
grok-4.694.9%42.9%56.5%80.3%88.4%50.7%——
kimi-k393.5%46.9%59.5%88.7%85.0%46.0%——
glm-5.3-flash91.2%39.9%51.6%80.0%84.3%47.2%——
muse-spark-1.2-contributor90.4%45.5%57.4%79.0%80.1%34.8%——
deepseek-v4.1-flash—39.2%51.9%84.0%————
grok-4.593.1%42.7%55.0%79.3%81.6%42.1%——
gpt-6-luna—38.5%54.6%83.3%————
mimo-v2.6-flash—35.1%51.3%74.3%————
gpt-5.6-luna91.1%39.5%53.6%83.7%80.9%31.1%——
deepseek-v4-pro92.8%41.0%51.0%80.3%78.7%39.6%——
deepseek-v4-flash90.8%38.6%50.3%79.7%78.7%39.4%——
glm-5.289.5%41.1%51.2%78.3%77.9%34.6%——
qwen3.7-max92.3%40.5%49.5%79.0%74.5%11.8%——
minimax-m392.9%39.0%47.1%83.0%65.2%15.3%——
mimo-v2-pro87.0%30.4%—68.3%————
glm-582.0%29.3%—75.7%————
kimi-k2.691.1%37.5%51.5%81.0%65.9%23.3%——
qwen3.6-plus88.2%27.8%—78.3%61.4%20.8%——
glm-5.186.8%30.1%44.8%73.7%61.8%13.6%——
mimo-v2.5-pro86.6%35.7%50.6%79.7%65.2%9.9%——
mimo-v2.586.6%35.7%50.6%79.7%65.2%9.9%——
kimi-k2.7-code89.6%35.0%47.8%79.3%67.4%20.2%——
hy389.7%33.5%48.6%79.0%64.4%22.9%——
qwen3.7-plus90.0%35.6%46.1%73.0%61.0%17.5%——
mimo-v2-omni82.8%22.1%—75.0%————
kimi-k2.587.9%30.7%—78.0%45.7%14.2%——
minimax-m2.787.4%29.6%50.1%78.3%55.4%9.9%——
minimax-m2.584.8%20.5%—73.3%————
hy3-preview86.7%27.8%—64.7%————
longcat-2.078.0%33.7%36.3%65.0%50.2%13.2%——

Cost & Speed

What does smart cost, and how fast?

Intelligence Index vs cost per task (blended $/1M @ 3:1 input:output, log scale) · top-left green = most attractive quadrant · dotted line = cost-performance frontier

Input : output mix

Blended $/1M @ 3:1 — Artificial Analysis default. Scatter, table, and $/intel all recompute from your raw input/output prices.

★ Most attractive — high intelligence, low cost0102030405060$0.08$0.1$0.15$0.2$0.3$0.4$0.5$0.6$0.8$1$1.5$2$3$4$5$6$8Cost per Task @ 3:1 (USD, Log Scale)Intelligence IndexMiniMax-M3 · Intel 29.2 · $0.52/1Mminimax-m3MiniMax-M2.7 · Intel 22.8 · $0.52/1Mminimax-m2.7MiniMax-M2.5 · Intel 22.8 · $0.52/1Mminimax-m2.5Kimi K3 (Max) · Intel 43.6 · $6.00/1Mkimi-k3Kimi K2.7 Code · Intel 25.8 · $1.71/1Mkimi-k2.7-codeKimi K2.6 (Reasoning) · Intel 27.0 · $1.71/1Mkimi-k2.6LongCat 2.0 · Intel 19.1 · $0.52/1Mlongcat-2.0Kimi K2.5 (Reasoning) · Intel 23.5 · $1.14/1Mkimi-k2.5GLM-5.2 (Max) · Intel 33.7 · $2.15/1Mglm-5.2GLM 5.3 Flash · Intel 41.8 · $0.24/1Mglm-5.3-flashGLM-5.3 (Max) · Intel 44.8 · $2.15/1Mglm-5.3GLM-5.1 (Reasoning) · Intel 26.1 · $1.98/1Mglm-5.1GLM-5 (Reasoning) · Intel 27.9 · $1.55/1Mglm-5DeepSeek V4 Pro 0813 (Max) · Intel 36.0 · $1.98/1Mdeepseek-v4-proDeepSeek V4 Flash 0731 (Max) · Intel 34.3 · $0.66/1Mdeepseek-v4-flashDeepSeek V4.1 Flash (Max) · Intel 39.5 · $0.52/1Mdeepseek-v4.1-flashQwen3.7 Max · Intel 29.5 · $3.75/1Mqwen3.7-maxQwen3.8 Max (0902) · Intel 45.4 · $3.00/1Mqwen3.8-maxQwen3.7 Plus · Intel 25.2 · $0.70/1Mqwen3.7-plusQwen3.6 Plus · Intel 27.0 · $1.13/1Mqwen3.6-plusMiMo-V2.6-Pro · Intel 46.3 · $0.54/1Mmimo-v2.6-proMiMo-V2.6-Flash · Intel 37.9 · $0.18/1Mmimo-v2.6-flashMiMo-V2.5-Pro (Reasoning) · Intel 26.0 · $0.54/1Mmimo-v2.5-proMiMo-V2.5-Pro (Reasoning) · Intel 26.0 · $0.54/1Mmimo-v2.5Hy3 · Intel 25.3 · $0.24/1Mhy3Hy3-preview (Reasoning) · Intel 22.7 · $0.10/1Mhy3-previewGPT-5.6 Luna (Max) · Intel 37.3 · $0.45/1Mgpt-5.6-lunaGrok 4.5 (High) · Intel 38.8 · $3.00/1Mgrok-4.5Grok 4.7 (Xhigh) · Intel 46.4 · $3.00/1Mgrok-4.7Grok 4.6 (High) · Intel 44.3 · $3.00/1Mgrok-4.6Muse Spark 1.3 (Max) · Intel 48.1 · $2.00/1Mmuse-spark-1.3-contributorMuse Spark 1.2 (Xhigh) · Intel 39.6 · $2.00/1Mmuse-spark-1.2-contributorGPT-6 Luna (Max) · Intel 38.1 · $0.20/1Mgpt-6-luna
green halo = released within the last 7 days

Output speed tokens/sec · higher is better

All models

The full lineup

The signal strip reads intelligence · coding · value · speed at a glance (value follows the 3:1 mix above the scatter). ● NEW marks models released within the last 7 days.

muse-spark-1.3-contributor
Meta · Muse Spark 1.3 (Max)
48.175.8—$2.00$0.042144.029.85s
grok-4.7
SpaceXAI · Grok 4.7 (Xhigh)
46.4——$3.00$0.06573.610.74s
mimo-v2.6-pro
Xiaomi · MiMo-V2.6-Pro
46.3——$0.54$0.01246.64.65s
qwen3.8-max
Alibaba · Qwen3.8 Max (0902)
45.476.2—$3.00$0.06636.91.84s
glm-5.3
Z AI · GLM-5.3 (Max)
44.874.8—$2.15$0.04875.62.88s
grok-4.6
SpaceXAI · Grok 4.6 (High)
44.376.8—$3.00$0.068——
kimi-k3
Kimi · Kimi K3 (Max)
43.676.2—$6.00$0.13844.52.57s
glm-5.3-flash
Z AI · GLM 5.3 Flash
41.871.5—$0.24$0.00651.22.73s
muse-spark-1.2-contributor
Meta · Muse Spark 1.2 (Xhigh)
39.672.2—$2.00$0.051——
deepseek-v4.1-flash
DeepSeek · DeepSeek V4.1 Flash (Max)
39.5——$0.52$0.013214.20.80s
grok-4.5
SpaceXAI · Grok 4.5 (High)
38.872.4—$3.00$0.077——
gpt-6-luna
OpenAI · GPT-6 Luna (Max)
38.1——$0.20$0.005135.775.43s
mimo-v2.6-flash
Xiaomi · MiMo-V2.6-Flash
37.9——$0.18$0.00553.33.38s
gpt-5.6-luna
OpenAI · GPT-5.6 Luna (Max)
37.371.4—$0.45$0.012——
deepseek-v4-pro
DeepSeek · DeepSeek V4 Pro 0813 (Max)
36.068.8—$1.98$0.055106.01.10s
deepseek-v4-flash
DeepSeek · DeepSeek V4 Flash 0731 (Max)
34.369.1—$0.66$0.019——
glm-5.2
Z AI · GLM-5.2 (Max)
33.768.8—$2.15$0.064——
qwen3.7-max
Alibaba · Qwen3.7 Max
29.566.0—$3.75$0.127——
minimax-m3
MiniMax · MiniMax-M3
29.258.6—$0.52$0.01880.21.18s
mimo-v2-pro
Xiaomi · MiMo-V2-Pro
28.6——————
glm-5
Z AI · GLM-5 (Reasoning)
27.9——$1.55$0.056——
kimi-k2.6
Kimi · Kimi K2.6 (Reasoning)
27.061.8—$1.71$0.063——
qwen3.6-plus
Alibaba · Qwen3.6 Plus
27.054.5—$1.13$0.042——
glm-5.1
Z AI · GLM-5.1 (Reasoning)
26.155.8—$1.98$0.076——
mimo-v2.5-pro
Xiaomi · MiMo-V2.5-Pro (Reasoning)
26.060.2—$0.54$0.021——
mimo-v2.5
Xiaomi · MiMo-V2.5-Pro (Reasoning)
26.060.2—$0.54$0.021——
kimi-k2.7-code
Kimi · Kimi K2.7 Code
25.860.8—$1.71$0.066100.61.20s
hy3
Tencent · Hy3
25.358.8—$0.24$0.01090.22.01s
qwen3.7-plus
Alibaba · Qwen3.7 Plus
25.255.9—$0.70$0.02855.51.58s
mimo-v2-omni
Xiaomi · MiMo-V2-Omni
23.9——————
kimi-k2.5
Kimi · Kimi K2.5 (Reasoning)
23.546.8—$1.14$0.048——
minimax-m2.7
MiniMax · MiniMax-M2.7
22.852.6—$0.52$0.023——
minimax-m2.5
MiniMax · MiniMax-M2.5
22.8——$0.52$0.023——
hy3-preview
Tencent · Hy3-preview (Reasoning)
22.7——$0.10$0.004——
longcat-2.0
LongCat · LongCat 2.0
19.145.3—$0.52$0.027——
deepseek-flash
missing on AA
———————
deepseek-v4-flash-vision-exp
missing on AA
———————
qwen3.8-flash
missing on AA
———————
qwen3.5-plus
missing on AA
———————
space-bunny-free
missing on AA
———————
longcat-2.5-preview-free
missing on AA
———————
hy4-preview
missing on AA
———————
omen-alpha
missing on AA
———————