MegaBrain Leaderboard

Best AI Models
in MegaBrain

Compare live model rankings by real coding performance. See which models developers choose for planning, debugging, review, and agentic work across 453+ hosted options.

453+models24free0%inference markup

MegaBrain Bench

Cost vs completion rate across the most capable coding models · TerminalBench 2.0

MegaBrain Bench — Top 10 Most Capable Models

#ModelProviderTerminal Bench
1OPGPT-5.5openai
74.2%
2ANClaude Opus 4.7anthropic
70.1%
3ANClaude Opus 4.8anthropic
67.6%
4GOGemini 3.5 Flashgoogle
64.7%
5ANClaude Sonnet 4.6anthropic
55.1%
6MOKimi K2.6moonshot
54.4%
7X-Grok Build 0.1x-ai
50.6%
8ZHGLM 5.1zhipu
49.4%
9MIMiMo-V2.5-Promimo
47.6%
10MIMiniMax M3minimax
47.6%

Top Models by Mode

Code

1OPGPT-5.5
2ANClaude Opus 4.8
3ANClaude Opus 4.7
4GOGemini 3.5 Flash
5ANClaude Sonnet 4.6
6MOKimi K2.6
7X-Grok Build 0.1
8DEDeepSeek R1
9QWQwen3 235B
10MELlama 4 Maverick

Plan

1ANClaude Opus 4.7
2OPGPT-5.5
3ANClaude Opus 4.8
4DEDeepSeek R1
5ANClaude Sonnet 4.6
6GOGemini 3.5 Flash
7QWQwen3 235B
8MOKimi K2.6
9MELlama 4 Maverick
10ZHGLM 5.1

Debug

1ANClaude Sonnet 4.6
2OPGPT-5.5
3ANClaude Opus 4.8
4GOGemini 3.5 Flash
5ANClaude Opus 4.7
6X-Grok Build 0.1
7DEDeepSeek R1
8MIMiMo-V2.5-Pro
9ZHGLM 5.1
10GOGemini 2.5 Flash

Ask

1OPGPT-5.5
2ANClaude Sonnet 4.6
3GOGemini 3.5 Flash
4ANClaude Opus 4.7
5MOKimi K2.6
6GOGemini 2.5 Flash
7ANClaude Haiku 4.5
8MELlama 4 Maverick
9MIMiniMax M3
10QWQwen3 235B

Review

1ANClaude Opus 4.7
2OPGPT-5.5
3ANClaude Opus 4.8
4DEDeepSeek R1
5ANClaude Sonnet 4.6
6GOGemini 3.5 Flash
7X-Grok Build 0.1
8QWQwen3 235B
9MOKimi K2.6
10ZHGLM 5.1

Orchestrator

1ANClaude Opus 4.8
2OPGPT-5.5
3ANClaude Opus 4.7
4ANClaude Sonnet 4.6
5GOGemini 3.5 Flash
6DEDeepSeek R1
7MOKimi K2.6
8X-Grok Build 0.1
9MELlama 4 Maverick
10MIMiMo-V2.5-Pro

Methodology

Models earn their rank from the developers using them, not from a spec sheet. Synthetic benchmarks measure one capability at one moment. This leaderboard measures what developers come back to: real coding work, long planning sessions, debugging, review, and agentic tasks across 500+ models.

Benchmark

TerminalBench 2.0

Modes

Code, Plan, Ask, Debug, Review

Catalog

500+ models

Read the leaderboard like a pro

01

Start with usage

Rankings reflect real token usage by MegaBrain developers, not synthetic benchmarks.

02

Filter by mode

Use Top Models by Mode to see which models lead in Code, Plan, Debug, Ask, and Orchestrator.

03

Open the model page

Each model links to a dedicated page with benchmark scores, pricing, context length, and speed data.

04

Switch in MegaBrain Gateway

All 500+ models are available in MegaBrain. Switch from any model at any time.

AI model FAQ

How is MegaBrain Bench different from other benchmarks?

Most leaderboards score models on synthetic prompts. MegaBrain Bench uses TerminalBench 2.0 — a coding evaluation built around real software engineering tasks: completing functions, fixing bugs, adding tests, refactoring. The score reflects what the model can actually do in an agentic coding loop.

What does the Cost vs Performance chart show?

Each dot is a model. X-axis is TerminalBench 2.0 completion rate (higher = more capable). Y-axis is cost per attempt in dollars (log scale). Models in the bottom-right corner are the best value: high capability, low cost.

Are free models actually free?

Yes. Models marked free have $0 input and $0 output pricing. They are hosted by providers at no charge — typically to drive adoption of a new model family. Availability can change; MegaBrain Auto Free routes across free models and adapts when availability shifts.

How often is pricing updated?

Pricing data is pulled from OpenRouter every 5 minutes. Benchmark scores update when new TerminalBench results are published.

Can I use any of these models through MegaBrain today?

Yes. All 500+ models listed here are available through the MegaBrain Gateway at exact provider rates. Sign up, get an API key, and change one line of code.

What's the difference between Auto Frontier, Auto Balanced, and Auto Free?

Frontier routes every request to the most capable available model — for complex reasoning and architecture tasks. Balanced picks the best cost-effective model for your task type — the best default for daily development. Free rotates across the best free models — ideal for experimentation.

All 500+ models available in MegaBrain

Switch from any model at any time. One API key, one endpoint, no code changes.

Code for free

All Models

Browse and compare all 453 available models

Auto Frontier

mb-auto/frontier

200k ctxin $0.33/Mout $1.95/Mno training

Auto Free

mb-auto/free

256k ctxin Freeout Free

StepFun: Step 3.7 Flash (free)

stepfun/step-3.7-flash:free

262k ctxin Freeout Free

Anthropic: Claude Opus 4.8

anthropic/claude-opus-4.8

1M ctxin $5/Mout $25/Mno training

Stealth: Claude Opus 4.8 (20% off)

stealth/claude-opus-4.8

1M ctxin $4/Mout $20/M

Stealth: Claude Opus 4.7 (20% off)

stealth/claude-opus-4.7

1M ctxin $4/Mout $20/M

Stealth: Claude Sonnet 4.6 (20% off)

stealth/claude-sonnet-4.6

1M ctxin $2.40/Mout $12/M

Stealth: Claude Opus 4.6 (20% off)

stealth/claude-opus-4.6

1M ctxin $4/Mout $20/M

MoonshotAI: Kimi K3

moonshotai/kimi-k3

1M ctxin $1.95/Mout $10.92/Mno training

Anthropic: Claude Sonnet 4.6

anthropic/claude-sonnet-4.6

1M ctxin $3/Mout $15/Mno training

OpenAI: GPT-5.5

openai/gpt-5.5

1M ctxin $2.50/Mout $15/Mno training

Google: Gemini 3.1 Pro Preview

google/gemini-3.1-pro-preview

1M ctxin $1/Mout $6/Mno training

MiniMax: MiniMax M3

minimax/minimax-m3

1M ctxin $0.30/Mout $1.20/Mno training

Qwen: Qwen3.7 Plus (20% off)

qwen/qwen3.7-plus

1M ctxin $0.32/Mout $1.28/Mno training

Stealth: Qwen3.6 Plus (50% off)

stealth/qwen3.6-plus

1M ctxin $0.25/Mout $1.50/M

Z.ai: GLM 5.2

z-ai/glm-5.2

1M ctxin $1.40/Mout $4.40/Mno training

PrismML: Ternary Bonsai 2 27B

prism-ml/ternary-bonsai-2-27b

262k ctxin $0.07/Mout $0.50/Mno training

Pareto

unbiased/pareto

262k ctxin $2.50/Mout $7.50/Mno training

DeepSeek: DeepSeek Pro Latest

~deepseek/deepseek-pro-latest

1M ctxin $0.58/Mout $1.73/Mno training

DeepSeek: DeepSeek Flash Latest

~deepseek/deepseek-flash-latest

1M ctxin $0.14/Mout $0.54/Mno training

Inference.net: Schematron V2 Turbo

inference-net/schematron-v2-turbo

128k ctxin $0.03/Mout $0.15/Mno training

Inference.net: Schematron V2 Small

inference-net/schematron-v2-small

128k ctxin $0.05/Mout $0.23/Mno training

OpenAI: GPT Astra Latest ($$$$)

~openai/gpt-astra-latest

1M ctxin $10/Mout $50/Mno training

OpenAI: GPT Sol Latest

~openai/gpt-sol-latest

1M ctxin $2/Mout $10/Mno training

OpenAI: GPT Terra Latest

~openai/gpt-terra-latest

1M ctxin $2/Mout $12/Mno training

OpenAI: GPT Luna Latest

~openai/gpt-luna-latest

1M ctxin $0.20/Mout $1.20/Mno training

Sakana: Fugu Ultra v2

sakana/fugu-ultra-v2

1M ctxin $5/Mout $30/Mno training

Sakana: Fugu Max

sakana/fugu-max

1M ctxin $2/Mout $6/Mno training

inclusionAI: Ling 3.0 Flash VL

inclusionai/ling-3.0-flash-vl

131k ctxin $0.06/Mout $0.18/Mno training

inclusionAI: Ling 3.0 Flash VL (free)

inclusionai/ling-3.0-flash-vl:free

262k ctxin Freeout Free

DeepSeek: DeepSeek V4.1 Flash

deepseek/deepseek-v4.1-flash

1M ctxin $0.30/Mout $1.20/Mno training

Inception: Mercury 2.5

inception/mercury-2.5

260k ctxin $0.20/Mout $0.75/Mno training

Nex AGI: Nex-N2.5-Mini (free)

nex-agi/nex-n2.5-mini:free

262k ctxin Freeout Free

Nex AGI: Nex-N2.5-Pro (free)

nex-agi/nex-n2.5-pro:free

262k ctxin Freeout Free

OpenAI: GPT-6 Astra

openai/gpt-6-astra

1M ctxin $5/Mout $25/Mno training

OpenAI: GPT-6 Astra (batch)

openai/gpt-6-astra:batch

1M ctxin $5/Mout $25/Mno training

OpenAI: GPT-6 Astra Pro

openai/gpt-6-astra-pro

1M ctxin $5/Mout $25/Mno training

OpenAI: GPT-6 Astra Pro (batch)

openai/gpt-6-astra-pro:batch

1M ctxin $5/Mout $25/Mno training

inclusionAI: Ling 3.0 Flash Sante (free)

inclusionai/ling-3.0-flash-sante:free

262k ctxin Freeout Free

Qwen: Qwen3.8 Max (0902)

qwen/qwen3.8-max-0902

1M ctxin $2/Mout $6/Mno training

Meta: Muse Spark 1.3 Contributor

meta/muse-spark-1.3-contributor

1M ctxin $0.10/Mout $0.20/Mno training

Meta: Muse Spark 1.3

meta/muse-spark-1.3

1M ctxin $1.25/Mout $4.25/Mno training

Google: Gemini 3.8 Flash

google/gemini-3.8-flash

1M ctxin $0.75/Mout $3.75/Mno training

Google: Gemini 3.8 Flash (batch)

google/gemini-3.8-flash:batch

1M ctxin $0.75/Mout $3.75/Mno training

Anthropic: Claude Fable 5.1 ($$$$)

anthropic/claude-fable-5.1

1M ctxin $10/Mout $50/Mno training

Anthropic: Claude Fable 5.1 (batch)

anthropic/claude-fable-5.1:batch

1M ctxin $5/Mout $25/Mno training

IBM: Granite 4.2 8B

ibm-granite/granite-4.2-8b

131k ctxin $0.06/Mout $0.25/Mno training

Tencent: Hy4 preview

tencent/hy4-preview

1M ctxin $0.83/Mout $2.50/Mno training

inclusionAI: Ling 3.0 Flash Fin

inclusionai/ling-3.0-flash-fin

262k ctxin $0.06/Mout $0.18/Mno training

inclusionAI: Ling 3.0 Flash Fin (free)

inclusionai/ling-3.0-flash-fin:free

262k ctxin Freeout Free

Z.ai: GLM Flash Latest

~z-ai/glm-flash-latest

1M ctxin $0.07/Mout $0.25/Mno training

Qwen: Qwen3.8 Flash

qwen/qwen3.8-flash

1M ctxin $0.15/Mout $0.47/Mno training

Z.ai: GLM 5.3 Flash

z-ai/glm-5.3-flash

1M ctxin $0.15/Mout $0.50/Mno training

Z.ai: GLM 5.3 Flash (batch)

z-ai/glm-5.3-flash:batch

1M ctxin $0.07/Mout $0.25/Mno training

Meta: Muse Spark 1.2 Contributor

meta/muse-spark-1.2-contributor

1M ctxin $0.10/Mout $0.20/Mno training

DeepSeek: DeepSeek V4 Flash Vision Exp

deepseek/deepseek-v4-flash-vision-exp

1M ctxin $0.22/Mout $0.65/Mno training

DeepSeek: DeepSeek V4 Flash Vision Exp (batch)

deepseek/deepseek-v4-flash-vision-exp:batch

1M ctxin $0.11/Mout $0.33/Mno training

Tencent: Hy-MT2-1.8B

tencent/hy-mt2-1.8b

8k ctxin $0.04/Mout $0.18/Mno training

Tencent: Hy-MT2-30B-A3B

tencent/hy-mt2-30b-a3b

8k ctxin $0.07/Mout $0.29/Mno training

Z.ai: GLM Latest

~z-ai/glm-latest

1M ctxin $0.89/Mout $2.80/Mno training