LIVE

BODYBUILDER$-1000000.00▼ 5.4%

GLM-4-LONG$0.00▼ 14.1%

GEMINI-2.0-FLA$0.30▼ 20.0%

GEMINI-2.0-FLA$0.40▼ 6.6%

-CF$0.00▼ 9.7%

DEEPSEEK-V4-FL$0.28▼ 26.3%

LLAMA-4-MAVERI$0.60▼ 11.9%

GROK-4-FAST$0.50▼ 6.0%

QWEN-PLUS$0.78▼ 0.6%

MINIMAX-M2.1$0.95▲ 2.6%

MINIMAX-M2$1.00▲ 2.9%

MINIMAX-01$1.10▲ 8.0%

GPT-4.1-MINI$1.60▼ 13.7%

GROK-4-1-FAST-$0.50▲ 8.9%

GPT-4.1-MINI-2$1.60▼ 31.8%

XAI$0.00▲ 4.5%

MINIMAX-M2.7$0.00▼ 9.9%

QWEN3.5-FLASH-$0.26▼ 31.0%

GPT-4.1-NANO$0.40▼ 13.0%

GEMINI-2.5-FLA$0.40▼ 0.8%

DEEPSEEK-AI$0.00▼ 27.0%

GROK-4.1-FAST$0.50▼ 16.4%

MINIMAX-M3$1.20▼ 6.1%

GOOGLE$1.50▼ 3.4%

AUTO$0.00▲ 9.7%

GOOGLE$0.00▼ 27.8%

MINIMAX-M2.7$1.20▼ 7.2%

GOOGLE$0.00▼ 0.6%

BODYBUILDER$-1000000.00▼ 5.4%

GLM-4-LONG$0.00▼ 14.1%

GEMINI-2.0-FLA$0.30▼ 20.0%

GEMINI-2.0-FLA$0.40▼ 6.6%

-CF$0.00▼ 9.7%

DEEPSEEK-V4-FL$0.28▼ 26.3%

LLAMA-4-MAVERI$0.60▼ 11.9%

GROK-4-FAST$0.50▼ 6.0%

QWEN-PLUS$0.78▼ 0.6%

MINIMAX-M2.1$0.95▲ 2.6%

MINIMAX-M2$1.00▲ 2.9%

MINIMAX-01$1.10▲ 8.0%

GPT-4.1-MINI$1.60▼ 13.7%

GROK-4-1-FAST-$0.50▲ 8.9%

GPT-4.1-MINI-2$1.60▼ 31.8%

XAI$0.00▲ 4.5%

MINIMAX-M2.7$0.00▼ 9.9%

QWEN3.5-FLASH-$0.26▼ 31.0%

GPT-4.1-NANO$0.40▼ 13.0%

GEMINI-2.5-FLA$0.40▼ 0.8%

DEEPSEEK-AI$0.00▼ 27.0%

GROK-4.1-FAST$0.50▼ 16.4%

MINIMAX-M3$1.20▼ 6.1%

GOOGLE$1.50▼ 3.4%

AUTO$0.00▲ 9.7%

GOOGLE$0.00▼ 27.8%

MINIMAX-M2.7$1.20▼ 7.2%

GOOGLE$0.00▼ 0.6%

kimi-k2.6 vs Llama-4-Maverick-17B-128E-Instruct

Llama-4-Maverick-17B-128E-Instruct

Llama-4-Maverick-17B-128E-Instruct

Higher efficiency

Llama-4-Maverick-17B-128E-Instruct

Metric

kimi-k2.6 Llama-4-Maverick-17B-128E-Instruct

Input price

$0.95

$0.00

Output price

$4.00

$0.00

Context window

262K

128K

Throughput

154 tok/s

156 tok/s

Availability

99.9%

98.6%

Cost / task

$0.004

$0.000

Efficiency score

89

89

Estimated monthly cost by workload

Metric

KIMI-K2.6

LLAMA-4-MAVERI

Chat assistant

$765.00

$0.00

RAG / long context

$1,500

$0.00

Agent / tool use

$2,124

$0.00

Efficiency score: Llama-4-Maverick-17B-128E-Instruct

Across price, speed and reliability, Llama-4-Maverick-17B-128E-Instruct offers the stronger overall balance for most workloads — but the right pick depends on your exact mix of input, output and latency needs.

Figures are illustrative demo data, not financial advice.

Frequently asked questions

Is kimi-k2.6 or Llama-4-Maverick-17B-128E-Instruct cheaper?+

Llama-4-Maverick-17B-128E-Instruct has the lower input price — $0.00 vs $0.95 per 1M tokens — so for most blended workloads it is the more cost-effective of the two. Figures are illustrative demo data.

Which should I choose, kimi-k2.6 or Llama-4-Maverick-17B-128E-Instruct?+

Across price, speed and reliability, Llama-4-Maverick-17B-128E-Instruct offers the stronger overall balance for most workloads — but the right pick depends on your exact mix of input, output and latency needs.