The best open LLMs for your use case:

1Kimi K3Moonshot

Moonshot AI's 2.8-trillion-parameter open-weight frontier model with native vision, 1M-token context, and state-of-the-art agentic coding and tool use.

Speed:

Intelligence:

Price: (1M Tokens)

$3.00 / 15.00

Cached input: (1M Tokens)

$0.30

Context: (tokens)

1,048,576

Inputs:

ImageText

Benchmarks:

#1

GPQA-Diamond

General Knowledge

93.5
#1

HLE

General Knowledge

43.5
#1

MMMU-Pro

Multimodal - Vision

81.6
#1

Terminal-Bench 2.1

Coding Agents

88.3
#1

FrontierSWE

Coding Agents

81.2
#1

DeepSWE

Coding Agents

69
#1

Program Bench

Coding Agents

77.8
#1

MCP-Atlas

Agents and Function Calling

84.2
#1

GPQA-Diamond

General Knowledge

93.5
#1

HLE

General Knowledge

43.5
#1

MMMU-Pro

Multimodal - Vision

81.6
#1

Terminal-Bench 2.1

Coding Agents

88.3
#1

FrontierSWE

Coding Agents

81.2
#1

DeepSWE

Coding Agents

69
#1

Program Bench

Coding Agents

77.8
#1

MCP-Atlas

Agents and Function Calling

84.2
2Kimi K2.6Moonshot

1T-parameter MoE flagship from Moonshot with long-horizon coding, agent swarms scaling to 300 sub-agents, and state-of-the-art reasoning.

Speed:

Intelligence:

Price: (1M Tokens)

$1.20 / 4.50

Cached input: (1M Tokens)

$0.20

Context: (tokens)

262,144

Inputs:

ImageText

Benchmarks:

#3

EQBench

Creative Writing

1561
#1

SciCode

Coding Agents

52.2
#1

MCP-Mark

Agents and Function Calling

55.9
#1

Apex Agents

Agents and Function Calling

27.9
#1

FrontierCode

Coding Agents

3.8
#2

LiveCodeBench

Coding Agents

89.6
#2

Terminal-Bench 2.0

Coding Agents

66.7
#2

MMMU-Pro

Multimodal - Vision

79.4
#3

EQBench

Creative Writing

1561
#1

SciCode

Coding Agents

52.2
#1

MCP-Mark

Agents and Function Calling

55.9
#1

Apex Agents

Agents and Function Calling

27.9
#1

FrontierCode

Coding Agents

3.8
#2

LiveCodeBench

Coding Agents

89.6
#2

Terminal-Bench 2.0

Coding Agents

66.7
#2

MMMU-Pro

Multimodal - Vision

79.4

Use case:

Creative Writing

Features:

Low Latency