The best open LLMs for your use case:
Moonshot AI's 2.8-trillion-parameter open-weight frontier model with native vision, 1M-token context, and state-of-the-art agentic coding and tool use.
Speed:
Intelligence:
Price: (1M Tokens)
$3.00 / 15.00Cached input: (1M Tokens)
$0.30Context: (tokens)
1,048,576Inputs:
Benchmarks:
GPQA-Diamond
General Knowledge
HLE
General Knowledge
MMMU-Pro
Multimodal - Vision
Terminal-Bench 2.1
Coding Agents
FrontierSWE
Coding Agents
DeepSWE
Coding Agents
Program Bench
Coding Agents
MCP-Atlas
Agents and Function Calling
GPQA-Diamond
General Knowledge
HLE
General Knowledge
MMMU-Pro
Multimodal - Vision
Terminal-Bench 2.1
Coding Agents
FrontierSWE
Coding Agents
DeepSWE
Coding Agents
Program Bench
Coding Agents
MCP-Atlas
Agents and Function Calling
1T-parameter MoE flagship from Moonshot with long-horizon coding, agent swarms scaling to 300 sub-agents, and state-of-the-art reasoning.
Speed:
Intelligence:
Price: (1M Tokens)
$1.20 / 4.50Cached input: (1M Tokens)
$0.20Context: (tokens)
262,144Inputs:
Benchmarks:
EQBench
Creative Writing
SciCode
Coding Agents
MCP-Mark
Agents and Function Calling
Apex Agents
Agents and Function Calling
FrontierCode
Coding Agents
LiveCodeBench
Coding Agents
Terminal-Bench 2.0
Coding Agents
MMMU-Pro
Multimodal - Vision
EQBench
Creative Writing
SciCode
Coding Agents
MCP-Mark
Agents and Function Calling
Apex Agents
Agents and Function Calling
FrontierCode
Coding Agents
LiveCodeBench
Coding Agents
Terminal-Bench 2.0
Coding Agents
MMMU-Pro
Multimodal - Vision
Use case:
Creative Writing
Features:
Low Latency