The best open LLMs for your use case:
Agentic-focused refinement of GLM-5.1 from Z.ai with improved coding, tool use, and reasoning, plus extended 256K context.
Speed:
Intelligence:
Price: (1M Tokens)
$1.40 / 4.40Cached input: (1M Tokens)
$0.26Context: (tokens)
262,144Inputs:
Benchmarks:
MCP-Atlas
Agents and Function Calling
HLE
General Knowledge
SWE-Bench Pro
Coding Agents
DeepSWE
Coding Agents
Terminal-Bench 2.1
Coding Agents
FrontierSWE
Coding Agents
EQBench
Creative Writing
GPQA-Diamond
General Knowledge
MCP-Atlas
Agents and Function Calling
HLE
General Knowledge
SWE-Bench Pro
Coding Agents
DeepSWE
Coding Agents
Terminal-Bench 2.1
Coding Agents
FrontierSWE
Coding Agents
EQBench
Creative Writing
GPQA-Diamond
General Knowledge
Next-generation reasoning model from MiniMax with frontier agentic, coding, and multimodal performance. Strong scores on SWE-Bench, BrowseComp, OmniDocBench, and IMO/USAMO competition reasoning.
Speed:
Intelligence:
Price: (1M Tokens)
$0.30 / 1.20Cached input: (1M Tokens)
$0.06Context: (tokens)
524,288Inputs:
Benchmarks:
Apex Agents
Agents and Function Calling
Claw-Eval
Agents and Function Calling
GPQA-Diamond
General Knowledge
Video-MME v2
Multimodal - Vision
SWE-Bench Verified
Coding Agents
SWE-Bench Pro
Coding Agents
MMMU-Pro
Multimodal - Vision
Apex Agents
Agents and Function Calling
Claw-Eval
Agents and Function Calling
GPQA-Diamond
General Knowledge
Video-MME v2
Multimodal - Vision
SWE-Bench Verified
Coding Agents
SWE-Bench Pro
Coding Agents
MMMU-Pro
Multimodal - Vision
Use case:
Agents and Function Calling
Features:
Low Latency