The best open LLMs for your use case:
Agentic-focused refinement of GLM-5.1 from Z.ai with improved coding, tool use, and reasoning, plus extended 256K context.
Speed:
Intelligence:
Price: (1M Tokens)
$1.40 / 4.40Cached input: (1M Tokens)
$0.26Context: (tokens)
262,144Inputs:
Benchmarks:
FrontierSWE
Coding Agents
Terminal-Bench 2.1
Coding Agents
DeepSWE
Coding Agents
SWE-Bench Pro
Coding Agents
HLE
General Knowledge
MCP-Atlas
Agents and Function Calling
EQBench
Creative Writing
GPQA-Diamond
General Knowledge
FrontierSWE
Coding Agents
Terminal-Bench 2.1
Coding Agents
DeepSWE
Coding Agents
SWE-Bench Pro
Coding Agents
HLE
General Knowledge
MCP-Atlas
Agents and Function Calling
EQBench
Creative Writing
GPQA-Diamond
General Knowledge
Coding-specialized variant of Moonshot's Kimi K2.7 with long-context support and strong agentic software engineering performance.
Speed:
Intelligence:
Price: (1M Tokens)
$0.95 / 4.00Cached input: (1M Tokens)
$0.19Context: (tokens)
262,144Inputs:
Benchmarks:
MLS Bench Lite
Coding Agents
Program Bench
Coding Agents
Kimi Code Bench v2
Coding Agents
MCP-Mark Verified
Agents and Function Calling
Kimi Claw 24/7 Bench
Agents and Function Calling
MCP-Atlas
Agents and Function Calling
MLS Bench Lite
Coding Agents
Program Bench
Coding Agents
Kimi Code Bench v2
Coding Agents
MCP-Mark Verified
Agents and Function Calling
Kimi Claw 24/7 Bench
Agents and Function Calling
MCP-Atlas
Agents and Function Calling
Use case:
Coding Agents
Features:
Long Context Handling