The best open LLMs for your use case:

1Qwen3.5 397B-A17BQwen

Qwen's native multimodal MoE model with 397B total parameters and 17B active, featuring hybrid Gated Delta Networks for strong reasoning and vision capabilities.

Speed:

Intelligence:

Price: (1M Tokens)

$0.60 / 3.60

Context: (tokens)

262,144

Inputs:

ImageText

Benchmarks:

#1

Multilingual MMLU

Multilingual

88.5
#1

MMLU-Pro

General Knowledge

87.8
#1

AA-LCR

Summarization

68.7
#1

LongBenchv2

Summarization

63.2
#1

MMMU

Multimodal - Vision

85
#1

TAU2-Bench

Agents and Function Calling

86.7
#1

BFCL

Agents and Function Calling

72.9
#2

MMMU-Pro

Multimodal - Vision

79
#1

Multilingual MMLU

Multilingual

88.5
#1

MMLU-Pro

General Knowledge

87.8
#1

AA-LCR

Summarization

68.7
#1

LongBenchv2

Summarization

63.2
#1

MMMU

Multimodal - Vision

85
#1

TAU2-Bench

Agents and Function Calling

86.7
#1

BFCL

Agents and Function Calling

72.9
#2

MMMU-Pro

Multimodal - Vision

79
2Gemma 4 31BGoogle

Google's open multimodal model with strong multilingual coverage, 256K context, and text + image input.

Speed:

Intelligence:

Price: (1M Tokens)

$0.39 / 0.97

Context: (tokens)

262,144

Inputs:

ImageText

Benchmarks:

#2

Multilingual MMLU

Multilingual

88.4
#1

MRCR 128k

Summarization

66.4
#3

MMLU-Pro

General Knowledge

85.2
#3

TAU2-Bench

Agents and Function Calling

76.9
#4

LiveCodeBench

Coding Agents

80
#4

MMMU-Pro

Multimodal - Vision

76.9
#5

HLE

General Knowledge

19.5
#6

GPQA-Diamond

General Knowledge

84.3
#2

Multilingual MMLU

Multilingual

88.4
#1

MRCR 128k

Summarization

66.4
#3

MMLU-Pro

General Knowledge

85.2
#3

TAU2-Bench

Agents and Function Calling

76.9
#4

LiveCodeBench

Coding Agents

80
#4

MMMU-Pro

Multimodal - Vision

76.9
#5

HLE

General Knowledge

19.5
#6

GPQA-Diamond

General Knowledge

84.3

Use case:

Multilingual

Features:

Long Context Handling