Skip to content

EthenEthenEthen

· Flagships Analysis

Multilingual AI Model Benchmark - Compare Leading LLMs by Language

Intelligence, Performance & Price Analysis

Canonical slug: multilingual · Canonical model registry at build time

Executive Assessment

Routing Verdict & Tradeoffs

Ethen Routing Verdict

General purpose assistant

Compare the multilingual performance, pricing, and speed of leading AI language models. See how top LLMs like Gemini 3.1 Pro Preview, Gemini 3 Pro Preview (high), and Claude Opus 4.6 (max) perform across many languages.

Route complex or high-value tasks here when the extra capability justifies the cost.

Task Fit Assessment

No fit matrix is published for this profile.

Evidence charts

Profile visualizations

Charts are the same canonical evidence cards previously published for this slug, contained inside the D18D page grammar.

Output Speed

Output tokens per second · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

088175263350gpt-oss-120b (high): 314314AIgpt-oss-120b (high)Llama 4 Maverick: 118118MLlama 4 MaverickMiniMax-M2.5: 8686?MiniMax-M2.5MiniMax-M2.1: 7878?MiniMax-M2.1Claude Opus 4.5: 6161AClaude Opus 4.5Claude 4.5 Sonnet: 5151AClaude 4.5 SonnetMagistral Medium 1.2: 4040?Magistral Medium 1.2

End-to-End Response Time

Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sMagistral Medium 1.2: 13s13s?Magistral Medium 1.2Claude 4.5 Sonnet: 9.8s9.8sAClaude 4.5 SonnetClaude Opus 4.5: 8.2s8.2sAClaude Opus 4.5MiniMax-M2.1: 6.4s6.4s?MiniMax-M2.1MiniMax-M2.5: 5.8s5.8s?MiniMax-M2.5Llama 4 Maverick: 4.2s4.2sMLlama 4 Maverickgpt-oss-120b (high): 1.6s1.6sAIgpt-oss-120b (high)

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sMagistral Medium 1.2: 51s51s?Magistral Medium 1.2MiniMax-M2.1: 26s26s?MiniMax-M2.1MiniMax-M2.5: 23s23s?MiniMax-M2.5gpt-oss-120b (high): 6.4s6.4sAIgpt-oss-120b (high)Llama 4 Maverick: 0s0sMLlama 4 MaverickClaude 4.5 Sonnet: 0s0sAClaude 4.5 SonnetClaude Opus 4.5: 0s0sAClaude Opus 4.5

Pricing: Cache Hit, Input, and Output

Price (USD per M Tokens) · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0250k500k750k1MGrok 4: 3333xGrok 4Claude Opus 4.5: 3131AClaude Opus 4.5Claude 4.5 Sonnet: 1818AClaude 4.5 SonnetGPT-5.2 (medium): 1717AIGPT-5.2 (medium)Gemini 3 Pro Preview (high): 1414GGemini 3 Pro Preview (hi…Magistral Medium 1.2: 77?Magistral Medium 1.2MiniMax-M2.5: 1.51.5?MiniMax-M2.5MiniMax-M2.1: 1.51.5?MiniMax-M2.1Llama 4 Maverick: 1.51.5MLlama 4 Maverickgpt-oss-120b (high): 0.90.9AIgpt-oss-120b (high)inputoutputcache Hit

Methodology

Methodology & Provenance

This page is rendered from the normalized profile and page JSON for Multilingual AI Model Benchmark - Compare Leading LLMs by Language.

Benchmark values are preserved as normalized; only layout, disclosure ordering, and typography are adjusted for readability.