Skip to content

EthenEthenEthen

xAI · Flagships Analysis

Grok 3 mini Reasoning (high)

Intelligence, Performance & Price Analysis

Canonical slug: grok-3-mini-reasoning · Canonical model registry at build time

Intelligence

23 score

Speed

83.0 output tokens/sec

Latency

0.74s TTFT

Input Price

$0.30 / 1M tokens

Output Price

$0.50 / 1M tokens

Executive Assessment

Routing Verdict & Tradeoffs

Ethen Routing Verdict

Strong general-purpose model

Grok 3 mini Reasoning (high) scores 23 (estimated) on the Artificial Analysis Intelligence Index, placing it above average among other reasoning models in a similar price tier (median: 15). Grok 3 mini Reasoning (high) generates output at 83.0 tokens per second (based on xAI's API), which is below average compared to other reasoning models in a similar price tier (median: 99.0 t/s). Grok 3 mini Reasoning (high) costs $0.30 per 1M input tokens (somewhat higher than average, median: $0.25) and $0.50 per 1M output tokens (better than average, median: $0.87), based on xAI's API.

Suitable for most production tasks, but high-volume or repetitive work should still be compared against cheaper routes.

Task Fit Assessment

WorkloadRatingNotes
Complex reasoning & agentic workflowsoptimalIntelligence score 23 supports capable reasoning, but very hard tasks may benefit from higher-tier models.
High-volume chat & customer-facingoptimalOutput speed 83.0 tokens/sec is adequate for chat.
Latency-sensitive applicationsoptimalTTFT 0.74s — among the lowest latencies, suitable for interactive latency-critical use cases.
Cost-sensitive pipelinesoptimalOutput pricing at $0.50 is reasonable for moderate volume.

Cost Pressure Analysis

Low

Pricing is competitive — input $0.30, output $0.50. Suitable for sustained production use.

Technical Specifications

Architecture and Limits

SpecificationValueValidation Authority
Model typeProprietaryinferred
ReasoningYesfaq
Input modalitiesGrok 3 mini Reasoning (high) supports text input.faq
Output modalitiesGrok 3 mini Reasoning (high) supports text and image output.faq
Context windowGrok 3 mini Reasoning (high) has a context window of 1.0M tokens. This determines how much text and conversation history the model can process in a single request.faq
Open weights / sourceNo, Grok 3 mini Reasoning (high) is proprietary. The model weights are not publicly available.faq
ParametersGrok 3 mini Reasoning (high) is a proprietary model and xAI has not disclosed the model size or parameter count.faq
API availabilityYes, Grok 3 mini Reasoning (high) is available via API through 3 providers.faq

Evidence charts

Profile visualizations

Charts are the same canonical evidence cards previously published for this slug, contained inside the D18D page grammar.

AA-Omniscience Index

AA-Omniscience Index (higher is better) measures knowledge reliability and hallucination. It rewards correct answers, penalizes hallucinations, and has no penalty for refusing to answer. Scores range from -100 to 100, where 0 means as many correct as incorrect answers, and negative scores mean more incorrect than correct. · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

-200204060Claude Fable 5 (with fallback): 4040AClaude Fable 5 (with fal…Gemini 3.5 Flash: 2323GGemini 3.5 FlashGPT-5.5 (xhigh): 2020AIGPT-5.5 (xhigh)Grok 4.3 (high): 1818xGrok 4.3 (high)Kimi K2.6: 6.46.4KKimi K2.6Muse Spark: 4.14.1?Muse SparkGLM-5.2 (max): 44?GLM-5.2 (max)MiMo-V2.5-Pro: 3.63.6?MiMo-V2.5-ProQwen3.7 Plus: 2.42.4QQwen3.7 PlusMiniMax-M3: 1.41.4?MiniMax-M3Grok 3 mini Reasoning (high): -6-6xGrok 3 mini Reasoning (h…

Artificial Analysis Intelligence Index by Open Weights / Proprietary

Artificial Analysis Intelligence Index v4.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkMiMo-V2.5-Pro: 4242?MiMo-V2.5-ProNex-N2-Pro: 4141?Nex-N2-Pro

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkMiMo-V2.5-Pro: 4242?MiMo-V2.5-ProNex-N2-Pro: 4141?Nex-N2-Pro

Intelligence

Artificial Analysis Intelligence Index · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkNemotron 3 Ultra: 3838NNemotron 3 UltraGrok 4.3 (high): 3838xGrok 4.3 (high)Grok 3 mini Reasoning (high): 2323xGrok 3 mini Reasoning (h…

Output Speed

Output tokens per second · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0113225338450Step 3.7 Flash: 410410?Step 3.7 FlashNVIDIA Nemotron 3 Super: 364364NNVIDIA Nemotron 3 SuperGemini 3.1 Flash-Lite: 347347GGemini 3.1 Flash-Litegpt-oss-120b (high): 314314AIgpt-oss-120b (high)Nemotron 3 Ultra: 250250NNemotron 3 UltraGLM-5.2 (max): 218218?GLM-5.2 (max)Gemini 3.5 Flash: 192192GGemini 3.5 FlashQwen3.6 35B A3B: 166166QQwen3.6 35B A3BGrok 4.3 (high): 164164xGrok 4.3 (high)GPT-5.4 nano (xhigh): 153153AIGPT-5.4 nano (xhigh)Grok 3 mini Reasoning (high): 8383xGrok 3 mini Reasoning (h…

Speed

Output tokens per second · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

088175263350gpt-oss-120b (high): 314314AIgpt-oss-120b (high)Nemotron 3 Ultra: 250250NNemotron 3 UltraGLM-5.2 (max): 218218?GLM-5.2 (max)Gemini 3.5 Flash: 192192GGemini 3.5 FlashGrok 4.3 (high): 164164xGrok 4.3 (high)MiniMax-M3: 9999?MiniMax-M3GPT-5.5 (xhigh): 8888AIGPT-5.5 (xhigh)Grok 3 mini Reasoning (high): 8383xGrok 3 mini Reasoning (h…Kimi K2.6: 7676KKimi K2.6DeepSeek V4 Pro (max): 7272DDeepSeek V4 Pro (max)

End-to-End Response Time

Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sGLM-4.6: 10s10s?GLM-4.6Qwen3.7 Plus: 9.9s9.9sQQwen3.7 PlusMiniMax-M2.7: 9.8s9.8s?MiniMax-M2.7MiMo-V2.5-Pro: 9.5s9.5s?MiMo-V2.5-ProGrok 3 mini Reasoning (high): 6s6sxGrok 3 mini Reasoning (h…Nex-N2-Pro: 5.4s5.4s?Nex-N2-ProMiniMax-M3: 5s5s?MiniMax-M3DeepSeek V4 Flash (max): 4.5s4.5sDDeepSeek V4 Flash (max)GLM-4.7: 4s4s?GLM-4.7Ring-2.6-1T: 3.7s3.7s?Ring-2.6-1T

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sDeepSeek V4 Flash (max): 50s50sDDeepSeek V4 Flash (max)MiniMax-M2.7: 48s48s?MiniMax-M2.7GLM-4.6: 41s41s?GLM-4.6Qwen3.7 Plus: 40s40sQQwen3.7 PlusMiMo-V2.5-Pro: 38s38s?MiMo-V2.5-ProQwen3.6 35B A3B: 33s33sQQwen3.6 35B A3BGrok 3 mini Reasoning (high): 24s24sxGrok 3 mini Reasoning (h…Nex-N2-Pro: 22s22s?Nex-N2-ProMiniMax-M3: 20s20s?MiniMax-M3GLM-4.7: 16s16s?GLM-4.7

Context Window

Context window: tokens limit · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0250k500k750k1MGrok 3 mini Reasoning (high): 1M1MxGrok 3 mini Reasoning (h…Gemini 3.5 Flash: 1M1MGGemini 3.5 FlashClaude Fable 5 (with fallback): 1M1MAClaude Fable 5 (with fal…GLM-5.2 (max): 1M1M?GLM-5.2 (max)DeepSeek V4 Pro (max): 1M1MDDeepSeek V4 Pro (max)Grok 4.3 (high): 1M1MxGrok 4.3 (high)MiniMax-M3: 1M1M?MiniMax-M3MiMo-V2.5-Pro: 1M1M?MiMo-V2.5-ProDeepSeek V4 Flash (max): 1M1MDDeepSeek V4 Flash (max)Qwen3.7 Plus: 1M1MQQwen3.7 Plus

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

00.51GPT-5.5 (xhigh): 0.10.1AIGPT-5.5 (xhigh)Gemini 3.5 Flash: 0.10.1GGemini 3.5 FlashGLM-5.2 (max): 00?GLM-5.2 (max)Ring-2.6-1T: 00?Ring-2.6-1TKimi K2.6: 00KKimi K2.6Qwen3.6 35B A3B: 00QQwen3.6 35B A3BNemotron 3 Ultra: 00NNemotron 3 UltraGLM-4.7: 00?GLM-4.7GLM-4.6: 00?GLM-4.6MiniMax-M3: 00?MiniMax-M3

Cost per Task

Weighted average cost (USD) per Intelligence Index task · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

$0.00$1.00$2.00$3.00$4.00$5.00Claude Fable 5 (with fallback): $2.75$2.75AClaude Fable 5 (with fal…GPT-5.5 (xhigh): $0.86$0.86AIGPT-5.5 (xhigh)Gemini 3.5 Flash: $0.59$0.59GGemini 3.5 FlashGLM-5.2 (max): $0.37$0.37?GLM-5.2 (max)Kimi K2.6: $0.35$0.35KKimi K2.6Nemotron 3 Ultra: $0.24$0.24NNemotron 3 UltraGrok 4.3 (high): $0.14$0.14xGrok 4.3 (high)MiniMax-M3: $0.12$0.12?MiniMax-M3gpt-oss-120b (high): $0.06$0.06AIgpt-oss-120b (high)DeepSeek V4 Pro (max): $0.04$0.04DDeepSeek V4 Pro (max)

Cost to Run Artificial Analysis Intelligence Index

Cost (USD) to run all evaluations in the Artificial Analysis Intelligence Index · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

050100150200Qwen3.6 35B A3B: 165165QQwen3.6 35B A3BGPT-5.5 (xhigh): 132132AIGPT-5.5 (xhigh)Gemini 3.5 Flash: 114114GGemini 3.5 FlashNemotron 3 Ultra: 3434NNemotron 3 UltraGLM-5.2 (max): 3333?GLM-5.2 (max)Kimi K2.6: 2525KKimi K2.6Ring-2.6-1T: 2424?Ring-2.6-1TGLM-4.7: 1717?GLM-4.7GLM-4.6: 1616?GLM-4.6MiniMax-M3: 1515?MiniMax-M3

Pricing: Cache Hit, Input, and Output

Price (USD per M Tokens) · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0250k500k750k1MNemotron 3 Ultra: 3.63.6NNemotron 3 UltraGLM-4.7: 3.33.3?GLM-4.7Nex-N2-Pro: 3.33.3?Nex-N2-ProRing-2.6-1T: 2.82.8?Ring-2.6-1TGLM-4.6: 2.82.8?GLM-4.6GPT-5 mini (high): 2.32.3AIGPT-5 mini (high)Gemini 3.1 Flash-Lite: 1.81.8GGemini 3.1 Flash-LiteQwen3.6 35B A3B: 1.71.7QQwen3.6 35B A3BQwen3.7 Plus: 1.61.6QQwen3.7 PlusMiniMax-M3: 1.61.6?MiniMax-M3Grok 3 mini Reasoning (high): 0.880.88xGrok 3 mini Reasoning (h…cache Hitinputoutput

Time per Intelligence Index Task

Weighted average decode time (minutes) per task; excludes TTFT and overhead time · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

02468DeepSeek V4 Flash (max): 6.36.3DDeepSeek V4 Flash (max)MiMo-V2.5-Pro: 6.16.1?MiMo-V2.5-ProGLM-4.6: 5.95.9?GLM-4.6MiniMax-M2.7: 5.55.5?MiniMax-M2.7Nex-N2-Pro: 4.84.8?Nex-N2-ProClaude Fable 5 (with fallback): 4.84.8AClaude Fable 5 (with fal…MiniMax-M3: 3.93.9?MiniMax-M3GPT-5.5 (xhigh): 33AIGPT-5.5 (xhigh)Ring-2.6-1T: 2.92.9?Ring-2.6-1TGLM-5.2 (max): 2.82.8?GLM-5.2 (max)

Methodology

Methodology & Provenance

This page is rendered from the normalized profile and page JSON for Grok 3 mini Reasoning (high).

Benchmark values are preserved as normalized; only layout, disclosure ordering, and typography are adjusted for readability.

Frequently Asked Questions

Model FAQs & Technical Disclosures

Grok 3 mini Reasoning (high) was released on February 19, 2025.