Skip to content

EthenEthenEthen

Google · Flagships Analysis

Gemini 2.5 Flash-Lite

Intelligence, Performance & Price Analysis

Canonical slug: gemini-2-5-flash-lite-reasoning · Canonical model registry at build time

Intelligence

11 score

Speed

287.0 output tokens/sec

Latency

29.36s TTFT

Input Price

$0.10 / 1M tokens

Output Price

$0.40 / 1M tokens

Executive Assessment

Routing Verdict & Tradeoffs

Ethen Routing Verdict

Capable everyday model

Gemini 2.5 Flash-Lite (Reasoning) scores 11 (estimated) on the Artificial Analysis Intelligence Index, placing it below average among other reasoning models in a similar price tier (median: 15). Gemini 2.5 Flash-Lite (Reasoning) generates output at 287.0 tokens per second (based on Google's API), which is well above average compared to other reasoning models in a similar price tier (median: 99.1 t/s). Gemini 2.5 Flash-Lite (Reasoning) costs $0.10 per 1M input tokens (very competitive, median: $0.25) and $0.40 per 1M output tokens (very competitive, median: $0.87), based on Google's API.

Good for routine tasks; route complex reasoning and premium workloads to stronger models.

Task Fit Assessment

WorkloadRatingNotes
Complex reasoning & agentic workflowsviableIntelligence score 11 handles routine reasoning but may struggle with open-ended agentic tasks.
High-volume chat & customer-facingoptimalOutput speed 287.0 tokens/sec and capable intelligence make this suitable for real-time chat at scale.
Latency-sensitive applicationsviableTTFT 29.36s — latency may be noticeable in interactive use.
Cost-sensitive pipelinesoptimalOutput pricing at $0.40 is very competitive for high-volume workloads.

Cost Pressure Analysis

Low

Pricing is competitive — input $0.10, output $0.40. Suitable for sustained production use.

Technical Specifications

Architecture and Limits

SpecificationValueValidation Authority
Model typeProprietaryinferred
ReasoningYesfaq
Input modalitiesGemini 2.5 Flash-Lite (Reasoning) supports text, image, speech, and video input.faq
Output modalitiesGemini 2.5 Flash-Lite (Reasoning) supports text output.faq
Context windowGemini 2.5 Flash-Lite (Reasoning) has a context window of 1.0M tokens. This determines how much text and conversation history the model can process in a single request.faq
Open weights / sourceNo, Gemini 2.5 Flash-Lite (Reasoning) is proprietary. The model weights are not publicly available.faq
ParametersGemini 2.5 Flash-Lite (Reasoning) is a proprietary model and Google has not disclosed the model size or parameter count.faq
Knowledge cutoffGemini 2.5 Flash-Lite (Reasoning) has a knowledge cutoff of January 2025. The model's training data includes information up to this date.faq
API availabilityYes, Gemini 2.5 Flash-Lite (Reasoning) is available via API through 1 provider.faq

Evidence charts

Profile visualizations

Charts are the same canonical evidence cards previously published for this slug, contained inside the D18D page grammar.

AA-Omniscience Index

AA-Omniscience Index (higher is better) measures knowledge reliability and hallucination. It rewards correct answers, penalizes hallucinations, and has no penalty for refusing to answer. Scores range from -100 to 100, where 0 means as many correct as incorrect answers, and negative scores mean more incorrect than correct. · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 4040AClaude Fable 5 (with fal…Gemini 3.5 Flash: 2323GGemini 3.5 FlashGPT-5.5 (xhigh): 2020AIGPT-5.5 (xhigh)Grok 4.3 (high): 1818xGrok 4.3 (high)Kimi K2.6: 6.46.4KKimi K2.6Muse Spark: 4.14.1?Muse SparkGLM-5.2 (max): 44?GLM-5.2 (max)MiMo-V2.5-Pro: 3.63.6?MiMo-V2.5-ProQwen3.7 Plus: 2.42.4QQwen3.7 PlusMiniMax-M3: 1.41.4?MiniMax-M3

Artificial Analysis Intelligence Index by Open Weights / Proprietary

Artificial Analysis Intelligence Index v4.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkMiMo-V2.5-Pro: 4242?MiMo-V2.5-ProNex-N2-Pro: 4141?Nex-N2-Pro

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkMiMo-V2.5-Pro: 4242?MiMo-V2.5-ProNex-N2-Pro: 4141?Nex-N2-Pro

Intelligence

Artificial Analysis Intelligence Index · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0204060Claude Fable 5 (with fallback): 6060AClaude Fable 5 (with fal…GPT-5.5 (xhigh): 5555AIGPT-5.5 (xhigh)GLM-5.2 (max): 5151?GLM-5.2 (max)Gemini 3.5 Flash: 5050GGemini 3.5 FlashMiniMax-M3: 4444?MiniMax-M3DeepSeek V4 Pro (max): 4444DDeepSeek V4 Pro (max)Kimi K2.6: 4444KKimi K2.6Muse Spark: 4343?Muse SparkNemotron 3 Ultra: 3838NNemotron 3 UltraGrok 4.3 (high): 3838xGrok 4.3 (high)Gemini 2.5 Flash-Lite: 1111GGemini 2.5 Flash-Lite

Output Speed

Output tokens per second · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0113225338450Step 3.7 Flash: 410410?Step 3.7 FlashNVIDIA Nemotron 3 Super: 356356NNVIDIA Nemotron 3 SuperGemini 3.1 Flash-Lite: 339339GGemini 3.1 Flash-Litegpt-oss-120b (high): 316316AIgpt-oss-120b (high)Gemini 2.5 Flash-Lite: 287287GGemini 2.5 Flash-LiteNemotron 3 Ultra: 248248NNemotron 3 UltraGLM-5.2 (max): 216216?GLM-5.2 (max)Gemini 3.5 Flash: 192192GGemini 3.5 FlashQwen3.6 35B A3B: 166166QQwen3.6 35B A3BGrok 4.3 (high): 164164xGrok 4.3 (high)

Speed

Output tokens per second · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

088175263350gpt-oss-120b (high): 316316AIgpt-oss-120b (high)Gemini 2.5 Flash-Lite: 287287GGemini 2.5 Flash-LiteNemotron 3 Ultra: 248248NNemotron 3 UltraGLM-5.2 (max): 216216?GLM-5.2 (max)Gemini 3.5 Flash: 192192GGemini 3.5 FlashGrok 4.3 (high): 164164xGrok 4.3 (high)MiniMax-M3: 9595?MiniMax-M3GPT-5.5 (xhigh): 8888AIGPT-5.5 (xhigh)Kimi K2.6: 7676KKimi K2.6DeepSeek V4 Pro (max): 7272DDeepSeek V4 Pro (max)

End-to-End Response Time

Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sGLM-4.6: 10s10s?GLM-4.6Qwen3.7 Plus: 9.9s9.9sQQwen3.7 PlusMiniMax-M2.7: 9.8s9.8s?MiniMax-M2.7MiMo-V2.5-Pro: 9.5s9.5s?MiMo-V2.5-ProNex-N2-Pro: 5.4s5.4s?Nex-N2-ProMiniMax-M3: 5.2s5.2s?MiniMax-M3DeepSeek V4 Flash (max): 4.4s4.4sDDeepSeek V4 Flash (max)GLM-4.7: 3.8s3.8s?GLM-4.7Ring-2.6-1T: 3.7s3.7s?Ring-2.6-1TGPT-5.4 nano (xhigh): 3.3s3.3sAIGPT-5.4 nano (xhigh)Gemini 2.5 Flash-Lite: 1.7s1.7sGGemini 2.5 Flash-Lite

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0s15s30s45s60sDeepSeek V4 Flash (max): 50s50sDDeepSeek V4 Flash (max)MiniMax-M2.7: 48s48s?MiniMax-M2.7GLM-4.6: 41s41s?GLM-4.6Qwen3.7 Plus: 39s39sQQwen3.7 PlusMiMo-V2.5-Pro: 38s38s?MiMo-V2.5-ProQwen3.6 35B A3B: 33s33sQQwen3.6 35B A3BNex-N2-Pro: 22s22s?Nex-N2-ProMiniMax-M3: 21s21s?MiniMax-M3GLM-4.7: 15s15s?GLM-4.7Ring-2.6-1T: 15s15s?Ring-2.6-1TGemini 2.5 Flash-Lite: 0s0sGGemini 2.5 Flash-Lite

Context Window

Context window: tokens limit · Higher is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0250k500k750k1MGemini 2.5 Flash-Lite: 1M1MGGemini 2.5 Flash-LiteGemini 3.5 Flash: 1M1MGGemini 3.5 FlashClaude Fable 5 (with fallback): 1M1MAClaude Fable 5 (with fal…GLM-5.2 (max): 1M1M?GLM-5.2 (max)DeepSeek V4 Pro (max): 1M1MDDeepSeek V4 Pro (max)Grok 4.3 (high): 1M1MxGrok 4.3 (high)MiniMax-M3: 1M1M?MiniMax-M3MiMo-V2.5-Pro: 1M1M?MiMo-V2.5-ProDeepSeek V4 Flash (max): 1M1MDDeepSeek V4 Flash (max)Qwen3.7 Plus: 1M1MQQwen3.7 Plus

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

00.51GPT-5.5 (xhigh): 0.10.1AIGPT-5.5 (xhigh)Gemini 3.5 Flash: 0.10.1GGemini 3.5 FlashGLM-5.2 (max): 00?GLM-5.2 (max)Ring-2.6-1T: 00?Ring-2.6-1TKimi K2.6: 00KKimi K2.6Qwen3.6 35B A3B: 00QQwen3.6 35B A3BNemotron 3 Ultra: 00NNemotron 3 UltraGLM-4.7: 00?GLM-4.7GLM-4.6: 00?GLM-4.6MiniMax-M3: 00?MiniMax-M3

Cost per Task

Weighted average cost (USD) per Intelligence Index task · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

$0.00$1.00$2.00$3.00$4.00$5.00Claude Fable 5 (with fallback): $2.75$2.75AClaude Fable 5 (with fal…GPT-5.5 (xhigh): $0.86$0.86AIGPT-5.5 (xhigh)Gemini 3.5 Flash: $0.59$0.59GGemini 3.5 FlashGLM-5.2 (max): $0.37$0.37?GLM-5.2 (max)Kimi K2.6: $0.35$0.35KKimi K2.6Nemotron 3 Ultra: $0.24$0.24NNemotron 3 UltraGrok 4.3 (high): $0.14$0.14xGrok 4.3 (high)MiniMax-M3: $0.12$0.12?MiniMax-M3gpt-oss-120b (high): $0.06$0.06AIgpt-oss-120b (high)DeepSeek V4 Pro (max): $0.04$0.04DDeepSeek V4 Pro (max)

Cost to Run Artificial Analysis Intelligence Index

Cost (USD) to run all evaluations in the Artificial Analysis Intelligence Index · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

050100150200Qwen3.6 35B A3B: 165165QQwen3.6 35B A3BGPT-5.5 (xhigh): 132132AIGPT-5.5 (xhigh)Gemini 3.5 Flash: 114114GGemini 3.5 FlashNemotron 3 Ultra: 3434NNemotron 3 UltraGLM-5.2 (max): 3333?GLM-5.2 (max)Kimi K2.6: 2525KKimi K2.6Ring-2.6-1T: 2424?Ring-2.6-1TGLM-4.7: 1717?GLM-4.7GLM-4.6: 1616?GLM-4.6MiniMax-M3: 1515?MiniMax-M3

Pricing: Cache Hit, Input, and Output

Price (USD per M Tokens) · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

0250k500k750k1MNemotron 3 Ultra: 3.63.6NNemotron 3 UltraGLM-4.7: 3.33.3?GLM-4.7Nex-N2-Pro: 3.33.3?Nex-N2-ProRing-2.6-1T: 2.82.8?Ring-2.6-1TGLM-4.6: 2.82.8?GLM-4.6GPT-5 mini (high): 2.32.3AIGPT-5 mini (high)Gemini 3.1 Flash-Lite: 1.81.8GGemini 3.1 Flash-LiteQwen3.6 35B A3B: 1.71.7QQwen3.6 35B A3BQwen3.7 Plus: 1.61.6QQwen3.7 PlusMiniMax-M3: 1.61.6?MiniMax-M3Gemini 2.5 Flash-Lite: 0.510.51GGemini 2.5 Flash-Litecache Hitinputoutput

Time per Intelligence Index Task

Weighted average decode time (minutes) per task; excludes TTFT and overhead time · Lower is better · Evaluation results measured independently by Artificial Analysis

Retrieved 2026-07-08

Chart source and provenance are listed in Methodology & sources below.

Retrieval date:
2026-07-08
Methodology:
Independent test run by Artificial Analysis on dedicated hardware.

02468MiMo-V2.5-Pro: 6.16.1?MiMo-V2.5-ProGLM-4.6: 5.95.9?GLM-4.6DeepSeek V4 Flash (max): 5.85.8DDeepSeek V4 Flash (max)MiniMax-M2.7: 5.55.5?MiniMax-M2.7Nex-N2-Pro: 4.84.8?Nex-N2-ProClaude Fable 5 (with fallback): 4.84.8AClaude Fable 5 (with fal…MiniMax-M3: 44?MiniMax-M3GPT-5.5 (xhigh): 33AIGPT-5.5 (xhigh)Ring-2.6-1T: 2.92.9?Ring-2.6-1TGLM-5.2 (max): 2.82.8?GLM-5.2 (max)

Methodology

Methodology & Provenance

This page is rendered from the normalized profile and page JSON for Gemini 2.5 Flash-Lite.

Benchmark values are preserved as normalized; only layout, disclosure ordering, and typography are adjusted for readability.

Frequently Asked Questions

Model FAQs & Technical Disclosures

Gemini 2.5 Flash-Lite (Reasoning) was released on June 17, 2025.