Skip to content

EthenEthenEthen

Benchmark Dimension

Latency Benchmark

The Latency Benchmark measures time to first answer token (TTFT), representing how quickly a model begins generating after receiving a prompt. Lower latency is critical for interactive and real-time applications. Measured values include reasoning prep time where applicable.

Unit: s TTFT · Direction: Lower is better

Ranked standings

Certified model rankings

Certification

Not certified — No benchmark result

This benchmark has no certified, evidence-backed result. Rankings are unavailable until canonical results are published.