Google's most cost-effective model. Highest throughput at the lowest price in the Gemini lineup.
Performance
Time to first token
435ms
↓ 7%vs prior 24h
Total response time
659ms
↓ 3%vs prior 24h
Throughput
97.8tok/s
↑ 3%vs prior 24h
Inter-token latency
91.7ms
↑ 6%vs prior 24h