OpenAI's compact reasoning model. Fast reasoning at lower cost.
Performance
Time to first token
—ms
—vs prior 24h
Total response time
Throughput
—tok/s
Inter-token latency