OpenAI Models Speedtest

A client-side streaming benchmark for comparing time to first token, end-to-end latency, and output throughput. Your API key is used only for direct requests to api.openai.com.

Benchmark setup

Saved in this browser's localStorage. It never leaves the browser except in the Authorization header sent to api.openai.com.

Ready. Add models, prices, then run.

Models and editable prices

Prices are editable estimates and can drift; verify them against OpenAI's current pricing page before using this for cost decisions. “FAST” sends service_tier: "priority"; unsupported models are recorded as failures without substituting another model.

Test?ModelModeNormal input
$/1M
Normal output
$/1M
FAST input
$/1M
FAST output
$/1M

Cost is computed from returned input_tokens and output_tokens; cached/reasoning tokens are not separately priced by this lightweight tool.

Results

$0.0000 / $0.000 / 0 tries

Results will appear as attempts finish.

ModelModeDoneFailuresTTFT median
(min–max)
Total medianTokens/sec medianCostThroughputTTFT (lower is better)
No results yet.