OpenAI (Direct)
OpenAI AI Gateway Benchmark

Request Latency

Baseline
100% success

No gateway involved — a direct call to Anthropic's API, used as the no-gateway control every other provider is compared against.

Cold E2E
881ms
Median
Warm TTFT
643ms
Median
Tokens/sec
47
Median
DNS
2.9ms
Median
TCP
1.3ms
Median
TLS
3.9ms
Median

Cold Connection Breakdown

Where a fresh request's time goes: DNS lookup, TCP connect, TLS handshake, then time to first response byte, then first byte to first streamed token.

DNS2.9ms
TCP1.3ms
TLS3.9ms
To First Byte485.0ms
First Byte → Token381.7ms

Performance Over Time

Iteration Distribution

Cold E2E

Warm TTFT

Tokens/sec

DNS

TCP

TLS

Become a partner

Want your logo featured across our benchmarks? Become a ComputeSDK benchmark partner.

LatitudeGoogle Cloud Run
PartnersLatitude