Our benchmark suite is being finalized. So that the results mean something when they land, this page states the standard we hold them to up front: the exact setup, what each workload does on every platform, and the honest limits of what a given comparison can and can't tell you. We would rather publish numbers you can reproduce than numbers that flatter us.
fortio) at fixed concurrency with keep-alive, reporting full latency percentiles and checking for server saturation.text/event-stream response from a ReadableStream (the AI-gateway shape, minus the model).SELECT per request over the network, isolating the connection model (pooled session vs. reconnect-per-request vs. embedded store).