Benchmarks / AI Inference
AI Inference benchmarks
Time to first token, streaming throughput, cold starts and tail latency across AI inference providers — measured continuously, not quoted from launch blogs.
Planned — not yet measuringWant this category sooner?
The benchmark framework is one Go interface — category suites are designed in the open. Propose metrics or contribute tests.