Sort providers by cost, latency, or throughput on AI Gateway
You can now sort the providers behind a model by cost, time to first token (TTFT), or throughput (TPS) in .AI Gateway The default provider order blends provider reliability, quality of model output, cost, and speed of response. You can now use for explicit control over ranking criteria.sort For models with many providers and noticeable cost or speed variation, you can use to optimize on your dimension of choice. Ranking is computed at request time, so newly added providers, price changes, and
