Model performance

Cost, requests, tokens, latency, and time to first token for each model.

Model performance lists each model with its own totals and runtime metrics.

For each model you can see:

  • Cost
  • Number of requests
  • Total tokens
  • Latency
  • Time to first token (TTFT)
  • The other per-model metrics shown on the row
Photo 3 — Model performance
On this page

On this page

No Headings