How do I track usage and spend per model?
Last updated: September 17, 2026
Three ways, from quickest to most powerful:
Dashboard — click Model APIs in the sidebar for per-model views.
Usage endpoint —GET /v1/model_apis/usagereturns request counts and token counts (input, cached input, uncached input, output) in 1-minute, 1-hour, or 1-day buckets, grouped or filtered by API key, user, or model. This is the tool for "which key/teammate/model is spending the money."
Prometheus export — baseten_model_api_tokens_total and baseten_model_api_inference_requests_totalstream the same numbers into your own observability stack.
Note: usage data is available from August 5, 2026 onward — earlier Model API usage wasn't backfilled.
Docs: Usage · Metrics export