Contact Us 1-800-596-4880

Viewing Token Usage and Model Proxy Metrics

Token Usage reports and API Manager insights help you monitor how your model proxies consume resources and perform over time. Token Usage reports show the number of API tokens consumed per model, broken down by business group or application. The API Manager LLM Summary page provides hourly metrics including request volume, policy violations, errors, and average response time. For deeper analysis, you can build custom dashboards in Anypoint Monitoring.

View Token Usage Reports for Model Proxies

With Token Usage reports, you can view the amount of API tokens each Model Proxy uses for individual models.

To limit token usage, apply the LLM Token Based Rate Limit Policy to your Model Proxy.

To view token usage reports:

  1. In Anypoint Platform, click your Anypoint Platform profile icon (with your initials).

  2. Click Usage Reports.

  3. Select Model Proxy for Product.

  4. Filter between Usage by Business Group and Usage by Application to see how tokens are consumed.

To learn more about Usage Reports, see Usage Reports.

Gemini models use reasoning tokens. This number is included in Total tokens but isn’t individually listed.

View Model Proxy Metrics in API Manager

The API Manager LLM Summary page provides these metrics for your Model Proxy:

  • Total requests per hour

  • Total policy violations per hour

  • Total errors per hour

  • Average response time per hour

To view the LLM Summary page for a Model Proxy:

  1. From API Manager, click Model Proxies.

  2. Click the name of the Model Proxy you want to view.

To view more detailed metrics and build custom dashboards, click View more metrics in Anypoint Monitoring dashboard.