Monitor AI model performance on MicroShift

After your AI model is serving traffic, you can collect model-server metrics to identify bottlenecks and optimize resource allocation on your edge device.

MicroShift supports two methods for accessing model-server metrics: querying the Prometheus-format /metrics endpoint directly, or exporting metrics through OpenTelemetry if the microshift-observability RPM is installed.