Hi
@limited-pizza-33551,
I’m currently working on the GPU View stack for Portainer. The goal is to consume DCGM Exporter metrics and surface the relevant GPU information and utilization metrics directly in the Portainer dashboard.
Does Rancher have anything similar in terms of GPU monitoring/visibility that I could use as a reference for the dashboard design and the metrics they expose?
I’ve already built a script that collects and reports the GPU stack details across the Kubernetes cluster. You can take a look here:
github.com/ShubhamTatvamasi/gpu-operator/blob/…/gpu-cluster-report.sh
Any references or examples from Rancher would be really helpful as I shape the GPU view.
Thanks