This message was deleted.
# announcements
s
This message was deleted.
c
Hi Jules, we don’t have any plan on that yet, how do you deploy BentoML api server today? have you tried using a DCGM-Exporter side-car?
j
Hey @Chaoyu we have 3 different Bento services deployed on a GKE cluster at the moment. We've set up actually set up the
DCGM
exporter (which was quite painful) but we're now trying to scale each service independently based on its own GPU consumption.
We're having a hard time trying to filter metrics by services right now..
a
@Jules Belveze I assume the metrics you want to export are around GPU inference speed, usage of the models etc?
j
@Aaron Pham Yes exactly
a
We can definitely take advantage of
nvidia-smi
from the looks of things. I will get back to you as I find out more about this
👍 1
j
Awesome thanks Aaron!