This message was deleted.
# ask-for-help
s
This message was deleted.
v
Or is a single model shared across multiple runner workers
Trying to see how much memory to allocate in our deployments, thanks in advance!
c
One model instance per runner replica! Sharing one instance across multiple python processes is possible but is very hacky and can lead to many production issues 😊
gratitude thank you 1
Btw this is just the default behavior, BentoML is flexible enough that users can actually customize runners’ scheduling strategy as well
👍 1