Slackbot
09/07/2023, 8:08 PMJudah Rand
09/07/2023, 8:09 PMJudah Rand
09/07/2023, 8:10 PMJudah Rand
09/07/2023, 8:13 PMBENTOML_CONFIG_OPTIONS='runners.resources.cpu=1'Chaoyu
09/07/2023, 9:15 PMJudah Rand
09/08/2023, 8:56 AMYes that’s right, are you using a built-in runner? Most of the built in runner should be default to one workerNope! Using a custom runner which does a feature store interaction as well as an inference call
Judah Rand
09/08/2023, 8:56 AMChaoyu
09/08/2023, 1:29 PMChaoyu
09/08/2023, 1:31 PMChaoyu
09/08/2023, 1:32 PMChaoyu
09/08/2023, 1:33 PMJudah Rand
09/08/2023, 3:34 PMTypically we recommend out feature store fetching code in API server and leverage async. And in some cases, feature fetching code in its own runner to leverage batchingI'm not sure that this is generally a good recommendation. The Runner is batched and the API worker is not.
Judah Rand
09/08/2023, 3:35 PMJudah Rand
09/08/2023, 3:36 PMJudah Rand
09/08/2023, 3:36 PMJudah Rand
09/08/2023, 3:37 PMDid you set the support multi threading flag in runnable class?
SUPPORTS_CPU_MULTI_THREADING = FalseJian Shen Yap
09/09/2023, 2:41 AM