Slackbot
08/18/2023, 10:29 PMKaran Kumar
08/19/2023, 2:25 AMHowever, what I observed was that there were 10 times query timed out in the router than the one failed in the brokers due to this resource contention.It looks like in your usecase, you are okay with having timeouts for end users/Application. You can think of setting ``druid.server.http.enableRequestLimit`` which will cause broker to reject q's if it has reached its capacity. What we could do is not hit that broker if its overloaded. For this a new router strategy might be needed.
Didip Kerabat
08/20/2023, 5:58 PMKrishna
08/21/2023, 2:13 AMKai Sun
08/21/2023, 5:07 PMThere is a setting where you can turn off caching at broker layer. This pushed down some calculations to the historical, which will free up your brokers. For a large cluster, this is a must.This is not the case here. We already disabled the caching in the broker layer so to push the queries also to the leaf (historical nodes)
Kai Sun
08/21/2023, 5:11 PMWhat we could do is not hit that broker if its overloaded. For this a new router strategy might be needed.@Karan Kumar, can you illustrate a little bit more in detail, what strategy would be needed in this case?
Karan Kumar
08/22/2023, 5:17 AM