acoustic-doctor-87652
09/17/2025, 11:27 AMtall-airline-16756
09/17/2025, 11:27 AMtall-airline-16756
09/17/2025, 11:28 AMautoscaling configuration block. It leverages Kubernetes HPA based on CPU or memory usage. For instance, you can set a target CPU utilization of 60%, along with minimum and maximum replica counts to suit your needs.
• Ingress and Egress Services: These services also support autoscaling! For Ingress, you can use a custom Prometheus metric livekit_ingress_available to scale based on how ready the service is, rather than just CPU or memory. This requires exposing metrics via the prometheus_port.
• Graceful Scaling: When scaling down, pods enter a "draining mode," which means active rooms and sessions keep running smoothly while new connections are paused, helping to minimize any disruption.
• Traffic Distribution: For AWS EKS setups, the Helm chart works with AWS ALB (Application Load Balancer) for TLS termination and routing traffic efficiently, supporting autoscaling by distributing traffic effectively.
If you’d like, I’m here to walk you through any part of the setup or answer any questions you have about autoscaling!
Sources: ingress-sample.yaml | Self-hosting the Ingress Service | LiveKit Docs | Deploy to Kubernetes | LiveKit Docs | Distributed Setup | LiveKit Docs | Self-hosting the Egress Service | LiveKit Docsacoustic-doctor-87652
09/17/2025, 1:01 PMtall-airline-16756
09/17/2025, 1:02 PMacoustic-doctor-87652
09/17/2025, 1:03 PMtall-airline-16756
09/17/2025, 1:04 PMacoustic-doctor-87652
09/17/2025, 1:05 PMtall-airline-16756
09/17/2025, 1:05 PM