5 ms·
Limits are what give consistency when your pod gets scheduled on nodes with different amounts of load.
by tbrownaw 1mo ago
Limits are what give consistency when your pod gets scheduled on nodes with different amounts of load.
- acuteaura 1mo agoSome workloads will also consume all the resources you hand them without being latency sensitive at all. I've handled outages of CPU time available getting suddenly compressed (we were running pod priorities with staging/prod on one cluster and up to 70% spots in 2019) and then learning that some very important applications outgrew their original requests, gone unnoticed because limits were removed a year or so prior. You can fix this with monitoring/right-sizing tools, but that requires your org to not be dysfunctional, and my style of platform engineering usually has to account for the org being very dysfunctional.