Hello All, I'm hoping yall could help me diagnose...
# general
b
Hello All, I'm hoping yall could help me diagnose an issue I'm seeing with RKE 1.16.3-rancher-1-1 running on AWS. My setup is, I have 4 worker nodes, and 1 controller/etcd node, all t3 medium instances. About once a week, I see the controller node go completely haywire and the only way I can fix it is by restarting it. The CPU and disk latency on that machine go through the roof. It's not running any of my workloads, it only has controller and etcd enabled, so I find it odd that this keeps happening. Has anyone else had this issue before? / Have any ideas on what I could do to try and diagnose it more?