Overview
For two weeks, a log aggregation pod kept getting evicted whenever the node ran short on memory. It recovered quickly enough that, from the outside, it looked like a normal restart.
What looked fine at the application level turned out to be counterintuitive. The pod wasn't being