Major Incident Management experience — has actually run or driven a bridge/war-room for a P1/P2 outage, not just "participated." This is the core of the role; everything else is supporting it. Ask for a specific incident they led end-to-end.
RCA/Problem Management rigor — can articulate a structured RCA methodology.
Production troubleshooting under pressure across a mixed stack — comfortable reading logs/dashboards (Dynatrace-type tooling) and reasoning about cloud + on-prem + store/POS systems simultaneously, since a real incident here will span all three.
This listing is sourced from a third-party job board. Applying will redirect you to the original posting.