Start here. This is the direct spoken answer to practice first.
Why this question matters
Production incidents are team events. The answer should show calm coordination, not a room full of people trying random fixes.
During one incident, API latency spiked after a release and multiple people jumped in at once. I helped slow the room down by separating roles: one person checked release changes, one checked database metrics, one watched logs and traces, and one owned stakeholder updates. I focused on correlating slow endpoints with query changes and dependency latency so we could decide whether rollback or a targeted fix made more sense.