Resolved
Sep 21, 2026 at 2:05pm UTC
Prod was down for ~15 minutes during the afternoon, due to a failing node in the DB cluster which took longer than expected to rebalance to the replica node. The failing node was rebooted which quickly restored functionality. No data loss has been identified.