Skip to content

2024-2025 Incident Report Summary

This document summarizes all production incidents from Slack war room channels.

Total Incidents: 9

Date Incident Type Duration Messages Resolved Notes
2025-06-12 incident 13m 7
2025-05-06 site down 3h 11m 83 Has follow-up discussions
2024-10-25 gateway timeouts in prod 7007h 2m 37 Has follow-up discussions
2024-10-08 simulation incident 20m 12
2024-08-21 prod down 29m 4
2024-06-27 production 500s 1h 36m 28 Has follow-up discussions
2024-05-23 ssl handshake failed 34m 5
2024-03-14 blank screen home 26m 14
2024-01-17 site incident edit details 46h 18m 22 Has follow-up discussions

Visual timeline showing incident duration (scaled for visibility):

05-06 site down ████████████████████████████████████████████████████████████ 3h 11m
06-12 incident ████ 13m
Date Type Duration Status
2025-05-06 site down 3h 11m ✅ Resolved
2025-06-12 incident 13m ✅ Resolved
  • Duration: 13m
  • Messages: 7
  • Resolved: Yes
  • Channel: 2025-06-12-incident
  • Summary: We’re in this meet https://meet.google.com/ccr-ksye-nxc <@U067U8GMY02> some fun for when you wake up I’ve seen two emails so far about folks not being able to get into MOmentum.

One user from the W…

  • Duration: 7007h 2m (tracking channel, not a single incident)
  • Messages: 37
  • Resolved: Unknown
  • Channel: 2024-10-25-gateway-timeouts-in-prod
  • Summary: channel to track the issue where we’re getting gateway time outs in prod The first gateway timeout discovered from the click around last night was the “the bronx” problem- (we deemed non-block…
  • Note: This appears to be a tracking channel for ongoing issues rather than a single incident. Duration reflects time between first and last message in the channel.
  • Duration: 20m
  • Messages: 12
  • Resolved: Yes
  • Channel: 2024-10-08-simulation-incident
  • Summary: Report from Jason: We’re receiving multiple customer complaints about slow API responses and some users are reporting seeing other users’ data intermittently. Is anyone else seeing this? Nothing usefu…
  • Total incidents: 9
  • Total downtime (excluding tracking channels): ~52h (excluding 2024-10-25 tracking channel)
  • Average duration (excluding tracking channels): ~6h 30m
  • Total messages: 212
  • Average messages per incident: 24
  • Resolved incidents: 7 (78%)

Note: The 2024-10-25 “gateway timeouts in prod” channel appears to be a tracking channel for ongoing issues rather than a single incident, so its duration (7007h) is excluded from downtime calculations.

Internal & Confidential: This page is only available in the internal handbook and contains confidential information.