Global API Gateway Latency & SSL Handshake Failures
Detected
24 May · 14:20 UTC
Expected duration
2h 25m
Impact radius
Global (partial)
Stability
85%
Incident summary & impact
Impact analysis
Users in EU-West and US-East are experiencing intermittent failures when connecting to the core API. Response times increased by 35%, with approximately 5% of requests failing during SSL negotiation.
Affected components
External links
Grafana ↗
Elasticsearch ↗
Mitigation actions taken
- Traffic rerouted from impacted edge nodes to secondary standby clusters.
- Timeout thresholds increased for SSL handshakes at the load balancer.
- Regional CDN cache flushed to clear potentially corrupted session state.
Timeline of events
Last update: 19:45 UTCUpdate · 19:45 UTC
Infrastructure TeamMonitoring has begun. Request success rates stabilized over the last 30 minutes and error rates dropped below 0.01%.
Identification · 18:30 UTC
Security EngineeringA certificate rotation script failed to propagate the new intermediate CA bundle to regional edge nodes. The configuration rollback is complete.
Investigation · 17:55 UTC
DevOps LeadInvestigation centered on SSL negotiation after 504 Gateway Timeouts appeared for traffic hitting the primary gateway.
Update · 17:30 UTC
Incident ManagerConnectivity issues were reported for API services. Engineering began reviewing regional traffic logs.
