Resolved -
Between approximately 10:30 and 11:55 UTC on July 28, customers with telephony workloads in our London region experienced elevated error rates on SIP APIs. Small number of call transfers were affected.
We mitigated by removing the affected infrastructure from service at 11:48 UTC, and error rates returned to normal by 11:55 UTC.
Jul 28, 07:23 PDT
Resolved -
On July 27, our San Jose region experienced a minor networking incident beginning at 16:14 UTC. Some users connected to this region may have experienced disconnects and increased session start times. Traffic was routed away from this region by 16:39 UTC and errors have since returned to baseline.
Jul 27, 10:00 PDT
Resolved -
On July 23, between 07:50 and 08:07 UTC, media servers in our London region experienced a host-level network fault, causing them to fail health checks and restart. Users connected to these servers may have experienced brief connection interruptions, and egress operations may have seen increased latency. The affected servers were removed and replaced. The region has been operating normally since.
Jul 23, 01:00 PDT
Resolved -
Error rates for room service api requests in our London region have remained normal since 09:06 am UTC and we are no longer seeing any issues.
Jul 14, 03:41 PDT
Monitoring -
Between 08:54 am and 09:06 am UTC, 5.4% of room service api requests in our London region failed with errors. A fix was applied and error rates returned to normal by 09:06 am UTC. We are continuing to monitor.
Jul 14, 02:49 PDT