An incident at AWS resulted in the majority of the application servers to be taken out of action. While the remaining application servers took up the load and were able to response to most requests, with elevated response times, a configuration issue meant it took an 45 minutes before the replacement application servers were available.
We will be deploying infrastructure fixes this weekend which will avoid this scenario from repeating again in the future.