Production Outage - 17 September 2026

Timing report - certificate cleanup and service account credentials

Summary

A maintenance job with a new cleanup step removed old and stale certificates. A subsequent deployment to a staging environment forced a certificate reload, which caused the production server to become unavailable.

Restoring the certificates was not sufficient on its own. The service recovered fully once the service account credentials were re-entered and Windows picked up the restored certificates.

Measurable customer impact ran from 13:16 to 15:00 CEST (11:16 to 13:00 UTC) – roughly 1 hour 45 minutes, including two separate periods of disruption.

The incident has also highlighted unnecessary complexity in parts of our current server setup. This complexity was the direct cause of the outage, but simplifying the setup will make the environment easier to maintain and reduce the risk of similar issues in the future.

Follow-up actions

Following the incident, we are taking a number of corrective and preventive actions:

This work has already started and will continue over the coming days and into next week.

Current status

The production environment is running normally, and we do not currently expect the ongoing work to cause further disruption.

Some staging environments are not yet working 100%, and we are continuing to work on these.

If any other issues are currently being experienced, please report them to SpeedAdmin Support so they can be investigated.

Merged timeline

Points worth stating

Book online!