<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Northflank Status - Incident history</title>
    <link>https://status.northflank.com</link>
    <description>Northflank</description>
    <pubDate>Wed, 5 Aug 2026 13:27:13 +0000</pubDate>
    
<item>
  <title>Workloads Failing to Start Across Multiple Regions</title>
  <description>
    Type: Incident
    Duration: 1 hour and 1 minute

    Affected Components: Addons, Jobs, Services
    Aug 5, 13:27:13 GMT+0 - Investigating - We are seeing partial start up issues for workloads across multiple regions. 

We are currently investigating this incident. Aug 5, 13:42:39 GMT+0 - Monitoring - We have identified the issue and applied a mitigation. Previously stuck workloads are now starting successfully across affected regions. We are monitoring recovery and will confirm once all workloads are fully operational. Aug 5, 14:27:54 GMT+0 - Resolved - All affected workloads are now fully operational.  
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 1 minute</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:27:13&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are seeing partial start up issues for workloads across multiple regions. 

We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:42:39&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have identified the issue and applied a mitigation. Previously stuck workloads are now starting successfully across affected regions. We are monitoring recovery and will confirm once all workloads are fully operational..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:27:54&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  All affected workloads are now fully operational. .&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 5 Aug 2026 13:27:13 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmsg4fvtt034m1ao3nowjoker</link>
  <guid>https://status.northflank.com/incident/cmsg4fvtt034m1ao3nowjoker</guid>
</item>

<item>
  <title>Investigating node stability issues in the US - East region</title>
  <description>
    Type: Incident
    Duration: 1 hour and 22 minutes

    Affected Components: Addons, Jobs, Services
    Jul 22, 12:17:46 GMT+0 - Investigating - We have noticed increased failure rates of nodes in the US - East region and are investigating the issue. Jul 22, 13:11:06 GMT+0 - Monitoring - We are seeing recovery across compute and storage. We are working to get the remaining workloads back online. Jul 22, 13:39:58 GMT+0 - Resolved - We have recovered affected workloads in the region and there are no more on-going node failures. The root cause was an outage on the Google Cloud Platform in us-east4 which hosts Northflank&#039;s US - East region. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 22 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 22&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;12:17:46&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We have noticed increased failure rates of nodes in the US - East region and are investigating the issue..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 22&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:11:06&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We are seeing recovery across compute and storage. We are working to get the remaining workloads back online..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 22&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:39:58&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  We have recovered affected workloads in the region and there are no more on-going node failures. The root cause was an outage on the Google Cloud Platform in us-east4 which hosts Northflank&#039;s US - East region..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 22 Jul 2026 12:17:46 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmrw1sp1g00870smo8ztfcsfn</link>
  <guid>https://status.northflank.com/incident/cmrw1sp1g00870smo8ztfcsfn</guid>
</item>

<item>
  <title>Active DDoS Attack in US - Central</title>
  <description>
    Type: Incident
    Duration: 2 hours and 49 minutes

    Affected Components: Networking
    Jul 21, 11:09:21 GMT+0 - Identified - We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations. Jul 21, 11:20:15 GMT+0 - Monitoring - We have implemented mitigations are monitoring the situation. Jul 21, 11:41:41 GMT+0 - Resolved - The DDoS attack has stopped, we will continue to monitor the affected region. Jul 21, 13:23:09 GMT+0 - Identified - We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations. Jul 21, 13:43:46 GMT+0 - Monitoring - We have implemented mitigations are monitoring the situation. Jul 21, 13:58:28 GMT+0 - Resolved - This incident has been resolved. We will continue to monitor the affected region. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 49 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;11:09:21&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;11:20:15&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have implemented mitigations are monitoring the situation..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;11:41:41&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The DDoS attack has stopped, we will continue to monitor the affected region..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:23:09&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:43:46&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have implemented mitigations are monitoring the situation..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:58:28&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved. We will continue to monitor the affected region..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 21 Jul 2026 11:09:21 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmrujwsz307ud0kqm1nugymth</link>
  <guid>https://status.northflank.com/incident/cmrujwsz307ud0kqm1nugymth</guid>
</item>

<item>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>
  <description>
    Type: Incident
    Duration: 4 hours and 40 minutes

    Affected Components: Certificates
    Jun 18, 18:55:48 GMT+0 - Investigating - We are currently observing delays in certificate generation due to an ongoing incident with Let&#039;s Encrypt: &lt;https://letsencrypt.status.io/&gt;

This will affect provisioning of new addons with TLS enabled, the provisioning of BYOC clusters as well as domain certification provisioning. Jun 18, 23:02:45 GMT+0 - Monitoring - No recent API errors experienced and addon provisioning is working as expected. We&#039;re still closely monitoring Let&#039;s Encrypt API responses. Jun 18, 23:36:05 GMT+0 - Resolved - No more certificate creation errors have been observed from Let&#039;s Encrypt. This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 4 hours and 40 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:55:48&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently observing delays in certificate generation due to an ongoing incident with Let&#039;s Encrypt: &lt;https://letsencrypt.status.io/&gt;

This will affect provisioning of new addons with TLS enabled, the provisioning of BYOC clusters as well as domain certification provisioning..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:02:45&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  No recent API errors experienced and addon provisioning is working as expected. We&#039;re still closely monitoring Let&#039;s Encrypt API responses..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:36:05&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  No more certificate creation errors have been observed from Let&#039;s Encrypt. This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 18 Jun 2026 18:55:48 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmqjv1jt601p42okxsc6cjelx</link>
  <guid>https://status.northflank.com/incident/cmqjv1jt601p42okxsc6cjelx</guid>
</item>

<item>
  <title>Issues accessing app.northflank.com</title>
  <description>
    Type: Incident
    Duration: 17 minutes

    Affected Components: Northflank App
    Jun 10, 15:58:15 GMT+0 - Investigating - We are currently investigating this incident. Jun 10, 16:03:23 GMT+0 - Monitoring - We implemented a fix and are currently monitoring the result.

Users should now be able to log in and access the UI. Running workloads were unaffected. Jun 10, 16:14:52 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 17 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:58:15&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:03:23&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We implemented a fix and are currently monitoring the result.

Users should now be able to log in and access the UI. Running workloads were unaffected..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:14:52&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 10 Jun 2026 15:58:15 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmq896fck00lcqopy8micgt3z</link>
  <guid>https://status.northflank.com/incident/cmq896fck00lcqopy8micgt3z</guid>
</item>

<item>
  <title>Europe - West (London): Intermittent connectivity issues impacting Postgres cluster management</title>
  <description>
    Type: Incident
    Duration: 6 hours and 23 minutes

    Affected Components: Addons
    Jun 10, 08:00:23 GMT+0 - Investigating - We are currently investigating this incident.

You may see logs in Postgres with the following error message:

```
ERROR: Error communicating with DCS
``` Jun 10, 09:59:53 GMT+0 - Monitoring - We have observed a recovery starting about 20 minutes ago. Currently, Postgres addons are stable. Jun 10, 10:19:44 GMT+0 - Identified - We saw a period of recovery between 09:45 and 10:04 UTC, but the issue has since recurred.

This is tied to an ongoing incident affecting our cloud service provider in the region.

We are continuing to monitor and will share further updates as we have them Jun 10, 11:53:12 GMT+0 - Monitoring - We have not observed any further issues since 11:05 UTC. 

We are awaiting confirmation from our cloud provider that the issue is fully resolved on their side, and we will update this status accordingly. Jun 10, 14:23:08 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 6 hours and 23 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;08:00:23&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident.

You may see logs in Postgres with the following error message:

```
ERROR: Error communicating with DCS
```.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;09:59:53&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have observed a recovery starting about 20 minutes ago. Currently, Postgres addons are stable..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;10:19:44&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We saw a period of recovery between 09:45 and 10:04 UTC, but the issue has since recurred.

This is tied to an ongoing incident affecting our cloud service provider in the region.

We are continuing to monitor and will share further updates as we have them.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;11:53:12&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have not observed any further issues since 11:05 UTC. 

We are awaiting confirmation from our cloud provider that the issue is fully resolved on their side, and we will update this status accordingly..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:23:08&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 10 Jun 2026 08:00:23 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmq7v8suo0762p4uz2pxqysg8</link>
  <guid>https://status.northflank.com/incident/cmq7v8suo0762p4uz2pxqysg8</guid>
</item>

<item>
  <title>Europe - West (London): New resources not able to be created</title>
  <description>
    Type: Incident
    Duration: 1 hour and 26 minutes

    Affected Components: Addons, Jobs, Services
    May 13, 15:00:45 GMT+0 - Investigating - We are currently investigating this incident.

Existing running workloads are unaffected. May 13, 15:33:29 GMT+0 - Monitoring - We are seeing an improvement; workloads are starting to be created.

Connectivity has been restored, and we are continuing to monitor the situation. May 13, 16:27:13 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 26 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:00:45&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident.

Existing running workloads are unaffected..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:33:29&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We are seeing an improvement; workloads are starting to be created.

Connectivity has been restored, and we are continuing to monitor the situation..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:27:13&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 13 May 2026 15:00:45 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmp46smud00byp6qva4zwgzm6</link>
  <guid>https://status.northflank.com/incident/cmp46smud00byp6qva4zwgzm6</guid>
</item>

<item>
  <title>Active DDoS Attack in Europe - West - Netherlands</title>
  <description>
    Type: Incident
    Duration: 2 hours and 2 minutes

    Affected Components: Addons, Services
    May 10, 18:30:00 GMT+0 - Identified - We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations. May 10, 20:10:07 GMT+0 - Monitoring - We have put in place mitigations and are monitoring the affected region. May 10, 20:31:56 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 2 minutes</p>
    <p><strong>Affected Components:</strong> , </p>
    &lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:30:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;20:10:07&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have put in place mitigations and are monitoring the affected region..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;20:31:56&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Sun, 10 May 2026 18:30:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmp05l0ln00fo96bh1a4pufch</link>
  <guid>https://status.northflank.com/incident/cmp05l0ln00fo96bh1a4pufch</guid>
</item>

<item>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>
  <description>
    Type: Incident
    Duration: 2 hours and 36 minutes

    Affected Components: Certificates
    May 8, 19:00:00 GMT+0 - Identified - [Let&#039;s Encrypt is currently experiencing a partial service disruption](https://letsencrypt.status.io/) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons, domains and BYOC clusters. May 8, 21:36:09 GMT+0 - Resolved - Certificate generation is functioning normally again. We will continue to monitor the situation. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 36 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  [Let&#039;s Encrypt is currently experiencing a partial service disruption](https://letsencrypt.status.io/) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons, domains and BYOC clusters..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;21:36:09&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  Certificate generation is functioning normally again. We will continue to monitor the situation..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Fri, 8 May 2026 19:00:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmoxb63od01es6n0u18am5oyu</link>
  <guid>https://status.northflank.com/incident/cmoxb63od01es6n0u18am5oyu</guid>
</item>

<item>
  <title>Addon and volume availability issues in US - Central</title>
  <description>
    Type: Incident
    Duration: 52 minutes

    Affected Components: Addons, Services
    Apr 8, 06:15:00 GMT+0 - Identified - We are aware of an issue with a node affecting some stateful workloads in US - Central.

We are working on a fix for this incident. Apr 8, 07:07:27 GMT+0 - Resolved - We have recovered the node, and all workloads are now running again. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 52 minutes</p>
    <p><strong>Affected Components:</strong> , </p>
    &lt;p&gt;&lt;small&gt;Apr &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;06:15:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are aware of an issue with a node affecting some stateful workloads in US - Central.

We are working on a fix for this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Apr &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;07:07:27&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  We have recovered the node, and all workloads are now running again..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 8 Apr 2026 06:15:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmnpp03h1006jd0s28d2teq25</link>
  <guid>https://status.northflank.com/incident/cmnpp03h1006jd0s28d2teq25</guid>
</item>

<item>
  <title>Active DDoS Attack in Europe - West - Netherlands</title>
  <description>
    Type: Incident
    Duration: 3 hours and 21 minutes

    Affected Components: Addons, Jobs, Services
    Mar 25, 14:28:04 GMT+0 - Identified - We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations. Mar 25, 16:01:00 GMT+0 - Monitoring - The region is currently stable. We are still actively mitigating the ongoing attack. Mar 25, 17:48:48 GMT+0 - Resolved - The DDoS attack has stopped we will continue to monitor the affected region. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 3 hours and 21 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:28:04&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:01:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  The region is currently stable. We are still actively mitigating the ongoing attack..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:48:48&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The DDoS attack has stopped we will continue to monitor the affected region..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 25 Mar 2026 14:28:04 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmn651uh000a880hcl3ox9cxf</link>
  <guid>https://status.northflank.com/incident/cmn651uh000a880hcl3ox9cxf</guid>
</item>

<item>
  <title>App UI is experiencing loading issues</title>
  <description>
    Type: Incident
    Duration: 34 minutes

    Affected Components: Northflank App
    Mar 16, 18:16:28 GMT+0 - Resolved - This incident has been resolved. Mar 16, 18:15:58 GMT+0 - Monitoring - We identified the root cause as a system component causing disproportionate database load. This is affecting read performance leading to degraded performance for the app UI. We have implemented a fix and are monitoring the situation. Mar 16, 17:42:54 GMT+0 - Identified - We are investing an incident where the UI is failing to load correctly. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 34 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 16&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:16:28&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 16&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:15:58&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We identified the root cause as a system component causing disproportionate database load. This is affecting read performance leading to degraded performance for the app UI. We have implemented a fix and are monitoring the situation..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 16&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:42:54&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are investing an incident where the UI is failing to load correctly..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 16 Mar 2026 17:42:54 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmmth1qit0cwahxt9crdwuuo3</link>
  <guid>https://status.northflank.com/incident/cmmth1qit0cwahxt9crdwuuo3</guid>
</item>

<item>
  <title>Degraded performance in US Central region</title>
  <description>
    Type: Incident
    Duration: 29 minutes

    Affected Components: Networking, Addons, Services
    Mar 9, 12:06:32 GMT+0 - Investigating - We are currently investigating this incident. Mar 9, 12:35:56 GMT+0 - Resolved - The root cause has been identified as an issue with the GCP infrastructure control plane and has been resolved.  
  
There was minimal impact to running user workloads. The main impact was a delay in provisioning net new or redeploying existing workloads. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 29 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 9&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;12:06:32&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 9&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;12:35:56&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The root cause has been identified as an issue with the GCP infrastructure control plane and has been resolved.  
  
There was minimal impact to running user workloads. The main impact was a delay in provisioning net new or redeploying existing workloads..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 9 Mar 2026 12:06:32 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmmj4y7ct01zw1jix7jv6r4vd</link>
  <guid>https://status.northflank.com/incident/cmmj4y7ct01zw1jix7jv6r4vd</guid>
</item>

<item>
  <title>Issues with builds using the Heroku 24 Buildpack</title>
  <description>
    Type: Incident
    Duration: 1 hour and 34 minutes

    Affected Components: Builds
    Feb 19, 15:00:00 GMT+0 - Identified - We are aware of an issue where builds fail to start when using the Heroku 24 buildpack image.

We are currently working on identifying a solution. We will provide an update when the fix has been released. Feb 19, 16:34:09 GMT+0 - Resolved - We have released a fix for the issue. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 34 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Feb &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are aware of an issue where builds fail to start when using the Heroku 24 buildpack image.

We are currently working on identifying a solution. We will provide an update when the fix has been released..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Feb &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:34:09&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  We have released a fix for the issue..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 19 Feb 2026 15:00:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmltl7z5d002trjwmeqyabys9</link>
  <guid>https://status.northflank.com/incident/cmltl7z5d002trjwmeqyabys9</guid>
</item>

<item>
  <title>Node infrastructure stability - London Region</title>
  <description>
    Type: Incident
    Duration: 39 minutes

    Affected Components: Networking, , Addons, Jobs, Services, 
Northflank Platform →
    Dec 11, 13:54:32 GMT+0 - Identified - There are currently issues with node stability in the London region leading to partial outages for some workloads. The team has identified the issue and is implementing a mitigation. Dec 11, 14:13:28 GMT+0 - Monitoring - The mitigation is in place and the region is operating as expected. Dec 11, 14:33:03 GMT+0 - Resolved - The incident has been resolved and the cause identified. A node image security release lead to the host filesystem going into a read only mode when interacting with a specific set of workloads causing the node to become unresponsive. We have implemented a mitigation and are working on a permanent solution. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 39 minutes</p>
    <p><strong>Affected Components:</strong> , , , , </p>
    &lt;p&gt;&lt;small&gt;Dec &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;13:54:32&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  There are currently issues with node stability in the London region leading to partial outages for some workloads. The team has identified the issue and is implementing a mitigation..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Dec &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:13:28&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  The mitigation is in place and the region is operating as expected..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Dec &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:33:03&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The incident has been resolved and the cause identified. A node image security release lead to the host filesystem going into a read only mode when interacting with a specific set of workloads causing the node to become unresponsive. We have implemented a mitigation and are working on a permanent solution..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 11 Dec 2025 13:54:32 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmj1i24qr0106uhdob8a4bf2w</link>
  <guid>https://status.northflank.com/incident/cmj1i24qr0106uhdob8a4bf2w</guid>
</item>

<item>
  <title>Docker Hub is experiencing issues</title>
  <description>
    Type: Incident
    Duration: 1 hour and 44 minutes

    Affected Components: Addons, Jobs, Services
    Sep 24, 23:53:50 GMT+0 - Identified - Dockerhub is currently experiencing issues related to authentication. This may cause issues with starting jobs, addons and services. &lt;https://www.dockerstatus.com/pages/incident/533c6539221ae15e3f000031/68d47a2f93c09e05486d93a9&gt; Sep 25, 01:11:08 GMT+0 - Monitoring - We are observing that image pull requests for Docker Hub images are now succeeding. Sep 25, 01:38:09 GMT+0 - Resolved - Dockerhub is fully operational and image pulling is now working as expected. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 44 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Sep &lt;var data-var=&#039;date&#039;&gt; 24&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:53:50&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Dockerhub is currently experiencing issues related to authentication. This may cause issues with starting jobs, addons and services. &lt;https://www.dockerstatus.com/pages/incident/533c6539221ae15e3f000031/68d47a2f93c09e05486d93a9&gt;.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Sep &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;01:11:08&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We are observing that image pull requests for Docker Hub images are now succeeding..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Sep &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;01:38:09&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  Dockerhub is fully operational and image pulling is now working as expected..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 24 Sep 2025 23:53:50 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmfyn4drv02axox9ea8otzvu7</link>
  <guid>https://status.northflank.com/incident/cmfyn4drv02axox9ea8otzvu7</guid>
</item>

<item>
  <title>Delay in DNS propagation</title>
  <description>
    Type: Incident
    Duration: 18 hours and 29 minutes

    Affected Components: Addons, Services, Certificates
    Sep 16, 15:09:51 GMT+0 - Monitoring - Our DNS provider NS1 is experiencing delays in DNS propagation.

This is affecting service and addon creation.  
  
NS1 status page: &lt;https://status.ai-apps-comms.ibm.com/ns1connect&gt; Sep 17, 09:39:01 GMT+0 - Resolved - The incident has been resolved on NS1 and we have not been seeing any more related failures over the last hour. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 18 hours and 29 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Sep &lt;var data-var=&#039;date&#039;&gt; 16&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:09:51&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Our DNS provider NS1 is experiencing delays in DNS propagation.

This is affecting service and addon creation.  
  
NS1 status page: &lt;https://status.ai-apps-comms.ibm.com/ns1connect&gt;.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Sep &lt;var data-var=&#039;date&#039;&gt; 17&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;09:39:01&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The incident has been resolved on NS1 and we have not been seeing any more related failures over the last hour..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 16 Sep 2025 15:09:51 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmfmovpks003r739gifi0hnsh</link>
  <guid>https://status.northflank.com/incident/cmfmovpks003r739gifi0hnsh</guid>
</item>

<item>
  <title>Issue affecting local disk caching for builds running on Northflank&#039;s infrastructure</title>
  <description>
    Type: Incident
    Duration: 1 hour and 17 minutes

    Affected Components: Builds
    Aug 25, 21:48:19 GMT+0 - Identified - There is an issue with the underlying storage which provides the local disk-based caching feature for builds on Northflank&#039;s infrastructure. BYOC builds and builds without a local cache will be unaffected.

  
We are currently working on a fix for this. Aug 25, 22:45:24 GMT+0 - Monitoring - We implemented a fix and are currently monitoring the result. Aug 25, 23:04:54 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 17 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;21:48:19&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  There is an issue with the underlying storage which provides the local disk-based caching feature for builds on Northflank&#039;s infrastructure. BYOC builds and builds without a local cache will be unaffected.

  
We are currently working on a fix for this..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:45:24&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We implemented a fix and are currently monitoring the result..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 25&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:04:54&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 25 Aug 2025 21:48:19 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmernfevf002amj1q4pgwr3xn</link>
  <guid>https://status.northflank.com/incident/cmernfevf002amj1q4pgwr3xn</guid>
</item>

<item>
  <title>Log &amp; Metric Ingestion and Query Outage</title>
  <description>
    Type: Incident
    Duration: 1 hour

    Affected Components: Logs and Metrics
    Aug 5, 07:23:59 GMT+0 - Investigating - We are currently investigating this incident. Aug 5, 07:55:18 GMT+0 - Monitoring - We implemented a fix and are currently monitoring the result. Aug 5, 08:23:56 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;07:23:59&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;07:55:18&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We implemented a fix and are currently monitoring the result..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Aug &lt;var data-var=&#039;date&#039;&gt; 5&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;08:23:56&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 5 Aug 2025 07:23:59 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmdy7qu4001ym91lf2txeit0h</link>
  <guid>https://status.northflank.com/incident/cmdy7qu4001ym91lf2txeit0h</guid>
</item>

<item>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>
  <description>
    Type: Incident
    Duration: 11 hours and 31 minutes

    Affected Components: Certificates
    Jul 21, 19:13:49 GMT+0 - Monitoring - [Let&#039;s Encrypt is currently experiencing an outage](https://letsencrypt.status.io/pages/incident/55957a99e800baa4470002da/687e8d62b8a4e804fad85799) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons or domains. Jul 22, 06:44:30 GMT+0 - Resolved - Certificate generation is functioning without issue. We will continue to monitor the situation. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 11 hours and 31 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:13:49&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  [Let&#039;s Encrypt is currently experiencing an outage](https://letsencrypt.status.io/pages/incident/55957a99e800baa4470002da/687e8d62b8a4e804fad85799) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons or domains..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 22&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;06:44:30&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  Certificate generation is functioning without issue. We will continue to monitor the situation..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 21 Jul 2025 19:13:49 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmddhhx1500cbb87xjzw4t6ee</link>
  <guid>https://status.northflank.com/incident/cmddhhx1500cbb87xjzw4t6ee</guid>
</item>

<item>
  <title>Issues with Google Cloud platform - Affecting multiple services</title>
  <description>
    Type: Incident
    Duration: 2 hours and 45 minutes

    Affected Components: Builds, Addons, Logs and Metrics
    Jun 12, 20:40:23 GMT+0 - Monitoring - We are observing improved operations in the US - Central region. Addon creation and backup issues are now resolved. Ongoing Issues: \* GCP BYOC cluster processing Jun 12, 18:00:00 GMT+0 - Identified - We have noticed issues with Google Cloud APIs. This seems to be widespread across the Google Cloud platform. We are currently determining the impact of the outage, and we will provide updates when we know more. Currently, logs, metrics, builds, BYOC and addon backups may be affected. Jun 12, 19:21:45 GMT+0 - Monitoring - Google Cloud has acknowledged the incident. For detailed information, please visit: &lt;https://status.cloud.google.com/incidents/ow5i3PPK96RduMcb1SsW&gt;

Most services have been restored to normal operation. However, we are still experiencing the following issues: 

* Addon creation in the US - Central region
* Addon backups in the US - Central region
* GCP BYOC cluster processing

We will continue monitoring the situation and provide updates as they become available. Jun 12, 20:44:39 GMT+0 - Resolved - GCP BYOC cluster processing is now operational. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 45 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;20:40:23&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We are observing improved operations in the US - Central region. Addon creation and backup issues are now resolved. Ongoing Issues: \* GCP BYOC cluster processing.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We have noticed issues with Google Cloud APIs. This seems to be widespread across the Google Cloud platform. We are currently determining the impact of the outage, and we will provide updates when we know more. Currently, logs, metrics, builds, BYOC and addon backups may be affected..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:21:45&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Google Cloud has acknowledged the incident. For detailed information, please visit: &lt;https://status.cloud.google.com/incidents/ow5i3PPK96RduMcb1SsW&gt;

Most services have been restored to normal operation. However, we are still experiencing the following issues: 

* Addon creation in the US - Central region
* Addon backups in the US - Central region
* GCP BYOC cluster processing

We will continue monitoring the situation and provide updates as they become available..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;20:44:39&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  GCP BYOC cluster processing is now operational..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 12 Jun 2025 18:00:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmbtqad4j000l6didp4d0xkfy</link>
  <guid>https://status.northflank.com/incident/cmbtqad4j000l6didp4d0xkfy</guid>
</item>

<item>
  <title>Metrics ingestion and query outage</title>
  <description>
    Type: Incident
    Duration: 2 hours and 17 minutes

    Affected Components: Logs and Metrics
    Jun 11, 22:00:00 GMT+0 - Identified - Metrics query and ingestion are currently affected. Logs are unaffected.

We are currently applying a fix for the issue and monitoring the results. Jun 11, 23:16:34 GMT+0 - Monitoring - The fix has been rolled out and we are currently monitoring the ingestion rate and catch-up. Jun 12, 00:16:31 GMT+0 - Resolved - The ingestion rate has been stable for 30 minutes and historic writes have caught up.

We sincerely apologise for the service disruption you may have experienced. We are investigating the root cause to prevent similar issues in the future. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 17 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Metrics query and ingestion are currently affected. Logs are unaffected.

We are currently applying a fix for the issue and monitoring the results..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:16:34&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  The fix has been rolled out and we are currently monitoring the ingestion rate and catch-up..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jun &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;00:16:31&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  The ingestion rate has been stable for 30 minutes and historic writes have caught up.

We sincerely apologise for the service disruption you may have experienced. We are investigating the root cause to prevent similar issues in the future..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 11 Jun 2025 22:00:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmbsjf79f0025k18t40kcbyb0</link>
  <guid>https://status.northflank.com/incident/cmbsjf79f0025k18t40kcbyb0</guid>
</item>

<item>
  <title>Infrastructure stability issues</title>
  <description>
    Type: Incident
    Duration: 1 hour and 10 minutes

    Affected Components: Networking, Addons, Services
    May 21, 16:04:00 GMT+0 - Monitoring -  May 21, 17:13:33 GMT+0 - Resolved - **We sincerely apologise for the outage you experienced.**  

**What happened:**  
Yesterday, we rolled out an update to support ARM architecture across our platform.  
As part of that, a migration was initiated today to extend support to x86 architecture within our database schemas.  
This triggered a higher-than-expected volume of database rescheduling events, which in turn put excessive pressure on the underlying GCP infrastructure. Some nodes failed under this load, compounding the issue and causing degraded service for several customers.

**Timeline of events (UTC):**

* **15:36** — Northflank engineers were alerted to failing health checks on customer endpoints.
* **15:41** — Elevated 503 errors were detected from a specific cluster.
* **15:42** — An internal incident was formally declared.
* **15:57** — Replacement infrastructure began provisioning to recover failed nodes.
* **16:32** — All customer addons returned to a healthy state, with the exception of four databases which required manual intervention.

  
Thank you for your patience—we know how critical uptime is, and we are reviewing and improving our systems to ensure this doesn’t happen again.  
  
Will Stewart  
CEO @ Northflank 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 10 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:04:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  .&lt;/p&gt;
&lt;p&gt;&lt;small&gt;May &lt;var data-var=&#039;date&#039;&gt; 21&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:13:33&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  **We sincerely apologise for the outage you experienced.**  

**What happened:**  
Yesterday, we rolled out an update to support ARM architecture across our platform.  
As part of that, a migration was initiated today to extend support to x86 architecture within our database schemas.  
This triggered a higher-than-expected volume of database rescheduling events, which in turn put excessive pressure on the underlying GCP infrastructure. Some nodes failed under this load, compounding the issue and causing degraded service for several customers.

**Timeline of events (UTC):**

* **15:36** — Northflank engineers were alerted to failing health checks on customer endpoints.
* **15:41** — Elevated 503 errors were detected from a specific cluster.
* **15:42** — An internal incident was formally declared.
* **15:57** — Replacement infrastructure began provisioning to recover failed nodes.
* **16:32** — All customer addons returned to a healthy state, with the exception of four databases which required manual intervention.

  
Thank you for your patience—we know how critical uptime is, and we are reviewing and improving our systems to ensure this doesn’t happen again.  
  
Will Stewart  
CEO @ Northflank.&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 21 May 2025 16:04:00 +0000</pubDate>
  <link>https://status.northflank.com/incident/cmay5h4pg0052zbome4c6m5q6</link>
  <guid>https://status.northflank.com/incident/cmay5h4pg0052zbome4c6m5q6</guid>
</item>

<item>
  <title>US - Central - Inability to Create Disk Snapshots</title>
  <description>
    Type: Incident
    Duration: 1 hour and 17 minutes

    Affected Components: Addons
    Nov 8, 19:20:22 GMT+0 - Resolved - This incident has been resolved. Nov 8, 18:02:57 GMT+0 - Monitoring - Google Cloud is experiencing issues with snapshot upload in the US Central region.

This is causing delays for addon snapshot backups within the region 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 17 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:20:22&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:02:57&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Google Cloud is experiencing issues with snapshot upload in the US Central region.

This is causing delays for addon snapshot backups within the region.&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Fri, 8 Nov 2024 18:02:57 +0000</pubDate>
  <link>https://status.northflank.com/incident/cm391ojva000bzr1411i47iz0</link>
  <guid>https://status.northflank.com/incident/cm391ojva000bzr1411i47iz0</guid>
</item>

<item>
  <title>DNS Propagation Delay</title>
  <description>
    Type: Incident
    Duration: 1 hour and 27 minutes

    Affected Components: Services, Addons, Certificates
    Oct 11, 18:41:37 GMT+0 - Monitoring - Our DNS provider is experiencing issues with DNS propagation leading to delays in provisioning new workloads and certificates. Existing workloads are unaffected.

&lt;https://ns1connect.trust.pagerduty.com/posts/details/PK35B1W&gt; Oct 11, 20:09:04 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 27 minutes</p>
    <p><strong>Affected Components:</strong> , , </p>
    &lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:41:37&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Our DNS provider is experiencing issues with DNS propagation leading to delays in provisioning new workloads and certificates. Existing workloads are unaffected.

&lt;https://ns1connect.trust.pagerduty.com/posts/details/PK35B1W&gt;.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 11&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;20:09:04&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Fri, 11 Oct 2024 18:41:37 +0000</pubDate>
  <link>https://status.northflank.com/incident/cm252qf810036lb09qejd495o</link>
  <guid>https://status.northflank.com/incident/cm252qf810036lb09qejd495o</guid>
</item>

  </channel>
  </rss>