<?xml version="1.0" encoding="UTF-8"?>
<feed xml:lang="en-US" xmlns="http://www.w3.org/2005/Atom">
  <id>tag:status.northflank.com,2005:/history</id>
  <link rel="alternate" type="text/html" href="https://status.northflank.com"/>
  <link rel="self" type="application/atom+xml" href="https://status.northflank.com/history.atom"/>
  <title>Northflank Status - Incident history</title>
  <updated>2026-08-05T13:27:13.111+00:00</updated>
  <author>
    <name>Northflank</name>
  </author>
  
<entry>
  <id>tag:status.northflank.com,2005:Incident/cmsg4fvtt034m1ao3nowjoker</id>
  <published>2026-08-05T13:27:13.111+00:00</published>
  <updated>2026-08-05T13:27:13.111+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmsg4fvtt034m1ao3nowjoker"/>
  <title>Workloads Failing to Start Across Multiple Regions</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 1 minute</p>
    <p><strong>Affected Components:</strong> Addons, Jobs, Services</p>
    <p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>13:27:13</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are seeing partial start up issues for workloads across multiple regions. 

We are currently investigating this incident..</p>
<p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>13:42:39</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have identified the issue and applied a mitigation. Previously stuck workloads are now starting successfully across affected regions. We are monitoring recovery and will confirm once all workloads are fully operational..</p>
<p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>14:27:54</var> GMT+0</small><br /><strong>Resolved</strong> -
  All affected workloads are now fully operational. .</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmrw1sp1g00870smo8ztfcsfn</id>
  <published>2026-07-22T12:17:46.396+00:00</published>
  <updated>2026-07-22T12:17:46.396+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmrw1sp1g00870smo8ztfcsfn"/>
  <title>Investigating node stability issues in the US - East region</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 22 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Jobs, Services</p>
    <p><small>Jul <var data-var='date'> 22</var>, <var data-var='time'>12:17:46</var> GMT+0</small><br /><strong>Investigating</strong> -
  We have noticed increased failure rates of nodes in the US - East region and are investigating the issue..</p>
<p><small>Jul <var data-var='date'> 22</var>, <var data-var='time'>13:11:06</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We are seeing recovery across compute and storage. We are working to get the remaining workloads back online..</p>
<p><small>Jul <var data-var='date'> 22</var>, <var data-var='time'>13:39:58</var> GMT+0</small><br /><strong>Resolved</strong> -
  We have recovered affected workloads in the region and there are no more on-going node failures. The root cause was an outage on the Google Cloud Platform in us-east4 which hosts Northflank&#039;s US - East region..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmrujwsz307ud0kqm1nugymth</id>
  <published>2026-07-21T11:09:21.442+00:00</published>
  <updated>2026-07-21T11:09:21.442+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmrujwsz307ud0kqm1nugymth"/>
  <title>Active DDoS Attack in US - Central</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 49 minutes</p>
    <p><strong>Affected Components:</strong> Networking</p>
    <p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>11:09:21</var> GMT+0</small><br /><strong>Identified</strong> -
  We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..</p>
<p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>11:20:15</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have implemented mitigations are monitoring the situation..</p>
<p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>11:41:41</var> GMT+0</small><br /><strong>Resolved</strong> -
  The DDoS attack has stopped, we will continue to monitor the affected region..</p>
<p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>13:23:09</var> GMT+0</small><br /><strong>Identified</strong> -
  We are currently experiencing a DDoS attack in the US - Central region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..</p>
<p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>13:43:46</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have implemented mitigations are monitoring the situation..</p>
<p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>13:58:28</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved. We will continue to monitor the affected region..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmqjv1jt601p42okxsc6cjelx</id>
  <published>2026-06-18T18:55:48.117+00:00</published>
  <updated>2026-06-18T18:55:48.117+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmqjv1jt601p42okxsc6cjelx"/>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 4 hours and 40 minutes</p>
    <p><strong>Affected Components:</strong> Certificates</p>
    <p><small>Jun <var data-var='date'> 18</var>, <var data-var='time'>18:55:48</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently observing delays in certificate generation due to an ongoing incident with Let&#039;s Encrypt: &lt;https://letsencrypt.status.io/&gt;

This will affect provisioning of new addons with TLS enabled, the provisioning of BYOC clusters as well as domain certification provisioning..</p>
<p><small>Jun <var data-var='date'> 18</var>, <var data-var='time'>23:02:45</var> GMT+0</small><br /><strong>Monitoring</strong> -
  No recent API errors experienced and addon provisioning is working as expected. We&#039;re still closely monitoring Let&#039;s Encrypt API responses..</p>
<p><small>Jun <var data-var='date'> 18</var>, <var data-var='time'>23:36:05</var> GMT+0</small><br /><strong>Resolved</strong> -
  No more certificate creation errors have been observed from Let&#039;s Encrypt. This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmq896fck00lcqopy8micgt3z</id>
  <published>2026-06-10T15:58:15.402+00:00</published>
  <updated>2026-06-10T15:58:15.402+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmq896fck00lcqopy8micgt3z"/>
  <title>Issues accessing app.northflank.com</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 17 minutes</p>
    <p><strong>Affected Components:</strong> Northflank App</p>
    <p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>15:58:15</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident..</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>16:03:23</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We implemented a fix and are currently monitoring the result.

Users should now be able to log in and access the UI. Running workloads were unaffected..</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>16:14:52</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmq7v8suo0762p4uz2pxqysg8</id>
  <published>2026-06-10T08:00:23.675+00:00</published>
  <updated>2026-06-10T08:00:23.675+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmq7v8suo0762p4uz2pxqysg8"/>
  <title>Europe - West (London): Intermittent connectivity issues impacting Postgres cluster management</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 6 hours and 23 minutes</p>
    <p><strong>Affected Components:</strong> Addons</p>
    <p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>08:00:23</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident.

You may see logs in Postgres with the following error message:

```
ERROR: Error communicating with DCS
```.</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>09:59:53</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have observed a recovery starting about 20 minutes ago. Currently, Postgres addons are stable..</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>10:19:44</var> GMT+0</small><br /><strong>Identified</strong> -
  We saw a period of recovery between 09:45 and 10:04 UTC, but the issue has since recurred.

This is tied to an ongoing incident affecting our cloud service provider in the region.

We are continuing to monitor and will share further updates as we have them.</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>11:53:12</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have not observed any further issues since 11:05 UTC. 

We are awaiting confirmation from our cloud provider that the issue is fully resolved on their side, and we will update this status accordingly..</p>
<p><small>Jun <var data-var='date'> 10</var>, <var data-var='time'>14:23:08</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmp46smud00byp6qva4zwgzm6</id>
  <published>2026-05-13T15:00:45.533+00:00</published>
  <updated>2026-05-13T15:00:45.533+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmp46smud00byp6qva4zwgzm6"/>
  <title>Europe - West (London): New resources not able to be created</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 26 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Jobs, Services</p>
    <p><small>May <var data-var='date'> 13</var>, <var data-var='time'>15:00:45</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident.

Existing running workloads are unaffected..</p>
<p><small>May <var data-var='date'> 13</var>, <var data-var='time'>15:33:29</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We are seeing an improvement; workloads are starting to be created.

Connectivity has been restored, and we are continuing to monitor the situation..</p>
<p><small>May <var data-var='date'> 13</var>, <var data-var='time'>16:27:13</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmp05l0ln00fo96bh1a4pufch</id>
  <published>2026-05-10T18:30:00.000+00:00</published>
  <updated>2026-05-10T18:30:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmp05l0ln00fo96bh1a4pufch"/>
  <title>Active DDoS Attack in Europe - West - Netherlands</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 2 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Services</p>
    <p><small>May <var data-var='date'> 10</var>, <var data-var='time'>18:30:00</var> GMT+0</small><br /><strong>Identified</strong> -
  We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..</p>
<p><small>May <var data-var='date'> 10</var>, <var data-var='time'>20:10:07</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We have put in place mitigations and are monitoring the affected region..</p>
<p><small>May <var data-var='date'> 10</var>, <var data-var='time'>20:31:56</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmoxb63od01es6n0u18am5oyu</id>
  <published>2026-05-08T19:00:00.000+00:00</published>
  <updated>2026-05-08T19:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmoxb63od01es6n0u18am5oyu"/>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 36 minutes</p>
    <p><strong>Affected Components:</strong> Certificates</p>
    <p><small>May <var data-var='date'> 8</var>, <var data-var='time'>19:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  [Let&#039;s Encrypt is currently experiencing a partial service disruption](https://letsencrypt.status.io/) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons, domains and BYOC clusters..</p>
<p><small>May <var data-var='date'> 8</var>, <var data-var='time'>21:36:09</var> GMT+0</small><br /><strong>Resolved</strong> -
  Certificate generation is functioning normally again. We will continue to monitor the situation..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmnpp03h1006jd0s28d2teq25</id>
  <published>2026-04-08T06:15:00.000+00:00</published>
  <updated>2026-04-08T06:15:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmnpp03h1006jd0s28d2teq25"/>
  <title>Addon and volume availability issues in US - Central</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 52 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Services</p>
    <p><small>Apr <var data-var='date'> 8</var>, <var data-var='time'>06:15:00</var> GMT+0</small><br /><strong>Identified</strong> -
  We are aware of an issue with a node affecting some stateful workloads in US - Central.

We are working on a fix for this incident..</p>
<p><small>Apr <var data-var='date'> 8</var>, <var data-var='time'>07:07:27</var> GMT+0</small><br /><strong>Resolved</strong> -
  We have recovered the node, and all workloads are now running again..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmn651uh000a880hcl3ox9cxf</id>
  <published>2026-03-25T14:28:04.826+00:00</published>
  <updated>2026-03-25T14:28:04.826+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmn651uh000a880hcl3ox9cxf"/>
  <title>Active DDoS Attack in Europe - West - Netherlands</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 3 hours and 21 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Jobs, Services</p>
    <p><small>Mar <var data-var='date'> 25</var>, <var data-var='time'>14:28:04</var> GMT+0</small><br /><strong>Identified</strong> -
  We are currently experiencing a large DDoS attack in the Europe - West - Netherlands region. This is affecting public ingress into the cluster. Internal traffic is not affected. The team is working on additional mitigations..</p>
<p><small>Mar <var data-var='date'> 25</var>, <var data-var='time'>16:01:00</var> GMT+0</small><br /><strong>Monitoring</strong> -
  The region is currently stable. We are still actively mitigating the ongoing attack..</p>
<p><small>Mar <var data-var='date'> 25</var>, <var data-var='time'>17:48:48</var> GMT+0</small><br /><strong>Resolved</strong> -
  The DDoS attack has stopped we will continue to monitor the affected region..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmmth1qit0cwahxt9crdwuuo3</id>
  <published>2026-03-16T17:42:54.132+00:00</published>
  <updated>2026-03-16T18:16:28.120+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmmth1qit0cwahxt9crdwuuo3"/>
  <title>App UI is experiencing loading issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 34 minutes</p>
    <p><strong>Affected Components:</strong> Northflank App</p>
    <p><small>Mar <var data-var='date'> 16</var>, <var data-var='time'>18:16:28</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>
<p><small>Mar <var data-var='date'> 16</var>, <var data-var='time'>18:15:58</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We identified the root cause as a system component causing disproportionate database load. This is affecting read performance leading to degraded performance for the app UI. We have implemented a fix and are monitoring the situation..</p>
<p><small>Mar <var data-var='date'> 16</var>, <var data-var='time'>17:42:54</var> GMT+0</small><br /><strong>Identified</strong> -
  We are investing an incident where the UI is failing to load correctly..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmmj4y7ct01zw1jix7jv6r4vd</id>
  <published>2026-03-09T12:06:32.914+00:00</published>
  <updated>2026-03-09T12:06:32.914+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmmj4y7ct01zw1jix7jv6r4vd"/>
  <title>Degraded performance in US Central region</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 29 minutes</p>
    <p><strong>Affected Components:</strong> Networking, Addons, Services</p>
    <p><small>Mar <var data-var='date'> 9</var>, <var data-var='time'>12:06:32</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident..</p>
<p><small>Mar <var data-var='date'> 9</var>, <var data-var='time'>12:35:56</var> GMT+0</small><br /><strong>Resolved</strong> -
  The root cause has been identified as an issue with the GCP infrastructure control plane and has been resolved.  
  
There was minimal impact to running user workloads. The main impact was a delay in provisioning net new or redeploying existing workloads..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmltl7z5d002trjwmeqyabys9</id>
  <published>2026-02-19T15:00:00.487+00:00</published>
  <updated>2026-02-19T15:00:00.487+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmltl7z5d002trjwmeqyabys9"/>
  <title>Issues with builds using the Heroku 24 Buildpack</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 34 minutes</p>
    <p><strong>Affected Components:</strong> Builds</p>
    <p><small>Feb <var data-var='date'> 19</var>, <var data-var='time'>15:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  We are aware of an issue where builds fail to start when using the Heroku 24 buildpack image.

We are currently working on identifying a solution. We will provide an update when the fix has been released..</p>
<p><small>Feb <var data-var='date'> 19</var>, <var data-var='time'>16:34:09</var> GMT+0</small><br /><strong>Resolved</strong> -
  We have released a fix for the issue..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmj1i24qr0106uhdob8a4bf2w</id>
  <published>2025-12-11T13:54:32.305+00:00</published>
  <updated>2025-12-11T13:54:32.305+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmj1i24qr0106uhdob8a4bf2w"/>
  <title>Node infrastructure stability - London Region</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 39 minutes</p>
    <p><strong>Affected Components:</strong> Networking, , Addons, Jobs, Services, 
Northflank Platform →</p>
    <p><small>Dec <var data-var='date'> 11</var>, <var data-var='time'>13:54:32</var> GMT+0</small><br /><strong>Identified</strong> -
  There are currently issues with node stability in the London region leading to partial outages for some workloads. The team has identified the issue and is implementing a mitigation..</p>
<p><small>Dec <var data-var='date'> 11</var>, <var data-var='time'>14:13:28</var> GMT+0</small><br /><strong>Monitoring</strong> -
  The mitigation is in place and the region is operating as expected..</p>
<p><small>Dec <var data-var='date'> 11</var>, <var data-var='time'>14:33:03</var> GMT+0</small><br /><strong>Resolved</strong> -
  The incident has been resolved and the cause identified. A node image security release lead to the host filesystem going into a read only mode when interacting with a specific set of workloads causing the node to become unresponsive. We have implemented a mitigation and are working on a permanent solution..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmfyn4drv02axox9ea8otzvu7</id>
  <published>2025-09-24T23:53:50.567+00:00</published>
  <updated>2025-09-24T23:53:50.567+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmfyn4drv02axox9ea8otzvu7"/>
  <title>Docker Hub is experiencing issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 44 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Jobs, Services</p>
    <p><small>Sep <var data-var='date'> 24</var>, <var data-var='time'>23:53:50</var> GMT+0</small><br /><strong>Identified</strong> -
  Dockerhub is currently experiencing issues related to authentication. This may cause issues with starting jobs, addons and services. &lt;https://www.dockerstatus.com/pages/incident/533c6539221ae15e3f000031/68d47a2f93c09e05486d93a9&gt;.</p>
<p><small>Sep <var data-var='date'> 25</var>, <var data-var='time'>01:11:08</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We are observing that image pull requests for Docker Hub images are now succeeding..</p>
<p><small>Sep <var data-var='date'> 25</var>, <var data-var='time'>01:38:09</var> GMT+0</small><br /><strong>Resolved</strong> -
  Dockerhub is fully operational and image pulling is now working as expected..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmfmovpks003r739gifi0hnsh</id>
  <published>2025-09-16T15:09:51.058+00:00</published>
  <updated>2025-09-16T15:09:51.058+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmfmovpks003r739gifi0hnsh"/>
  <title>Delay in DNS propagation</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 18 hours and 29 minutes</p>
    <p><strong>Affected Components:</strong> Addons, Services, Certificates</p>
    <p><small>Sep <var data-var='date'> 16</var>, <var data-var='time'>15:09:51</var> GMT+0</small><br /><strong>Monitoring</strong> -
  Our DNS provider NS1 is experiencing delays in DNS propagation.

This is affecting service and addon creation.  
  
NS1 status page: &lt;https://status.ai-apps-comms.ibm.com/ns1connect&gt;.</p>
<p><small>Sep <var data-var='date'> 17</var>, <var data-var='time'>09:39:01</var> GMT+0</small><br /><strong>Resolved</strong> -
  The incident has been resolved on NS1 and we have not been seeing any more related failures over the last hour..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmernfevf002amj1q4pgwr3xn</id>
  <published>2025-08-25T21:48:19.642+00:00</published>
  <updated>2025-08-25T21:48:19.642+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmernfevf002amj1q4pgwr3xn"/>
  <title>Issue affecting local disk caching for builds running on Northflank&#039;s infrastructure</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 17 minutes</p>
    <p><strong>Affected Components:</strong> Builds</p>
    <p><small>Aug <var data-var='date'> 25</var>, <var data-var='time'>21:48:19</var> GMT+0</small><br /><strong>Identified</strong> -
  There is an issue with the underlying storage which provides the local disk-based caching feature for builds on Northflank&#039;s infrastructure. BYOC builds and builds without a local cache will be unaffected.

  
We are currently working on a fix for this..</p>
<p><small>Aug <var data-var='date'> 25</var>, <var data-var='time'>22:45:24</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We implemented a fix and are currently monitoring the result..</p>
<p><small>Aug <var data-var='date'> 25</var>, <var data-var='time'>23:04:54</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmdy7qu4001ym91lf2txeit0h</id>
  <published>2025-08-05T07:23:59.728+00:00</published>
  <updated>2025-08-05T07:23:59.728+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmdy7qu4001ym91lf2txeit0h"/>
  <title>Log &amp; Metric Ingestion and Query Outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour</p>
    <p><strong>Affected Components:</strong> Logs and Metrics</p>
    <p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>07:23:59</var> GMT+0</small><br /><strong>Investigating</strong> -
  We are currently investigating this incident..</p>
<p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>07:55:18</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We implemented a fix and are currently monitoring the result..</p>
<p><small>Aug <var data-var='date'> 5</var>, <var data-var='time'>08:23:56</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmddhhx1500cbb87xjzw4t6ee</id>
  <published>2025-07-21T19:13:49.958+00:00</published>
  <updated>2025-07-21T19:13:49.958+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmddhhx1500cbb87xjzw4t6ee"/>
  <title>Let&#039;s Encrypt Certificate Generation Issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 11 hours and 31 minutes</p>
    <p><strong>Affected Components:</strong> Certificates</p>
    <p><small>Jul <var data-var='date'> 21</var>, <var data-var='time'>19:13:49</var> GMT+0</small><br /><strong>Monitoring</strong> -
  [Let&#039;s Encrypt is currently experiencing an outage](https://letsencrypt.status.io/pages/incident/55957a99e800baa4470002da/687e8d62b8a4e804fad85799) leading to certificate generation delays and failures on the Northflank platform. This will impact provisioning of new addons or domains..</p>
<p><small>Jul <var data-var='date'> 22</var>, <var data-var='time'>06:44:30</var> GMT+0</small><br /><strong>Resolved</strong> -
  Certificate generation is functioning without issue. We will continue to monitor the situation..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmbtqad4j000l6didp4d0xkfy</id>
  <published>2025-06-12T18:00:00.000+00:00</published>
  <updated>2025-06-12T20:40:23.833+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmbtqad4j000l6didp4d0xkfy"/>
  <title>Issues with Google Cloud platform - Affecting multiple services</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 45 minutes</p>
    <p><strong>Affected Components:</strong> Builds, Addons, Logs and Metrics</p>
    <p><small>Jun <var data-var='date'> 12</var>, <var data-var='time'>20:40:23</var> GMT+0</small><br /><strong>Monitoring</strong> -
  We are observing improved operations in the US - Central region. Addon creation and backup issues are now resolved. Ongoing Issues: \* GCP BYOC cluster processing.</p>
<p><small>Jun <var data-var='date'> 12</var>, <var data-var='time'>18:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  We have noticed issues with Google Cloud APIs. This seems to be widespread across the Google Cloud platform. We are currently determining the impact of the outage, and we will provide updates when we know more. Currently, logs, metrics, builds, BYOC and addon backups may be affected..</p>
<p><small>Jun <var data-var='date'> 12</var>, <var data-var='time'>19:21:45</var> GMT+0</small><br /><strong>Monitoring</strong> -
  Google Cloud has acknowledged the incident. For detailed information, please visit: &lt;https://status.cloud.google.com/incidents/ow5i3PPK96RduMcb1SsW&gt;

Most services have been restored to normal operation. However, we are still experiencing the following issues: 

* Addon creation in the US - Central region
* Addon backups in the US - Central region
* GCP BYOC cluster processing

We will continue monitoring the situation and provide updates as they become available..</p>
<p><small>Jun <var data-var='date'> 12</var>, <var data-var='time'>20:44:39</var> GMT+0</small><br /><strong>Resolved</strong> -
  GCP BYOC cluster processing is now operational..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmbsjf79f0025k18t40kcbyb0</id>
  <published>2025-06-11T22:00:00.000+00:00</published>
  <updated>2025-06-11T22:00:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmbsjf79f0025k18t40kcbyb0"/>
  <title>Metrics ingestion and query outage</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 17 minutes</p>
    <p><strong>Affected Components:</strong> Logs and Metrics</p>
    <p><small>Jun <var data-var='date'> 11</var>, <var data-var='time'>22:00:00</var> GMT+0</small><br /><strong>Identified</strong> -
  Metrics query and ingestion are currently affected. Logs are unaffected.

We are currently applying a fix for the issue and monitoring the results..</p>
<p><small>Jun <var data-var='date'> 11</var>, <var data-var='time'>23:16:34</var> GMT+0</small><br /><strong>Monitoring</strong> -
  The fix has been rolled out and we are currently monitoring the ingestion rate and catch-up..</p>
<p><small>Jun <var data-var='date'> 12</var>, <var data-var='time'>00:16:31</var> GMT+0</small><br /><strong>Resolved</strong> -
  The ingestion rate has been stable for 30 minutes and historic writes have caught up.

We sincerely apologise for the service disruption you may have experienced. We are investigating the root cause to prevent similar issues in the future..</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cmay5h4pg0052zbome4c6m5q6</id>
  <published>2025-05-21T16:04:00.000+00:00</published>
  <updated>2025-05-21T16:04:00.000+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cmay5h4pg0052zbome4c6m5q6"/>
  <title>Infrastructure stability issues</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 10 minutes</p>
    <p><strong>Affected Components:</strong> Networking, Addons, Services</p>
    <p><small>May <var data-var='date'> 21</var>, <var data-var='time'>16:04:00</var> GMT+0</small><br /><strong>Monitoring</strong> -
  .</p>
<p><small>May <var data-var='date'> 21</var>, <var data-var='time'>17:13:33</var> GMT+0</small><br /><strong>Resolved</strong> -
  **We sincerely apologise for the outage you experienced.**  

**What happened:**  
Yesterday, we rolled out an update to support ARM architecture across our platform.  
As part of that, a migration was initiated today to extend support to x86 architecture within our database schemas.  
This triggered a higher-than-expected volume of database rescheduling events, which in turn put excessive pressure on the underlying GCP infrastructure. Some nodes failed under this load, compounding the issue and causing degraded service for several customers.

**Timeline of events (UTC):**

* **15:36** — Northflank engineers were alerted to failing health checks on customer endpoints.
* **15:41** — Elevated 503 errors were detected from a specific cluster.
* **15:42** — An internal incident was formally declared.
* **15:57** — Replacement infrastructure began provisioning to recover failed nodes.
* **16:32** — All customer addons returned to a healthy state, with the exception of four databases which required manual intervention.

  
Thank you for your patience—we know how critical uptime is, and we are reviewing and improving our systems to ensure this doesn’t happen again.  
  
Will Stewart  
CEO @ Northflank.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cm391ojva000bzr1411i47iz0</id>
  <published>2024-11-08T18:02:57.501+00:00</published>
  <updated>2024-11-08T19:20:22.960+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cm391ojva000bzr1411i47iz0"/>
  <title>US - Central - Inability to Create Disk Snapshots</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 17 minutes</p>
    <p><strong>Affected Components:</strong> Addons</p>
    <p><small>Nov <var data-var='date'> 8</var>, <var data-var='time'>19:20:22</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>
<p><small>Nov <var data-var='date'> 8</var>, <var data-var='time'>18:02:57</var> GMT+0</small><br /><strong>Monitoring</strong> -
  Google Cloud is experiencing issues with snapshot upload in the US Central region.

This is causing delays for addon snapshot backups within the region.</p>

        ]]>
  </content>
</entry>

<entry>
  <id>tag:status.northflank.com,2005:Incident/cm252qf810036lb09qejd495o</id>
  <published>2024-10-11T18:41:37.083+00:00</published>
  <updated>2024-10-11T18:41:37.083+00:00</updated>
  <link rel="alternate" type="text/html" href="https://status.northflank.com/incident/cm252qf810036lb09qejd495o"/>
  <title>DNS Propagation Delay</title>

  <content type="html">
  <![CDATA[
    <p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 27 minutes</p>
    <p><strong>Affected Components:</strong> Services, Addons, Certificates</p>
    <p><small>Oct <var data-var='date'> 11</var>, <var data-var='time'>18:41:37</var> GMT+0</small><br /><strong>Monitoring</strong> -
  Our DNS provider is experiencing issues with DNS propagation leading to delays in provisioning new workloads and certificates. Existing workloads are unaffected.

&lt;https://ns1connect.trust.pagerduty.com/posts/details/PK35B1W&gt;.</p>
<p><small>Oct <var data-var='date'> 11</var>, <var data-var='time'>20:09:04</var> GMT+0</small><br /><strong>Resolved</strong> -
  This incident has been resolved..</p>

        ]]>
  </content>
</entry>

</feed>