<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Medchat Status - Incident history</title>
    <link>https://status.medchatapp.com</link>
    <description>Medchat</description>
    <pubDate>Mon, 2 Mar 2026 17:47:09 +0000</pubDate>
    
<item>
  <title>Issues logging into Medchat</title>
  <description>
    Type: Incident
    Duration: 15 minutes

    Affected Components: Cloud Infrastructure, SMS Provider, Admin UI, Custom OIDC SSO, Admin UI, Web Application, Agent UI, Admin UI, Webhooks, Google SSO, Widget, Medchat Auth Application, Agent UI, Agent UI, Custom SAML SSO
    Mar 2, 17:47:09 GMT+0 - Investigating - We are currently investigating this incident. Mar 2, 17:52:38 GMT+0 - Identified - Users should be able to login successfully now. We are continuing to monitor the issue. Mar 2, 18:02:29 GMT+0 - Resolved - This incident has been resolved. Mar 2, 18:09:22 GMT+0 - Postmortem - **Date:** 3/2/26  
**Service:** Production Application (Azure App Service)  
**Duration of Impact:** 9:34 AM PST – 9:49 AM PST  
**Impact:** Intermittent connection loss and service instability

We experienced service instability affecting our production environment. The issue was traced to one of our Azure App Service resources within the active production environment. To restore stability quickly, we executed a blue/green environment swap, removing the impacted App Service from the production mix.

Following the swap, services returned to normal operation and stability was restored.

We are continuing to investigate the underlying cause of the App Service degradation. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 15 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 2&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:47:09&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 2&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:52:38&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Users should be able to login successfully now. We are continuing to monitor the issue..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 2&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:02:29&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Mar &lt;var data-var=&#039;date&#039;&gt; 2&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:09:22&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Postmortem&lt;/strong&gt; -
  **Date:** 3/2/26  
**Service:** Production Application (Azure App Service)  
**Duration of Impact:** 9:34 AM PST – 9:49 AM PST  
**Impact:** Intermittent connection loss and service instability

We experienced service instability affecting our production environment. The issue was traced to one of our Azure App Service resources within the active production environment. To restore stability quickly, we executed a blue/green environment swap, removing the impacted App Service from the production mix.

Following the swap, services returned to normal operation and stability was restored.

We are continuing to investigate the underlying cause of the App Service degradation..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 2 Mar 2026 17:47:09 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/cmm9h19ml07wz14nj76j4ch36</link>
  <guid>https://status.medchatapp.com/incident/cmm9h19ml07wz14nj76j4ch36</guid>
</item>

<item>
  <title>System login and performance issue</title>
  <description>
    Type: Incident
    Duration: 1 hour and 2 minutes

    Affected Components: , , Cloud Infrastructure, SMS Provider, Admin UI, Custom OIDC SSO, , , , Admin UI, , Web Application, Agent UI, Admin UI, Webhooks, Google SSO, Widget, Medchat Auth Application, Agent UI, Agent UI, Custom SAML SSO, 
Team Chat → 
Journeys → 
Live Chat → 
Text Chat → 
General → 
Authentication →
    Nov 13, 21:56:57 GMT+0 - Investigating - We are currently investigating this incident; sporadic system login and slowness -- likely caused by an outage in Azure.  Nov 13, 22:17:57 GMT+0 - Monitoring - We implemented a fix and are currently monitoring the result. Nov 13, 22:58:28 GMT+0 - Resolved - This incident has been resolved. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 2 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;21:56:57&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident; sporadic system login and slowness -- likely caused by an outage in Azure. .&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:17:57&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We implemented a fix and are currently monitoring the result..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:58:28&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 13 Nov 2025 21:56:57 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/cmhxyyp93085bvuvu6q5zx5zp</link>
  <guid>https://status.medchatapp.com/incident/cmhxyyp93085bvuvu6q5zx5zp</guid>
</item>

<item>
  <title>System performance</title>
  <description>
    Type: Incident
    Duration: 2 hours and 12 minutes

    Affected Components: Custom OIDC SSO, Cloud Infrastructure, Admin UI, SMS Provider, Google SSO, Medchat Auth Application, Custom SAML SSO, Agent UI, Admin UI, Widget, Admin UI, Agent UI, Web Application, Agent UI, Webhooks
    Jul 28, 17:22:00 GMT+0 - Identified - We are continuing to work on a fix for this incident. Jul 28, 16:00:23 GMT+0 - Investigating - We are currently investigating this incident. Jul 28, 17:45:00 GMT+0 - Monitoring - We have implemented mitigations to return to normal operations. We will continue to monitor the system over the next few hours to ensure all components are fully functional. Jul 28, 18:11:55 GMT+0 - Resolved - This incident has been resolved. Jul 29, 16:15:28 GMT+0 - Postmortem - Root cause was due to the application&#039;s capacity limits as a result of a recent increase in traffic. The system was operating as expected but required additional resources to support the elevated load.

Resolution was to scale out the app&#039;s Azure infrastructure to accommodate the increased demand. Services have since stabilized, and team is continuing to monitor performance closely to ensure continued reliability. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 12 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 28&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:22:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  We are continuing to work on a fix for this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 28&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:00:23&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 28&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:45:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have implemented mitigations to return to normal operations. We will continue to monitor the system over the next few hours to ensure all components are fully functional..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 28&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;18:11:55&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 29&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:15:28&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Postmortem&lt;/strong&gt; -
  Root cause was due to the application&#039;s capacity limits as a result of a recent increase in traffic. The system was operating as expected but required additional resources to support the elevated load.

Resolution was to scale out the app&#039;s Azure infrastructure to accommodate the increased demand. Services have since stabilized, and team is continuing to monitor performance closely to ensure continued reliability..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Mon, 28 Jul 2025 16:00:23 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/cmdndm2jk00b613yp9atbjp05</link>
  <guid>https://status.medchatapp.com/incident/cmdndm2jk00b613yp9atbjp05</guid>
</item>

<item>
  <title>System outage</title>
  <description>
    Type: Incident
    Duration: 1 hour and 44 minutes

    Affected Components: Webhooks, Admin UI, Agent UI, Custom OIDC SSO, Web Application, Google SSO, Medchat Auth Application, Cloud Infrastructure, Agent UI, Admin UI, SMS Provider, Custom SAML SSO, Admin UI, Widget, Agent UI
    Feb 13, 22:11:00 GMT+0 - Identified - Around 4:11 PM CST, Medchat&#039;s authentication service became unresponsive after a planned release. The team identified widespread database connection issues in one application. After initial triage, Medchat reverted to the previous application versions, which resulted in the system becoming responsive again around 4:39 PM CST. Feb 13, 22:39:00 GMT+0 - Monitoring - Around 4:11 PM CST, Medchat&#039;s authentication service became unresponsive after a planned release. The team identified widespread database connection issues in one application. After initial triage, Medchat reverted to the previous application versions, which resulted in the system becoming responsive again around 4:39 PM CST. Feb 13, 23:55:29 GMT+0 - Resolved - We will continue to monitor the system to ensure operational status. Medchat will perform a more detailed analysis of the root cause within the next 24 hours. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour and 44 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Feb &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:11:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Around 4:11 PM CST, Medchat&#039;s authentication service became unresponsive after a planned release. The team identified widespread database connection issues in one application. After initial triage, Medchat reverted to the previous application versions, which resulted in the system becoming responsive again around 4:39 PM CST..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Feb &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;22:39:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Around 4:11 PM CST, Medchat&#039;s authentication service became unresponsive after a planned release. The team identified widespread database connection issues in one application. After initial triage, Medchat reverted to the previous application versions, which resulted in the system becoming responsive again around 4:39 PM CST..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Feb &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:55:29&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  We will continue to monitor the system to ensure operational status. Medchat will perform a more detailed analysis of the root cause within the next 24 hours..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 13 Feb 2025 22:11:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/cm73xs7sz0013en9kx5spxab5</link>
  <guid>https://status.medchatapp.com/incident/cm73xs7sz0013en9kx5spxab5</guid>
</item>

<item>
  <title>Application Service Interruption</title>
  <description>
    Type: Incident
    Duration: 6 hours and 13 minutes

    Affected Components: Webhooks, Admin UI, Agent UI, Custom OIDC SSO, Web Application, Google SSO, Medchat Auth Application, Cloud Infrastructure, Agent UI, Admin UI, SMS Provider, Custom SAML SSO, Admin UI, Widget, Agent UI
    Jul 18, 23:06:53 GMT+0 - Investigating - We are currently investigating this incident.

At the moment, we&#039;re seeing widespread outage with Microsoft Azure, Medchat&#039;s Cloud Services provider. Microsoft has just sent out a status update that they are investigating issues in their Central US region. Jul 18, 23:20:47 GMT+0 - Identified - Confirmed that the issue is due to Microsoft Azure. From their website:

---

**Investigating issues in the Central US region**

**Impact Statement:** Starting at approximately 21:56 UTC on 18 Jul 2024, a subset of customers may experience issues with multiple Azure services in the Central US region including failures with service management operations and connectivity or availability of services.

**Current Status:** We are aware of this issue and are actively investigating. The next update will be provided in 60 minutes, or as events warrant.

This message was last updated at 23:17 UTC on 18 July 2024 Jul 18, 23:49:47 GMT+0 - Identified - Latest update from Microsoft @4:45 pm PT

---

**Impact Statement:** Starting at 21:56 UTC on 18 Jul 2024, a subset of customers may experience issues with multiple Azure services in the Central US region including failures with service management operations and connectivity or availability of services.

**Current Status:** We are aware of this issue and have engaged multiple teams to investigate. As part of the investigation, we are reviewing previous deployments, and are running other workstreams to investigate for an underlying cause. The next update will be provided in 60 minutes, or as events warrant. 

This message was last updated at 23:45 UTC on 18 July 2024 Jul 18, 23:57:05 GMT+0 - Identified - Update from Microsoft @4:56 pm PT:

---

**Current status:** We have determined this issue was impacted by an underlying storage outage in the Central US region that the services were dependent upon. Once the underlying storage outage is mitigated, this impact will be resolved. The next update will be in 2 hours, or as events warrants. Jul 19, 01:32:58 GMT+0 - Monitoring - Microsoft update at 6:31 pm PT:

---

**Current Status:** ... We’ve determined the underlying cause and are currently working towards mitigation. We will start to see incremental recovery in next 90 minutes. The next update will be provided in 60 minutes, or as events warrant. Jul 19, 02:51:14 GMT+0 - Monitoring - Microsoft update @7:33 pm PT

* **Status:** Incident is now mitigated
* **Next steps:** Engineers will continue to investigate to establish the full root cause and prevent future occurrences.

---

We are currently performing checks on the Medchat Application to ensure overall system health; we will continue monitoring for another hour or so. Jul 19, 04:25:16 GMT+0 - Monitoring - While the Medchat application is back online and seems to be back in a fully functional state, Microsoft is continuing to send status updates (last at 9:10 pm PT) that the incident is still active, and that customers should continue to see increasing recovery at this time as residual and downstream impact mitigation progresses.  
  
Medchat team will continue to monitor system health and status updates. Jul 19, 05:19:34 GMT+0 - Resolved - While Microsoft is keeping their incident tickets open, they have confirmed that majority of their impacted services have now recovered. They have an updated dashboard where all services on the affected region are now back online.

Root cause per Microsoft: The underlying cause was due to a backend cluster management workflow deployed a configuration change that caused backend access to be blocked between a subset of Azure Storage clusters and compute resources in the Central US region. This resulted in the compute resources automatically restarting when connectivity was lost to virtual disks.

The Medchat team also conducted some health checks to ensure the system is back up and fully functional. Closing this incident.

--- 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 6 hours and 13 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:06:53&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident.

At the moment, we&#039;re seeing widespread outage with Microsoft Azure, Medchat&#039;s Cloud Services provider. Microsoft has just sent out a status update that they are investigating issues in their Central US region..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:20:47&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Confirmed that the issue is due to Microsoft Azure. From their website:

---

**Investigating issues in the Central US region**

**Impact Statement:** Starting at approximately 21:56 UTC on 18 Jul 2024, a subset of customers may experience issues with multiple Azure services in the Central US region including failures with service management operations and connectivity or availability of services.

**Current Status:** We are aware of this issue and are actively investigating. The next update will be provided in 60 minutes, or as events warrant.

This message was last updated at 23:17 UTC on 18 July 2024.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:49:47&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Latest update from Microsoft @4:45 pm PT

---

**Impact Statement:** Starting at 21:56 UTC on 18 Jul 2024, a subset of customers may experience issues with multiple Azure services in the Central US region including failures with service management operations and connectivity or availability of services.

**Current Status:** We are aware of this issue and have engaged multiple teams to investigate. As part of the investigation, we are reviewing previous deployments, and are running other workstreams to investigate for an underlying cause. The next update will be provided in 60 minutes, or as events warrant. 

This message was last updated at 23:45 UTC on 18 July 2024.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 18&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:57:05&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  Update from Microsoft @4:56 pm PT:

---

**Current status:** We have determined this issue was impacted by an underlying storage outage in the Central US region that the services were dependent upon. Once the underlying storage outage is mitigated, this impact will be resolved. The next update will be in 2 hours, or as events warrants..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;01:32:58&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Microsoft update at 6:31 pm PT:

---

**Current Status:** ... We’ve determined the underlying cause and are currently working towards mitigation. We will start to see incremental recovery in next 90 minutes. The next update will be provided in 60 minutes, or as events warrant..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;02:51:14&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Microsoft update @7:33 pm PT

* **Status:** Incident is now mitigated
* **Next steps:** Engineers will continue to investigate to establish the full root cause and prevent future occurrences.

---

We are currently performing checks on the Medchat Application to ensure overall system health; we will continue monitoring for another hour or so..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;04:25:16&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  While the Medchat application is back online and seems to be back in a fully functional state, Microsoft is continuing to send status updates (last at 9:10 pm PT) that the incident is still active, and that customers should continue to see increasing recovery at this time as residual and downstream impact mitigation progresses.  
  
Medchat team will continue to monitor system health and status updates..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jul &lt;var data-var=&#039;date&#039;&gt; 19&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;05:19:34&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  While Microsoft is keeping their incident tickets open, they have confirmed that majority of their impacted services have now recovered. They have an updated dashboard where all services on the affected region are now back online.

Root cause per Microsoft: The underlying cause was due to a backend cluster management workflow deployed a configuration change that caused backend access to be blocked between a subset of Azure Storage clusters and compute resources in the Central US region. This resulted in the compute resources automatically restarting when connectivity was lost to virtual disks.

The Medchat team also conducted some health checks to ensure the system is back up and fully functional. Closing this incident.

---.&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Thu, 18 Jul 2024 23:06:53 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/clyrvs6du57078hcofvy6bpklb</link>
  <guid>https://status.medchatapp.com/incident/clyrvs6du57078hcofvy6bpklb</guid>
</item>

<item>
  <title>504 Errors During Authentication</title>
  <description>
    Type: Incident
    Duration: 2 hours and 16 minutes

    Affected Components: Custom SAML SSO, Custom OIDC SSO, Google SSO, Medchat Auth Application
    Jan 23, 19:14:00 GMT+0 - Investigating - Around 1:14 PM CST, MedChat customers began experiencing 504 errors when attempting to authenticate with MedChat. The outage is due to a exceptional and prolonged spike of database activity. The root cause is still under investigation. We have initiated scaling of DB resources to alleviate the outage. Jan 23, 19:56:00 GMT+0 - Monitoring - The DB scaling operations are complete and the system appears to be returning to normal operations. We will continue to monitor and investigate the root cause of the unusual activity. Jan 23, 21:30:08 GMT+0 - Resolved - When today&#039;s outage began, connection limit exceptions started appearing in our telemetry data. By scaling the DB resources, we were able to increase our maximum number of connections, allowing the system to recover. Production support personnel are continuing to investigate automations, metrics, and proactive alerts that may be used to avoid similar problems in the future.  
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 2 hours and 16 minutes</p>
    <p><strong>Affected Components:</strong> , , , </p>
    &lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 23&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:14:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  Around 1:14 PM CST, MedChat customers began experiencing 504 errors when attempting to authenticate with MedChat. The outage is due to a exceptional and prolonged spike of database activity. The root cause is still under investigation. We have initiated scaling of DB resources to alleviate the outage..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 23&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;19:56:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  The DB scaling operations are complete and the system appears to be returning to normal operations. We will continue to monitor and investigate the root cause of the unusual activity..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 23&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;21:30:08&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  When today&#039;s outage began, connection limit exceptions started appearing in our telemetry data. By scaling the DB resources, we were able to increase our maximum number of connections, allowing the system to recover. Production support personnel are continuing to investigate automations, metrics, and proactive alerts that may be used to avoid similar problems in the future. .&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 23 Jan 2024 19:14:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/clrqrxy51141929b8oi7lb3b5rl</link>
  <guid>https://status.medchatapp.com/incident/clrqrxy51141929b8oi7lb3b5rl</guid>
</item>

<item>
  <title>App Service Interruption</title>
  <description>
    Type: Incident
    Duration: 1 hour

    Affected Components: Webhooks, Admin UI, Widget, Custom SAML SSO, Agent UI, Custom OIDC SSO, Agent UI, Web Application, Google SSO, Medchat Auth Application, Cloud Infrastructure, Agent UI, Admin UI, SMS Provider, Admin UI
    Jan 13, 00:00:00 GMT+0 - Resolved - This incident has been resolved. Jan 12, 23:00:00 GMT+0 - Investigating - MedChat&#039;s App Services in the Azure East US region appear to be experiencing an outage resulting in 502 responses. We are preparing to swap services to West US to restore functionality. Jan 12, 23:30:00 GMT+0 - Monitoring - We have swapped our app services to the West US region. Initial checks show service has been restored. We will continue to monitor for lingering issues. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 hour</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 13&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;00:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  This incident has been resolved..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:00:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  MedChat&#039;s App Services in the Azure East US region appear to be experiencing an outage resulting in 502 responses. We are preparing to swap services to West US to restore functionality..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Jan &lt;var data-var=&#039;date&#039;&gt; 12&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;23:30:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  We have swapped our app services to the West US region. Initial checks show service has been restored. We will continue to monitor for lingering issues..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Fri, 12 Jan 2024 23:00:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/clrb9hcy26617ben4gtb0n1li</link>
  <guid>https://status.medchatapp.com/incident/clrb9hcy26617ben4gtb0n1li</guid>
</item>

<item>
  <title>Degraded keyword detection</title>
  <description>
    Type: Incident
    Duration: 4 hours and 33 minutes

    Affected Components: Agent UI
    Nov 8, 21:22:00 GMT+0 - Identified - An Azure hardware outage is affecting the keyword detection component of Live Chat. Azure is aware of the issue and is working to restore service. Nov 9, 00:59:54 GMT+0 - Monitoring - Affected Azure resources are coming back online. The keyword detection components are operational once again. We will continue monitoring the system for further disruptions. Nov 9, 01:55:02 GMT+0 - Resolved - All operations have returned to normal. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 4 hours and 33 minutes</p>
    <p><strong>Affected Components:</strong> </p>
    &lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 8&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;21:22:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Identified&lt;/strong&gt; -
  An Azure hardware outage is affecting the keyword detection component of Live Chat. Azure is aware of the issue and is working to restore service..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 9&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;00:59:54&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  Affected Azure resources are coming back online. The keyword detection components are operational once again. We will continue monitoring the system for further disruptions..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Nov &lt;var data-var=&#039;date&#039;&gt; 9&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;01:55:02&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  All operations have returned to normal..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Wed, 8 Nov 2023 21:22:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/cloqcvbor178578beollzd392fn</link>
  <guid>https://status.medchatapp.com/incident/cloqcvbor178578beollzd392fn</guid>
</item>

<item>
  <title>Sporadic 502 errors</title>
  <description>
    Type: Incident
    Duration: 57 minutes

    Affected Components: Webhooks, Admin UI, Admin UI, Widget, Custom SAML SSO, Agent UI, Custom OIDC SSO, Agent UI, Web Application, Google SSO, Medchat Auth Application, Cloud Infrastructure, Agent UI, Admin UI, SMS Provider
    Oct 10, 14:54:00 GMT+0 - Investigating - The site is experiencing sporadic 502 errors on UI app services Oct 10, 15:51:00 GMT+0 - Resolved - All of MedChat&#039;s Azure app services (development through production) experienced issues between 9:54 AM CST to 10:51 AM CST where service availability was intermittent. The root cause is still unknown at this time, but all app services have returned to normal operation. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 57 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;14:54:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  The site is experiencing sporadic 502 errors on UI app services.&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 10&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;15:51:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  All of MedChat&#039;s Azure app services (development through production) experienced issues between 9:54 AM CST to 10:51 AM CST where service availability was intermittent. The root cause is still unknown at this time, but all app services have returned to normal operation..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 10 Oct 2023 14:54:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/clnkils6789474c2n4fsb5iyrx</link>
  <guid>https://status.medchatapp.com/incident/clnkils6789474c2n4fsb5iyrx</guid>
</item>

<item>
  <title>System Inaccessible</title>
  <description>
    Type: Incident
    Duration: 1 day, 22 hours and 52 minutes

    Affected Components: Webhooks, Admin UI, Admin UI, Widget, Custom SAML SSO, Agent UI, Custom OIDC SSO, Agent UI, Web Application, Google SSO, Medchat Auth Application, Cloud Infrastructure, Agent UI, Admin UI, SMS Provider
    Oct 3, 16:35:00 GMT+0 - Investigating - We are currently investigating this incident. Users have been logged out of the application and are no longer able to sign in. Oct 3, 16:53:00 GMT+0 - Monitoring - The root cause appears to be a downtime from scaling out DB resources. The team is continuing to investigate the outage, which was not expected from the scaling operation.

The scaling operation is now complete and the system appears to be back to normal operations. Continuing to monitor. Oct 3, 17:21:51 GMT+0 - Resolved - Throughout the morning of 10/3/23 and in the weeks leading up to this incident, the support team received automated alerts from the system regarding elevated (though not critical) DB resource consumption. The elevated DB consumption would quickly return to acceptable levels without affecting any users.

After observing the pattern over several weeks, the support team applied a request to scale out the DB, an operation which was not expected to cause any downtime. The scaling operation took approximately 18 minutes, during which the system was largely inaccessible to MedChat users. While the databases were still available during the scaling operation, they appear to have had a reduced capacity (further investigation still in progress). Once scaling completed, the system returned to normal operations.

In the future, all DB scaling will be performed off-hours to avoid the potential for service disruptions. 
  </description>
  <content:encoded>
    <![CDATA[<p><strong>Type:</strong> Incident</p>
    <p><strong>Duration:</strong> 1 day, 22 hours and 52 minutes</p>
    <p><strong>Affected Components:</strong> , , , , , , , , , , , , , , </p>
    &lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 3&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:35:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Investigating&lt;/strong&gt; -
  We are currently investigating this incident. Users have been logged out of the application and are no longer able to sign in..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 3&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;16:53:00&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Monitoring&lt;/strong&gt; -
  The root cause appears to be a downtime from scaling out DB resources. The team is continuing to investigate the outage, which was not expected from the scaling operation.

The scaling operation is now complete and the system appears to be back to normal operations. Continuing to monitor..&lt;/p&gt;
&lt;p&gt;&lt;small&gt;Oct &lt;var data-var=&#039;date&#039;&gt; 3&lt;/var&gt;, &lt;var data-var=&#039;time&#039;&gt;17:21:51&lt;/var&gt; GMT+0&lt;/small&gt;&lt;br&gt;&lt;strong&gt;Resolved&lt;/strong&gt; -
  Throughout the morning of 10/3/23 and in the weeks leading up to this incident, the support team received automated alerts from the system regarding elevated (though not critical) DB resource consumption. The elevated DB consumption would quickly return to acceptable levels without affecting any users.

After observing the pattern over several weeks, the support team applied a request to scale out the DB, an operation which was not expected to cause any downtime. The scaling operation took approximately 18 minutes, during which the system was largely inaccessible to MedChat users. While the databases were still available during the scaling operation, they appear to have had a reduced capacity (further investigation still in progress). Once scaling completed, the system returned to normal operations.

In the future, all DB scaling will be performed off-hours to avoid the potential for service disruptions..&lt;/p&gt;
]]>
  </content:encoded>
  <pubDate>Tue, 3 Oct 2023 16:35:00 +0000</pubDate>
  <link>https://status.medchatapp.com/incident/clnakbhr460680bdoeciduaivm</link>
  <guid>https://status.medchatapp.com/incident/clnakbhr460680bdoeciduaivm</guid>
</item>

  </channel>
  </rss>