Degraded performance of JIRA Automation

Incident Report for Jira

Postmortem

Summary

On Jul 6, 2026, between 06:51 and 20:17 UTC, Atlassian customers using cloud products in European (EU) regions experienced service disruptions affecting automation rule execution, user search, user picker, and related workflows. The incident began when a core identity service experienced database saturation in EU regions, increasing latency and error rates. While responders worked to restore capacity, an emergency mitigation was applied which blocked automation rules from accessing the identity endpoint. In parallel, the elevated identity latency contributed to a cascading failure in a downstream user search service, degrading user search and user picker experiences across multiple products. Service was progressively restored after additional database read capacity was added, the emergency block was removed, and the user search service was manually scaled up.

Impact

The incident affected customers across multiple Atlassian products in EU regions on Jul 6, 2026 between 06:51 and 20:17 UTC.

  • Core Identity service degradation: elevated intermittent access denied rates were observed across products in EU regions.
  • Automation rule execution failures: automation rules using the default user (Automation for Jira) in Jira, Jira Service Management, and Jira Product Discovery in EU regions failed with a permissions error.
  • User search and user picker failures: user search success rates dropped significantly in EU regions, with user picker capability largely unavailable across products.

Root Cause

The incident originated from a recent configuration change that reduced the cache lifetime for a core identity service. As EU workday traffic ramped up on July 6, 2026, a larger share of requests began reaching the backing database directly rather than being served from cache. The EU databases did not have sufficient regional capacity to absorb the increased load, and database CPU reached saturation, causing elevated intermittent access denied rates.

As part of the mitigation, an emergency block was applied to reduce load on the saturated identity service database and prioritize restoring the core identity service. This inadvertently prevented automation rules from passing their pre-execution permission checks, causing them to fail. The extended duration was primarily due to a secondary wave of automation rule failures caused by the processing limit throttling.

In parallel, the sustained degradation caused increased latency and intermittent failures across products that depend on these checks, and contributed to a cascading failure in a downstream user search service. As a result, user search and user picker experiences became largely unavailable in EU regions until the permissions service recovered and the user search service was scaled up to restore capacity.

Remedial Actions Plan & Next Steps

We know outages impact customers' productivity. Atlassian is prioritizing the following actions to help prevent similar incidents in future:

  • Improve safeguards for high-impact emergency mitigations: Harden operational tooling so mitigation steps that could disable customer-facing workflows require additional review.
  • Strengthen identity infrastructure capacity and change safety: Review regional database sizing and capacity headroom to reduce the risk of saturation during peak traffic periods, and refine the assessment process for configuration changes so downstream load impact is evaluated before production rollout.
  • Strengthen resilience against cascading failures from upstream degradation: Audit backpressure handling, circuit breaker configuration, and autoscaling behavior in services, so that upstream degradation does not cause broader product impact across user search, user picker, and related experiences.
  • Harden replay mechanisms for Automation rule failures: Increase resilience of automation rules to reduce impact associated with incidents occurring in dependencies..

We recognize how critical reliable product workflows are for our customers, and we apologize to customers who were impacted by this incident.

Thanks,

Atlassian Customer Support

Posted Jul 27, 2026 - 06:57 UTC

Resolved

On July 6, 2026, some users in Europe regions may have experienced performance degradation with Automation for JIRA. The issue has now been resolved, and the service is operating normally for all affected customers.

If you believe the issue continues to persist, please contact our Support team for further assistance and troubleshooting.
Posted Jul 06, 2026 - 18:52 UTC

Monitoring

The performance degradation of Automation for JIRA in Europe regions has been resolved, and services are now operating normally for all affected customers. We will continue to monitor performance closely to confirm stability.
Posted Jul 06, 2026 - 18:41 UTC

Identified

We have identified the likely cause of the issue, and our teams are diligently working on a mitigation. Affected users may experience performance degradation affecting Automation for JIRA in Europe regions. We will continue to share additional updates here as more information is available.
Posted Jul 06, 2026 - 17:57 UTC

Update

We continue to investigate performance degradation affecting Automation for JIRA in Europe regions. We will share updates here as more information becomes available.
Posted Jul 06, 2026 - 17:12 UTC

Investigating

We are actively investigating reports of performance degradation affecting Automation for JIRA. We will share updates here as more information is available.
Posted Jul 06, 2026 - 16:11 UTC
This incident affected: Viewing content, Create and edit, Authentication and User Management, Search, Notifications, Administration, Marketplace, Mobile, Purchasing & Licensing, Signup, and Automation for Jira.