Use cases
Software Products E-commerce MSPs Schools Development & Marketing DevOps Agencies Help Desk
Company
Internet Status Blog Pricing Log in Get started free

Outage in Coalesce

Job Execution Failure

Resolved Major
August 17, 2026 - Started 18 days ago - Lasted about 1 hour
Official incident page

Incident Report

We have identified an issue with jobs executing and investigating the root cause. We will update this incident when resolved.
Components affected
Coalesce Scheduler

Trusted by 1,000+ teams

The Status Page Aggregator with Early Outage Detection

Stop finding out about outages from your users. Monitor 6,320+ cloud services and get alerted the second something breaks.

IsDown status aggregator dashboard
Latest Updates ( sorted recent to last )
RESOLVED 18 days ago - at 08/17/2026 01:40PM

Between approximately 10:00 and 13:30 UTC today, some customers experienced delayed or stuck job runs.

During a routine, provider-managed Kubernetes upgrade in one of our US regions, our cloud provider was unable to provision replacement compute capacity. Instead of aborting, the upgrade proceeded to take existing nodes out of rotation, leaving the region without capacity to start new jobs.

We rolled back the upgrade, restoring full scheduling capacity, and jobs are processing normally. We were forced to cancel older jobs that were unable to recover and recommend customers re-run any effected operations.

We are adjusting our upgrade configuration and maintenance windows to prevent this scenario from recurring.

MONITORING 18 days ago - at 08/17/2026 01:17PM

We have identified the issue and pushed a change to resolve job processing. More details to follow once fully resolved.

INVESTIGATING 18 days ago - at 08/17/2026 12:56PM

We are investigating an issue affecting job runs for customers in our US region. A portion of scheduled and API-triggered job runs are failing to start. An affected run may remain in a Waiting state and then fail with the message "Run timed out. Restart the
run, and if this issue persists contact support." Runs that do start are completing normally.

Customers in our EU, Canada, and APAC regions are not affected. Our engineering team is actively investigating. We will post an update by 9:30 AM ET

INVESTIGATING 18 days ago - at 08/17/2026 12:54PM

We have identified an issue with jobs executing and investigating the root cause. We will update this incident when resolved.

The Status Page Aggregator with Early Outage Detection

With IsDown, you can monitor all your critical services' official status pages from one centralized dashboard and receive instant alerts the moment an outage is detected. Say goodbye to constantly checking multiple sites for updates and stay ahead of outages with IsDown.

Start free trial

No credit card required · Cancel anytime · 6320 services available

Integrations with Slack Microsoft Teams Google Chat Datadog PagerDuty Zapier Discord Webhook