Use cases
Software Products E-commerce MSPs Schools Development & Marketing DevOps Agencies Help Desk
Company
Internet Status Blog Pricing Log in Get started free

Outage in Coveralls

Increased latency for large repos

Resolved Minor
July 10, 2026 - Started 14 days ago - Lasted 3 days
Official incident page

Incident Report

Summary AI Generated

Coveralls experienced increased latency for large repositories (5,000+ source files) over 72 hours due to elevated background job queue traffic, affecting the Coveralls.io Web and API. The issue escalated into a broader outage when a surge of large repo uploads overwhelmed individual web servers, causing widespread 504 errors for most users. Outlier repos were temporarily paused to relieve queue pressure, web servers self-recovered, and the incident was fully resolved, with a post-mortem planned.

We are monitoring increased latency for larger repos (5K+ source files) due to elevated traffic in background job queues. We will pause outlier repos (more than 2x average traffic) to allow queues to clear for the general population, then restore once queues have cleared. If you think your repo may have been one of those paused, two things: 1) Reach out to us at support@coveralls.io and we'll confirm; and 2) Consider purchasing a usage add-on so that your jobs are not throttled for fair-use, or an enterprise plan with isolated infrastructure for maximum performance: https://coveralls.io/pricing
Components affected
Coveralls .io Web Coveralls .io API

Trusted by 1,000+ teams

The Status Page Aggregator with Early Outage Detection

Stop finding out about outages from your users. Monitor 6,320+ cloud services and get alerted the second something breaks.

IsDown status aggregator dashboard
Latest Updates ( sorted recent to last )
RESOLVED 11 days ago - at 07/13/2026 02:56PM

This incident has been resolved but we believe it triggered a worsened incident overnight Sun night/Mon morning (US PDT) which has just been resolved. To be confirmed by full RCA, we believe a deluge of large repo uploads tied up individual web servers that handle frontline requests. Each web server is able to recover on its own, and did, but as volume increased all servers were eventually affected, only allowing short windows where requests could get through—rejecting most requests with 504 errors.

We'll post a post-mortem when we understand more about what happened and how to prevent it going forward.

MONITORING 14 days ago - at 07/10/2026 05:34PM

We are continuing to monitor this situation. Latency is much reduced but still elevated. We will close this incident when it's fully restored to normal.

IDENTIFIED 14 days ago - at 07/10/2026 03:00PM

We are monitoring increased latency for larger repos (5K+ source files) due to elevated traffic in background job queues. We will pause outlier repos (more than 2x average traffic) to allow queues to clear for the general population, then restore once queues have cleared. If you think your repo may have been one of those paused, two things:

1) Reach out to us at support@coveralls.io and we'll confirm; and
2) Consider purchasing a usage add-on so that your jobs are not throttled for fair-use, or an enterprise plan with isolated infrastructure for maximum performance: https://coveralls.io/pricing

The Status Page Aggregator with Early Outage Detection

With IsDown, you can monitor all your critical services' official status pages from one centralized dashboard and receive instant alerts the moment an outage is detected. Say goodbye to constantly checking multiple sites for updates and stay ahead of outages with IsDown.

Start free trial

No credit card required · Cancel anytime · 6320 services available

Integrations with Slack Microsoft Teams Google Chat Datadog PagerDuty Zapier Discord Webhook