Is Blacksmith down?

Last checked just now
Current status
Blacksmith has degraded performance

Affected component: Github → Webhooks

Official status page: https://status.blacksmith.sh · Polled every 5 minutes · 49 components tracked

Blacksmith is reporting degraded performance right now (last checked just now). Services are up but slower or partially failing.

Real-time Blacksmith status, recent outages, and incident history — pulled directly from Blacksmith's official status page at https://status.blacksmith.sh every 5 minutes. Pingoru tracks 49 Blacksmith services and has captured 76 incidents in the last 90 days (94.59% uptime). Get email, Slack, Discord, or webhook alerts the moment Blacksmith reports a new incident — free for 3 monitors, no credit card.

Users who monitor Blacksmith also follow these CI/CD services: PDQ Depot AppVeyor ServiceRocket Subsplash Happo PactFlow Qlty Software SimplyQ Progress Telerik View all 6,000+ providers
Blacksmith uptime 94.59% uptime · past 90 days
Mon Wed Fri
JulAugSepOct
Less More

Recent outages & incidents

Past 90 days
  1. Resolved 1h 54m
    Started Sep 29, 2026, 09:10 PM UTC · Resolved Sep 29, 2026, 11:04 PM UTC
    eu-central ARMeu-central x86us-west ARMus-west x86eu-west x86us-central MacOSeu-central Storage Clusterus-west Storage Clustereu-central Storage Clusterus-west Storage Cluster
    Timeline · 6 updates
    • investigating · Sep 29, 2026, 09:10 PM UTC

      Jobs across all regions are taking longer to start and cache operations are failing. We are continuing to investigate the root cause.

    • identified · Sep 29, 2026, 09:28 PM UTC

      We have applied a fix and are monitoring its effect. Customers may still see delayed job starts and failing cache operations while recovery completes. We will provide an update within the next 30 minutes.

    • monitoring · Sep 29, 2026, 09:43 PM UTC

      Our mitigation has taken effect and job start times and cache operations are improving. Customers may still see some delayed job starts and cache failures while recovery completes. We will provide an update within the next 30 minutes.

    • monitoring · Sep 29, 2026, 10:20 PM UTC

      Caching has recovered, and job queue times across EU regions have recovered. Queue times for larger jobs (16 and 32 vcpu jobs) in US west and US east continue to remain elevated. We are continuing to monitor for any regressions.

    • monitoring · Sep 29, 2026, 10:42 PM UTC

      This incident is resolved. Job starts and cache operations returned to normal in all regions, and we are continuing to monitor. Jobs that failed during the incident can be re-run.

    • resolved · Sep 29, 2026, 11:04 PM UTC

      This incident has been resolved.

    Latest: This incident has been resolved.

  2. Resolved 1h 32m
    Started Sep 29, 2026, 05:00 PM UTC · Resolved Sep 29, 2026, 06:32 PM UTC
    US West Cache
    Timeline · 5 updates
    • investigating · Sep 29, 2026, 05:00 PM UTC

      We are investigating failing GitHub Actions cache operations for jobs running in us-west since approximately 16:10 UTC. Affected jobs may see cache restores and saves fail and fall back to a full install, which can make those jobs run longer or fail, while jobs in other regions are not affected.

    • identified · Sep 29, 2026, 05:09 PM UTC

      We have identified the cause of the failures, and we are implementing a fix to restore cache service. Customers with jobs in us-west may still see cache restores and saves fail and fall back to a full install, which can make those jobs run longer or fail, while jobs in other regions are not affected.

    • monitoring · Sep 29, 2026, 05:35 PM UTC

      We have restored the cache infrastructure in us-west, and cache failures have decreased but are not yet back to normal. Jobs in us-west may still see some cache restores and saves fail and fall back to a full install, while the earlier delayed job starts have cleared and other regions are unaffected. We are bringing additional cache capacity online to fully restore service.

    • monitoring · Sep 29, 2026, 06:06 PM UTC

      Cache errors for jobs in us-west have largely subsided. As we complete recovery, some jobs may see a one-time cache miss on their next run. We are monitoring and will provide an update within the next hour.

    • resolved · Sep 29, 2026, 06:32 PM UTC

      This incident is resolved. GitHub Actions cache operations for jobs in us-west have returned to normal, and any jobs that failed during the incident can be re-run.

    Latest: This incident is resolved. GitHub Actions cache operations for jobs in us-west have returned to normal, and any jobs that failed during the incident can be re-run.

  3. Resolved 1h 22m
    Started Sep 28, 2026, 08:52 PM UTC · Resolved Sep 28, 2026, 10:14 PM UTC
    us-west ARMus-west x86
    Timeline · 3 updates
    • identified · Sep 28, 2026, 08:52 PM UTC

      A rack-level power failure at our us-west data center caused some jobs to fail with runner communication errors. The issue has been communicated with our upstream provider and we are identifying the affected machines to mitigate the impact to our customers.

    • monitoring · Sep 28, 2026, 09:53 PM UTC

      The affected rack remains offline while our upstream provider works to restore its power, and jobs in us-west are running normally on the rest of the fleet. Any jobs that failed with runner communication errors can be safely re-run.

    • resolved · Sep 28, 2026, 10:14 PM UTC

      This incident has been resolved.

    Latest: This incident has been resolved.

  4. Resolved 33m
    Started Sep 24, 2026, 07:16 PM UTC · Resolved Sep 24, 2026, 07:50 PM UTC
    us-west ARMus-west x86
    Timeline · 3 updates
    • investigating · Sep 24, 2026, 07:16 PM UTC

      We are investigating degraded network connectivity between our us-west region and GitHub. Customers running jobs in us-west may see slower-than-normal git checkouts, while jobs in other regions are not affected.

    • monitoring · Sep 24, 2026, 07:31 PM UTC

      After rerouting traffic, congestion on the upstream network link has subsided, and git checkout performance in us-west has returned to normal. We are monitoring and will resolve this incident once performance has continued to stay stable.

    • resolved · Sep 24, 2026, 07:50 PM UTC

      Git checkout performance in us-west has remained stable since we rerouted traffic away from the congested upstream network link, and this incident is now resolved.

    Latest: Git checkout performance in us-west has remained stable since we rerouted traffic away from the congested upstream network link, and this incident is now resolved.

  5. Resolved 9m
    Started Sep 21, 2026, 06:37 PM UTC · Resolved Sep 21, 2026, 06:47 PM UTC
    eu-central ARMeu-central x86us-west ARMus-west x86eu-west x86us-central MacOSAPIeu-central Storage Clusterus-west Storage Clustereu-central Storage Cluster
    Timeline · 2 updates
    • identified · Sep 21, 2026, 06:37 PM UTC

      We are seeing elevated errors in job adoption and other control plane APIs. This is yielding higher latencies for your job to be adopted. You may also see certain actions fail, such as actions caching and sticky disks. We have identified the issue and are rolling out a mitigation.

    • resolved · Sep 21, 2026, 06:47 PM UTC

      This issue has been resolved, and all services have operating normally. If any of your jobs failed during this time, please re-run them.

    Latest: This issue has been resolved, and all services have operating normally. If any of your jobs failed during this time, please re-run them.

See the full Blacksmith outage history

57 more incidents in the last 90 days, plus the full multi-year archive of per-service events and update timelines.

Browse Blacksmith outage history →

Or sign up free to get alerts when Blacksmith breaks · 10 free monitors · No credit card

Outage history

Past 90 days · 62 incidents View full outage history →