Is Blacksmith down?
Last checked just nowAffected component: Github → Webhooks
Blacksmith is reporting degraded performance right now (last checked just now). Services are up but slower or partially failing.
Real-time Blacksmith status, recent outages, and incident history — pulled directly from Blacksmith's official status page at https://status.blacksmith.sh every 5 minutes. Pingoru tracks 49 Blacksmith services and has captured 76 incidents in the last 90 days (94.59% uptime). Get email, Slack, Discord, or webhook alerts the moment Blacksmith reports a new incident — free for 3 monitors, no credit card.
Recent outages & incidents
Past 90 days- eu-central ARMeu-central x86us-west ARMus-west x86eu-west x86us-central MacOSeu-central Storage Clusterus-west Storage Clustereu-central Storage Clusterus-west Storage Cluster
Timeline · 6 updates
- investigating · Sep 29, 2026, 09:10 PM UTC
Jobs across all regions are taking longer to start and cache operations are failing. We are continuing to investigate the root cause.
- identified · Sep 29, 2026, 09:28 PM UTC
We have applied a fix and are monitoring its effect. Customers may still see delayed job starts and failing cache operations while recovery completes. We will provide an update within the next 30 minutes.
- monitoring · Sep 29, 2026, 09:43 PM UTC
Our mitigation has taken effect and job start times and cache operations are improving. Customers may still see some delayed job starts and cache failures while recovery completes. We will provide an update within the next 30 minutes.
- monitoring · Sep 29, 2026, 10:20 PM UTC
Caching has recovered, and job queue times across EU regions have recovered. Queue times for larger jobs (16 and 32 vcpu jobs) in US west and US east continue to remain elevated. We are continuing to monitor for any regressions.
- monitoring · Sep 29, 2026, 10:42 PM UTC
This incident is resolved. Job starts and cache operations returned to normal in all regions, and we are continuing to monitor. Jobs that failed during the incident can be re-run.
- resolved · Sep 29, 2026, 11:04 PM UTC
This incident has been resolved.
Latest: This incident has been resolved.
-
- US West Cache
Timeline · 5 updates
- investigating · Sep 29, 2026, 05:00 PM UTC
We are investigating failing GitHub Actions cache operations for jobs running in us-west since approximately 16:10 UTC. Affected jobs may see cache restores and saves fail and fall back to a full install, which can make those jobs run longer or fail, while jobs in other regions are not affected.
- identified · Sep 29, 2026, 05:09 PM UTC
We have identified the cause of the failures, and we are implementing a fix to restore cache service. Customers with jobs in us-west may still see cache restores and saves fail and fall back to a full install, which can make those jobs run longer or fail, while jobs in other regions are not affected.
- monitoring · Sep 29, 2026, 05:35 PM UTC
We have restored the cache infrastructure in us-west, and cache failures have decreased but are not yet back to normal. Jobs in us-west may still see some cache restores and saves fail and fall back to a full install, while the earlier delayed job starts have cleared and other regions are unaffected. We are bringing additional cache capacity online to fully restore service.
- monitoring · Sep 29, 2026, 06:06 PM UTC
Cache errors for jobs in us-west have largely subsided. As we complete recovery, some jobs may see a one-time cache miss on their next run. We are monitoring and will provide an update within the next hour.
- resolved · Sep 29, 2026, 06:32 PM UTC
This incident is resolved. GitHub Actions cache operations for jobs in us-west have returned to normal, and any jobs that failed during the incident can be re-run.
Latest: This incident is resolved. GitHub Actions cache operations for jobs in us-west have returned to normal, and any jobs that failed during the incident can be re-run.
-
- us-west ARMus-west x86
Timeline · 3 updates
- identified · Sep 28, 2026, 08:52 PM UTC
A rack-level power failure at our us-west data center caused some jobs to fail with runner communication errors. The issue has been communicated with our upstream provider and we are identifying the affected machines to mitigate the impact to our customers.
- monitoring · Sep 28, 2026, 09:53 PM UTC
The affected rack remains offline while our upstream provider works to restore its power, and jobs in us-west are running normally on the rest of the fleet. Any jobs that failed with runner communication errors can be safely re-run.
- resolved · Sep 28, 2026, 10:14 PM UTC
This incident has been resolved.
Latest: This incident has been resolved.
-
- us-west ARMus-west x86
Timeline · 3 updates
- investigating · Sep 24, 2026, 07:16 PM UTC
We are investigating degraded network connectivity between our us-west region and GitHub. Customers running jobs in us-west may see slower-than-normal git checkouts, while jobs in other regions are not affected.
- monitoring · Sep 24, 2026, 07:31 PM UTC
After rerouting traffic, congestion on the upstream network link has subsided, and git checkout performance in us-west has returned to normal. We are monitoring and will resolve this incident once performance has continued to stay stable.
- resolved · Sep 24, 2026, 07:50 PM UTC
Git checkout performance in us-west has remained stable since we rerouted traffic away from the congested upstream network link, and this incident is now resolved.
Latest: Git checkout performance in us-west has remained stable since we rerouted traffic away from the congested upstream network link, and this incident is now resolved.
-
- eu-central ARMeu-central x86us-west ARMus-west x86eu-west x86us-central MacOSAPIeu-central Storage Clusterus-west Storage Clustereu-central Storage Cluster
Timeline · 2 updates
- identified · Sep 21, 2026, 06:37 PM UTC
We are seeing elevated errors in job adoption and other control plane APIs. This is yielding higher latencies for your job to be adopted. You may also see certain actions fail, such as actions caching and sticky disks. We have identified the issue and are rolling out a mitigation.
- resolved · Sep 21, 2026, 06:47 PM UTC
This issue has been resolved, and all services have operating normally. If any of your jobs failed during this time, please re-run them.
Latest: This issue has been resolved, and all services have operating normally. If any of your jobs failed during this time, please re-run them.
-
See the full Blacksmith outage history
57 more incidents in the last 90 days, plus the full multi-year archive of per-service events and update timelines.
Browse Blacksmith outage history →Or sign up free to get alerts when Blacksmith breaks · 10 free monitors · No credit card
- Started Sep 29, 2026, 09:10 PM UTC · Resolved Sep 29, 2026, 11:04 PM UTC · 1h 54m
- US west cache failure ResolvedStarted Sep 29, 2026, 05:00 PM UTC · Resolved Sep 29, 2026, 06:32 PM UTC · 1h 32m
- Job failures in us-west ResolvedStarted Sep 28, 2026, 08:52 PM UTC · Resolved Sep 28, 2026, 10:14 PM UTC · 1h 22m
- Started Sep 24, 2026, 07:16 PM UTC · Resolved Sep 24, 2026, 07:50 PM UTC · 33m
- Started Sep 21, 2026, 06:37 PM UTC · Resolved Sep 21, 2026, 06:47 PM UTC · 9m
- Degraded US East AWS connectivity ResolvedStarted Sep 13, 2026, 01:11 AM UTC · Resolved Sep 13, 2026, 01:58 AM UTC · 46m
- Degraded US East AWS connectivity ResolvedStarted Sep 13, 2026, 01:11 AM UTC · Resolved Sep 13, 2026, 01:11 AM UTC · —
- Started Sep 11, 2026, 05:30 AM UTC · Resolved Sep 11, 2026, 05:30 AM UTC · —