GitHub is having an outage and this results in a high number of failed jobs. We are watching the situation closely.
GitHub is back operational.
We're investigating. Some active metrics might be stale.
The network provider has fixed the incident. It was congestion on one of the uplinks.
The write-path has been stabilised, meaning metrics are properly ingested again. We'll continue to stabilise the whole situation.
Investigations already started earlier this morning, while unfortunately at approx. 12:30 the situation worsened - impacting ingestion at that point.