We're investigating. Some active metrics might be stale.
The network provider has fixed the incident. It was congestion on one of the uplinks.
The write-path has been stabilised, meaning metrics are properly ingested again. We'll continue to stabilise the whole situation.
Investigations already started earlier this morning, while unfortunately at approx. 12:30 the situation worsened - impacting ingestion at that point.