We were made aware of an issue affecting a significant subset of PrivateLink customers not supporting HTTP/2. This issue persisted from ~14:45 UTC to 18:20 UTC. Customers affected may notice a gap in their data during this time.
Due to this bug reported in https://github.com/kubernetes/kubernetes/issues/127370, we were affected by an issue causing K8S service endpoints not getting updated when pods are stopped/started if there are more than 1k pods matching the service.
This caused a temporary outage in Mimir gossiping services, which further resulted in failures to ingest and query metrics for a short time.
This issue has been resolved.