Two scrapers, one copy of the truth — replica labels and deduplication at write or query
A redundant scraper pair sends every sample twice under different replica labels, so the cluster must keep one stream or every sum doubles — either by accepting only an elected replica and paying a short failover gap, or by storing both and merging at query time for double the storage.
Now that the cluster keeps three safe copies of whatever it receives, the next question is what happens when two redundant scrapers hand it every sample twice.
Scene 10a
Two scrapers, one copy of the truth
- Watch
- Try it
- Predict
- Capture
How do you run redundant scrapers without counting everything twice? One scraper is a single point of failure, so teams run two of them side by side: same configuration, same targets, same interval. Both push everything they collect into the cluster, and the cluster from the last scene faithfully stores whatever it is handed. Watch the fleet requests-per-second line fill in against the dashed line of the traffic that is really happening.
Where this sits in Metrics / Monitoring System
Scene 10a of 18, in the Scale out act — Remote-write, ingesters, dedup, blocks, split queries.. A redundant scraper pair sends every sample twice, so the cluster keeps one stream: elect one replica at write time and accept a short failover gap, or store both and merge at query time for twice the storage.
Up next. Now that exactly one stream per series lands in memory, the next question is how the cluster keeps months of it without holding it all in memory.
All 18 scenes in Metrics / Monitoring System · Every curriculum