I’m curious as to which tools and technologies you all are using to keep track of all those services you are deploying, whether it be resource tracking, network traffic, logs, traces, or uptime.

As a bonus question, how have you organized your network or your services to reduce the overhead of implementing observability?

    • ℍ𝕂-𝟞𝟝@sopuli.xyz
      link
      fedilink
      English
      arrow-up
      2
      ·
      edit-2
      4 hours ago

      Mimir is just the DB, no metric scrapers, but it’s distributed and capable of high volume ingestion, retention and querying. Can’t scrape metrics though.

      Alloy has no storage persistence, so it’s just the scraper part, but is massively scalable, its basically a Prometheus Agent on steroids. Can do logs, traces, Pyroscope profiles too.

      So instead of deploying one server to scrape and store metrics, you deploy one to store and one to scrape.