I’m curious as to which tools and technologies you all are using to keep track of all those services you are deploying, whether it be resource tracking, network traffic, logs, traces, or uptime.

As a bonus question, how have you organized your network or your services to reduce the overhead of implementing observability?

        • ℍ𝕂-𝟞𝟝@sopuli.xyz
          link
          fedilink
          English
          arrow-up
          2
          ·
          edit-2
          4 hours ago

          Mimir is just the DB, no metric scrapers, but it’s distributed and capable of high volume ingestion, retention and querying. Can’t scrape metrics though.

          Alloy has no storage persistence, so it’s just the scraper part, but is massively scalable, its basically a Prometheus Agent on steroids. Can do logs, traces, Pyroscope profiles too.

          So instead of deploying one server to scrape and store metrics, you deploy one to store and one to scrape.