What it is
Dashboards with real-time charts of server and service health.
How we apply it
To see the load, the error count and whether everything is alive — on one screen, without logging into every server separately.
Dashboards with real-time charts of server and service health.
To see the load, the error count and whether everything is alive — on one screen, without logging into every server separately.
Grafana with Prometheus is part of our infrastructure stack for projects with many servers and metrics. In the smaller projects among our cases monitoring is simpler and needs no dashboard: container health checks and a Telegram message on failure. That is how the finance assistant works, with an external check every 5 minutes, and Peregovorka, where a failed database restore check reaches the administrator as a message.
For one server with a couple of bots, health checks and Telegram messages are enough: a dashboard would be pretty but unnecessary.
Terms explained — now let's get to work: tell us about the task and we'll turn it into a clear work plan.