Prometheus: Metrics Collection ๐
Metrics collection system for early warning about problems.
Functions:
- ๐ฅ๏ธ Server metrics: CPU, RAM, disk, network via node_exporter
- ๐ Service availability: blackbox HTTP/TCP/ICMP checks
- ๐ Collecting metrics from applications: Nextcloud, Home Assistant, and others
- ๐จ Alertmanager: notifications in Telegram/Discord when thresholds are exceeded
- ๐ Queries via PromQL for deep analysis
How it works:
- Prometheus collects metrics on a schedule (scrape)
- Charts are displayed in Grafana
- On anomaly, an alert is triggered - the administrator receives a notification
For administrators: Alert rules, recording rules, federation of metrics, long-term storage via Thanos.
Access: via Grafana (grafana.potatoenergy.ru) โข according to Potato Energy credentials (management is only based on the rights of the admin group)