I will stand up a monitoring stack for your servers and services: Prometheus and Grafana with Node Exporter and cAdvisor, ready-made dashboards, alert rules that matter (not noise), and centralized logs (ELK/Loki on request). You'll see problems before your users do.
What you get:
- Prometheus + Grafana (Node Exporter, cAdvisor) via Docker Compose
- Dashboards for hosts and your key services
- Alert rules with sensible thresholds and first-action notes
- A short runbook
What I need from you:
- Host/service access, what's critical to watch, and an alert channel (Telegram/email/etc)