I made a Grafana Master-dashboard to monitor my Kubernetes cluster.
Features:
- Health bar at the top, that will turn amber or red when things go wrong, and send push notifications to my phone
- CPU, Memory and Network utilization
- Database and Storage backup statuses
I also added CPU/Memory usage request monitoring so that I can tune how much CPU/Memory a pod requests in the cluster.

And active alerts

I’ve muted these two, I need to add a 3rd node with more CPU and ram, currently if one node goes down there isn’t enough redundancy for the cluster to just keep working.
Lots more detail here: https://erasmus.works/ It’s all Open-Source, leave a Star on my repo if you like what you see.


Yeah we could have some more of that here on Lemmy because people tend to be quite negative.
Honestly that’s awesome, I didn’t manage to get grafana working when I tried exactly this, I’ll definitely check out any guides you used ^^
Thanks!
I took a lot of inspiration from here https://github.com/onedr0p/home-ops, there is also a discord server on there where you can ask people for help, they are quite helpful.
I mostly shopped around and found parts and implementations I liked and coppied that and made it my own