# How to graph the resource usage of servers?

**URL:** <https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563>\
**Category:** Homelab\
**Created:** [June 5, 2021, 4:00am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563 "2021-06-05T04:00:39Z")\
**Posts on this page:** 20\
**Page:** 1

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [June 5, 2021, 4:00am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/1 "2021-06-05T04:00:39Z")

</div>

Sorry if this is a dumb question… I don’t even know the correct terms to google for.

Is there an easy way to have different machines report their usage level to a single place. From there, is there a good tool to graph that information on a single page?

At this point, I am interested in tracking memory usage, CPU usage, and storage usage across a proxmox server and two different NASs, one Synology and one qnap.

Even just the correct keywords to search for would be appreciated.

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [June 9, 2021, 2:04pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/2 "2021-06-09T14:04:48Z")

</div>

To answer my own question. There seems to be a good example at [https://github.com/Mbarmem/Grafana.Dashboard](https://github.com/Mbarmem/Grafana.Dashboard)

---

<div class="post-metadata">

**Author:** ![jay](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/jay/32/649_2.png) [@jay](https://community.learnlinux.tv/u/jay)\
**Post date:** [June 10, 2021, 3:51pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/3 "2021-06-10T15:51:22Z")

</div>

I haven’t looked into it yet, but I’ve seen people use Grafana for this use-case.

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [June 10, 2021, 9:09pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/4 "2021-06-10T21:09:16Z")

</div>

Thanks,

This ranks pretty high on my accidental home lab priorities. I would be nice to have a single pane of glass to show memory, cpu, and storage consumption across all of my servers. Then if something seems to be nearing capacity I can shift things around or buy more of a specific resource.

I’ll go though the link I posted above a see what I can figure out. But, to be honest, I like your video style. They feel like they bring me up to the necessary level to understand have the various parts of a system fit together before I get down in the weeds.

David

---

<div class="post-metadata">

**Author:** ![jay](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/jay/32/649_2.png) [@jay](https://community.learnlinux.tv/u/jay)\
**Post date:** [June 16, 2021, 6:18pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/5 "2021-06-16T18:18:03Z")

</div>

I’m hoping to cover this topic as well as centralized logging as soon as I can. There’s some things I’ll catch up on and I’ll try to see where everything fits into my schedule. Sometimes the production quality going up can slow me down a bit, so I’m trying to figure out how to automate some of my video production so I can be quicker to cover topics.

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [June 17, 2021, 4:01am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/6 "2021-06-17T04:01:11Z")

</div>

Great.

The last time I really played with my network was in the early 2000’s when the NSLU2 came out and flashing router firmware was all the rage. Then, life and work seemed to take over.

Home labs and home networking (and google search) are now at a stage where one can set up a surprising polished system with just a few hours per week. Most importantly, for me, the tools are good enough that when something goes wrong I can find the issue or a least revert the change until I get some time to look into things.

---

<div class="post-metadata">

**Author:** ![Mr\_McBride](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/m/a698b9/32.png) [@Mr\_McBride](https://community.learnlinux.tv/u/Mr_McBride)\
**Post date:** [July 12, 2021, 12:52pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/7 "2021-07-12T12:52:57Z")

</div>

Grafana, InfluxDB, and Telegraf.

Telegraf is the agent you install on the host to push metrics into InfluxDB. Telegraf has a huge config file, but is very easy to configure. Grafana is the dashboard engine.

I installed all three on a Debian server last week and used Ansible to install Telegram on 4 VM’s. Then I configured Telegram. Had everything up an running within 20 minutes. I did, however, have to configure /etc/telegraf/telegraf.conf to enable network interface monitoring and specify the name of the interface you want to monitor.

Here is the article I used to perform the install and configuration:

> **[Monitor Linux System with Grafana and Telegraf | ComputingForGeeks](https://computingforgeeks.com/monitor-linux-system-with-grafana-and-telegraf/)**
>
> In this article, we're going to look at how to Monitor a Linux System with Grafana and Telegraf. Telegraf metrics will be stored on InfluxDB, then we can

---

<div class="post-metadata">

**Author:** ![ameinild](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/a/51bf81/32.png) [@ameinild](https://community.learnlinux.tv/u/ameinild)\
**Post date:** [July 12, 2021, 2:03pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/8 "2021-07-12T14:03:27Z")

</div>

I’m personally using Grafana and Prometheus as Docker containers, with prometheus-node-exporter running on each machine I’m monitoring. This works very well for me. 😊

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [July 12, 2021, 2:57pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/9 "2021-07-12T14:57:35Z")

</div>

Thanks,  
I appreciate the suggestion. I’ll check out Prometheus.

Unless I am mistaken, monitoring and reporting seem to be one of the things that users and companies are often willing to pay to do correctly. I went through the Graylog homelab episode over the weekend and go a good feeling for how that works.

---

<div class="post-metadata">

**Author:** ![KI7MT](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/ki7mt/32/260_2.png) [@KI7MT](https://community.learnlinux.tv/u/KI7MT)\
**Post date:** [July 15, 2021, 8:37am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/10 "2021-07-15T08:37:39Z")

</div>

Depending on your needs, but, many of the heavy hitters have already been mentioned in this thread. The traditional ELK Stack (Producers, Storage, Processors, Grapher’s ) is a good starting point.

I ran across this one and bookmarked it some time back due to it’s Jupyter Notebook / Spark components.

- [The HELK Project Docs](https://thehelk.com/intro.html)
- [The Github Project](https://github.com/Cyb3rWard0g/HELK)

I wouldn’t say it’s an entry level approach, but it certainly covers a lot of ground in the logging / analytics space.

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [July 15, 2021, 8:46am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/11 "2021-07-15T08:46:54Z")

</div>

Thanks, I’ll look into it.

I am thinking that I spent a ton of time working on getting the home/free versions of a vmware based infrastructure working. While it might be great in a large network, it was overly complex for my basic needs (and basic hardware.)

Proxmox meets my needs/willingness to deal with complexity almost perfectly. I am kind of hoping there is a similar type of monitoring tool.

---

<div class="post-metadata">

**Author:** ![Mr\_McBride](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/m/a698b9/32.png) [@Mr\_McBride](https://community.learnlinux.tv/u/Mr_McBride)\
**Post date:** [July 15, 2021, 12:07pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/12 "2021-07-15T12:07:44Z")

</div>

For monitoring anything, Zabbix is an excellent tool and will run on a raspberry pi.

ELK or Splunk is a bit advanced. If you are new to monitoring, I’d recommend starting off with Grafana or Zabbix. But there are other tools out there. Nagios is popular as well.

I’ve been working in the application and network monitoring area for the past 15 years, so let me know if I can help.

---

<div class="post-metadata">

**Author:** ![dfarning](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/d/b782af/32.png) [@dfarning](https://community.learnlinux.tv/u/dfarning)\
**Post date:** [July 17, 2021, 8:32am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/13 "2021-07-17T08:32:10Z")

</div>

If anyone else is learning along…

There is an excellent explanation of the Prometheus Architecture at [How Prometheus Monitoring works | Prometheus Architecture explained - YouTube](https://www.youtube.com/watch?v=h4Sl21AKiDg&list=RDCMUCdngmbVKX1Tgre699-XLlUA&start_radio=1)

I am going to try to set up a simple monitoring system that monitors my Proxmox server, my primary Synology NAS, my remote backup NAS, and my primary laptop with a couple of different systems.

As odd as it seems, my goal is a ‘set it and forget it’ system where I can trust my network is chugging along without any intervention by me.

Going to set up prometheus based system today… will try zabbix asap.

---

<div class="post-metadata">

**Author:** ![Mr\_McBride](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/m/a698b9/32.png) [@Mr\_McBride](https://community.learnlinux.tv/u/Mr_McBride)\
**Post date:** [July 17, 2021, 1:02pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/14 "2021-07-17T13:02:36Z")

</div>

Prometheus is also very popular and Grafana integrates very well it.

---

<div class="post-metadata">

**Author:** ![jay](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/jay/32/649_2.png) [@jay](https://community.learnlinux.tv/u/jay)\
**Post date:** [July 24, 2021, 11:29pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/15 "2021-07-24T23:29:06Z")

</div>

I like Grafana quite a bit. I was thinking of setting that up on my servers (which usually leads to a video) but I’ve gotten quite a bit behind schedule at the moment.

---

<div class="post-metadata">

**Author:** ![Mr\_McBride](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/m/a698b9/32.png) [@Mr\_McBride](https://community.learnlinux.tv/u/Mr_McBride)\
**Post date:** [July 26, 2021, 11:24am UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/16 "2021-07-26T11:24:54Z")

</div>

I’m currently running Grafana and InfluxDB from a VM, but they both run great on a RPi-4.

---

<div class="post-metadata">

**Author:** ![hypoiodous](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/h/f475e1/32.png) [@hypoiodous](https://community.learnlinux.tv/u/hypoiodous)\
**Post date:** [October 11, 2021, 4:22pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/17 "2021-10-11T16:22:39Z")

</div>

I would be interested in a video using Grafana and how to setup own graphs. More specifically, what things should we make available at a glance, given the limited “above the fold” space.

---

<div class="post-metadata">

**Author:** ![ameinild](https://community.learnlinux.tv/letter_avatar_proxy/v4/letter/a/51bf81/32.png) [@ameinild](https://community.learnlinux.tv/u/ameinild)\
**Post date:** [October 21, 2021, 12:20pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/18 "2021-10-21T12:20:25Z")

</div>

I made my own Home Screen, using the [Prometheus Node Exporter Full Dashboard](https://grafana.com/grafana/dashboards/1860) as a template.

 ![Grafana](https://community.learnlinux.tv/uploads/default/original/1X/c008b241e0ec01f6904eee99bdce7bfe65232b53.jpeg)

This has basic stats for my machines, as well as basic HDD and ZFS stats. And I link directly to the Node Exporter Full Dashboard for more detailed information.

---

<div class="post-metadata">

**Author:** ![KI7MT](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/ki7mt/32/260_2.png) [@KI7MT](https://community.learnlinux.tv/u/KI7MT)\
**Post date:** [December 5, 2021, 9:25pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/19 "2021-12-05T21:25:41Z")

</div>

@ [ameinild](https://community.learnlinux.tv/u/ameinild) - First, nice Dashboard !

I can’t got into great detail, but, my team manages several critical applications where we are responsible for all aspects of the deployment(s) (Bare metal servers, instances, patching, application performance, SLA’s etc). We have hundreds of nodes across virtually all regions where we have infrastructure (that list is growing rapidly). Prometheus and Grafana are two tools my team use extensively. When problems happen (it’s not a matter of if, it’s when), these tools (and others) prove invaluable for getting resources back on line and operating in peek condition.

In low volume scenarios, like most of us have in our home labs, the items you’ve chosen to monitor are solid choices (they apply equally at scale also). However, in High-Volume situation, things like: Networking, HTTP Requests, Stack Tracing, Thread Monitoring, Response Time Metrics, Database transactions, and more, all become critical in troubleshooting issues that may have originated down stream. Setting these metrics up can be no small task, but, the effort is well worth it when a problem crops up.

Anyone looking to work at scale in the DevOps / Cloud Infrastructure / Application Orchestration World would be wise to gain a good understanding of at least the basics of Prometheus and Grafana. Replicating the dashboard(s) you’ve done here would be a very good place to start.

---

<div class="post-metadata">

**Author:** ![jay](https://community.learnlinux.tv/user_avatar/community.learnlinux.tv/jay/32/649_2.png) [@jay](https://community.learnlinux.tv/u/jay)\
**Post date:** [December 11, 2021, 10:08pm UTC](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563/20 "2021-12-11T22:08:19Z")

</div>

That dashboard is great, how long did that take?

[Next page](https://community.learnlinux.tv/t/how-to-graph-the-resource-usage-of-servers/563.md?page=2)
