4 Metrics for Measuring Your Service Level Agreements
Summary
To keep up with demand, one of our partners created a PRTG plugin for SLA monitoring and back in March of this year, my colleague Sascha wrote a blog post on it. To understand SLA monitoring a bit better, let’s dive into what these numbers are and what the difference between them is: The first metric measures how much time has elapsed before an error occurs. And if we are referring to something that cannot be salvaged after a failure, we call it “mean time to failure” (Oh, and in case you were wondering: yes, my source is Wikipedia). And if you still ask yourself the question of why it is so important to actually track your SLAs, heres another reason: Even if you have services that are running smoothly, having figures to prove this is more than just a nice-to-have. Another useful feature of PRTG in regards to monitoring how well you comply with your SLA is actually the possibility to set thresholds as required.