lookuppierre
Member
- Joined:
- Aug 2024
- Posts:
- 113
- From:
- Lyon, France
Ow to say... I am using the CloudCone monitoring and is... how to say... acceptable. The «intelligent alerting» is not so intelligent. I drop the h on purpose sometimes, is habit. But the real solution I pind is very simple: hysteresis. You set threshold at 80% for alert ON, 60% for alert OFF. No flap. No spam. The Machine Learning is just... Oui, I say oui, is overkill for most.
Random Words because why not.
prix fixe infrastructure: €5/mo
uma
Member
99.99% or bust
- Joined:
- Jun 2024
- Posts:
- 324
- From:
- Dublin, IE
I've tested twelve monitoring solutions in the last three years. Uptime percentages: 99.97% average across my fleet. Alert fatigue is real. The ML-based ones promise the world and deliver 40% false positive rates in my experience. Status pages become meaningless when everything is "degraded."
What actually works: hysteresis, as mentioned. Also: require two consecutive failures before alerting. Exponential backoff on flapping checks. The simple math outperforms neural nets for infrastructure blips.
436 days. reboot is surrender.