Woke up to 47 Nagios alerts screaming that my Contabo box was down. Fired off an angry ticket, got a calm reply that their end showed 99.99% uptime. Traced it back to my own check_command timeout being too aggressive after I "optimized" it last week. Embarrassing.
Anyone else burned by their own monitoring stack crying wolf? I'm running 6 boxes across 4 providers and this is the third false alarm this month. Starting to think my $3 VPS with better YABS disk scores than my $10 box is the only thing I configured right. 1874 MB/s on that bad boy.
What's your alert fatigue story?