I run a small anycast setup across three regions for a DNS resolver project. Two nodes in my country, one in Europe. Past week I have been seeing intermittent timeouts from users in SEA but my monitoring shows 100% uptime.
Using OVHcloud for the European node, Contabo for local. Monitoring is simple ICMP from a fourth box to the anycast IP. The flapping seems to happen during peak hours only. Traceroutes from affected users sometimes land on the European node instead of local.
I checked BGP sessions; all stable. No withdrawals in logs. Where should I look next? Is my monitoring just checking the wrong instance entirely?