Ngl I've been tracking my support tickets across three providers for my inference cluster. Weird pattern: the "fast" responses (under 15 min) are wrong like 70% of the time. The slow ones (2-4 hrs) actually fix stuff. Is this just me? I got VRAM errors, network topology questions, CUDA version mismatches. Fast answers were always "restart it" or "upgrade your plan" anyone else log this?
CUDA cores are my love language