Last month I was running postgres on "cheap" non-ECC box at Contabo, going to tell what happened. Silent bit flip, corrupted whole database, lost 3 days of data. I thought I was smart saving $20/month.
dmidecode shows non-ECC:
Memory Device
Array Handle: 0x1000
Error Information Handle: Not Provided
Total Width: 64 bits
Data Width: 64 bits
Size: 32 GB
Form Factor: DIMM
Set: None
Locator: DIMM_A1
Bank Locator: BANK 0
Type: DDR4
Type Detail: Synchronous
Speed: 3200 MT/s
Manufacturer: Not Specified
Serial Number: Not Specified
Asset Tag: Not Specified
Part Number: Not Specified
Rank: 2
Configured Memory Speed: 3200 MT/s
Minimum Voltage: 1.2 V
Maximum Voltage: 1.2 V
Configured Voltage: 1.2 V
Memory Technology: DRAM
Memory Operating Mode Capability: Volatile memory
Firmware Version: Not Specified
Module Manufacturer ID: Bank 1, Hex 0xCE
Module Product ID: Unknown
Memory Subsystem Controller Manufacturer ID: Unknown
Memory Subsystem Controller Product ID: Unknown
Non-Volatile Size: None
Volatile Size: 32 GB
Cache Size: None
Logical Size: NoneKernel log caught it too late:
[Hardware Error] Machine check events logged
[Hardware Error] Corrected error, no action requiredExcept it was NOT corrected. Migrated to ECC, sleep better now. That $20 savings cost me weekend of restore from Bob level disaster.