GeorgeNmp
Member
OP
AS64512
- Joined:
- May 2024
- Posts:
- 218
- From:
- Ashburn, US
I've been network-booting our edge nodes for three years. Root on iSCSI over dedicated VLAN, local NVMe reserved for container ephemeral storage only. The economics:
- 480GB boot NVMe: $8/month from most providers
- 1Gbps unmetered iSCSI target: $3/month, shared across 8 nodes
- PXE boot time with iPXE: 12 seconds to kernel, 4 seconds to systemd
The "local boot" requirement is a holdover from bare-metal thinking. With proper LACP to your TOR and redundant targets, network boot exceeds SATA SSD reliability. My current fleet: 47 nodes, zero boot-related outages in 36 months.
Yes, this excludes database workloads. I know. Don't @ me yet.
iBGP, eBGP, don't care, just peer
uma
Member
99.99% or bust
- Joined:
- Jun 2024
- Posts:
- 324
- From:
- Dublin, IE
Network boot. Single switch failure. 47 nodes dead. Ticket open. Waiting.
436 days. reboot is surrender.
GeorgeNmp
Member
OP
AS64512
- Joined:
- May 2024
- Posts:
- 218
- From:
- Ashburn, US
Peering back @ned69
1Gbps for 8 nodes = 125MB/s each
Sequential worst-case, sure. Boot traffic is not sequential. IPXE loads ~50MB kernel+initrd, then systemd hits the target for rootfs. Average sustained during boot: 18MB/s per node. Post-boot: near zero for stateless containers.
Single db query spill to disk = saturation
Did you miss the part where I said this excludes database workloads? Reading comprehension.
Transit pricing for that ispsci vlan
Dedicated VLAN on same TOR = no transit. Layer 2. You know how switches work, you're from Oslo.
V4 DHCP: stateless, 60-second lease, option 66/67. Not hard.
iBGP, eBGP, don't care, just peer