Hi,
We have a EVPN VXLAN Setup using Spine/Leaf. Most of the Leafs until now have been using VSX. They perform really well, but for a new Rack we wanted something cheaper, and opted for the 6300 (24xSFP+). The connected servers are all vmware, and they do not use LACP LAGs. They just move over the MAC to the other NIC when the link goes down. The gateway is an active gateway.
I have opted to configure the two 6300 switches independently for the following reasons:
- We do not need multichassis lags
- Replacing a standalone switch is much easier then replacing a VSF Member (especially in Central)
- Upgrading VSF reboot's both members at the same time, and causes downtime. (I have seen some roadmap changing this in a future version, but upgrading both independently still seems easier)
The setup with two independent switches works well in normal circumstances. If I disconnect Cables from a server in 1 switch, all Mac addresses move instantly to the other switch and everything stays reachable. So far so good.
However when I reboot a switch the failover does not work. I see the MAC addresses of the server, and also the ARP entry pointing to the second switch, but still I can't ping servers that used to reside on Switch 1.
If I look at the bgp paths, I see that MAC and IP entries are present for both switches. So maybe the traffic is still sent to the old switch that is down. I can think of a couple of solutions but I'd like to hear some thoughts on this.
- Make them into a VSF stack and accept the hassle that comes with that.
- Use neighbor x.x.x.x fall-over bfd. Without bfd it would not work because the switches have a default route in the underlay. (Maybe I could get rid of those though)
What would be the cleanest solution for this according to you guys? Or is there another approach that I could take here?
------------------------------
Jelmer Hartman
------------------------------