Hello,
Thank you for the response.
Active Gateway is predominantly associated with Layer 3 configurations ( as per my understanding), particularly in scenarios involving core switching functionalities. In our current setup, we've integrated the topology into the client's network at Layer 2, utilizing static VXLAN for the purpose. This decision was driven by the specific requirements of our deployment, which did not necessitate Layer 3 capabilities at the edge of our network.
During our installation today, as outlined in the provided network diagram, we connected an Arista switch to the right-side Aruba pair using LAG LACP and MC-LAG configurations, respectively. This setup was then tested by connecting a laptop to the Arista switch and pinging a real Active Directory server, which was linked via LAG to the left side of our topology. The connection between these points was facilitated by Metro Ethernet links. Initially, the ping tests to the AD server were successful, indicating a healthy connection through the top link.
However, an issue arose when the top link was disconnected; the ping tests failed to continue, suggesting a failure in traffic rerouting through the remaining bottom link. This led me to consider several potential explanations:
- The WAN link failure might not have triggered a corresponding change in the VSX LAG state, causing the primary switch to retain traffic handling responsibilities without a viable delivery path.
- Traffic might be persistently directed towards the failed WAN link by the primary switch, which fails to reroute it through the secondary switch as expected.
- Considering the LAG setup on the Arista switch, I expected traffic from the laptop to utilize both links towards the Aruba pair. The primary switch's continued operation might be influencing traffic to favor it, raising questions about the load balancing behavior in this configuration.
I'm currently delving into these areas to pinpoint the exact cause of the issue and would greatly appreciate any further insights or suggestions you might have. Specifically, I'm interested in understanding if there are nuances in the MC-LAG or VSX configurations that could be influencing this behavior, or if there are recommended practices for such deployments that I might have overlooked.
Any advice or shared experiences with similar topologies would be incredibly valuable as I work towards resolving this challenge.
Original Message:
Sent: Jan 24, 2024 06:41 PM
From: ProbeRequest
Subject: Challenges with LACP, MCLAG, and VXLAN Configuration on Aruba 8100 in VSX Mode
Good details and well explained.
How have you troubleshot this so far?
Does VSX appear to be established and the inter-switch links show in tact and operational?
Are you using Active Gateway in this configuration?
If you established LACP between a Win Server and two (or more) interfaces on a single 8100 switch do you have a problem or does it only arise when the LAG is across two switches?