Job Overview
We are seeking an experienced Network Consulting Engineer (NCE) specializing in VXLAN EVPN, AI Data Center Networking, and high-performance computing (HPC) environments. The successful candidate will be responsible for designing, implementing, and optimizing next-generation data center networks that support AI/ML workloads, GPU clusters, distributed computing, and hyperscale infrastructure.
This role requires deep expertise in modern data center networking technologies, including VXLAN EVPN, BGP, RoCEv2, RDMA, Cisco Nexus, Arista EOS, and fabric-based architectures. The ideal candidate will have hands-on experience building highly available, low-latency networks optimized for AI training, inference, and large-scale compute environments.
Key Responsibilities
- Design, deploy, and support VXLAN EVPN-based data center fabrics for AI, HPC, and cloud-scale environments.
- Architect and implement scalable spine-leaf network designs using platforms such as Cisco Nexus 9000, Cisco ACI, and Arista EOS.
- Configure and optimize Layer 2/Layer 3 segmentation, workload mobility, and network virtualization.
- Design and tune networks supporting GPU clusters, distributed AI training, and high-performance workloads.
- Optimize network performance for RoCEv2, RDMA, GPUDirect, and low-latency traffic flows.
- Implement and troubleshoot routing protocols including BGP, OSPF, IS-IS, and multicast.
- Deploy and manage data center fabric integrations with virtualization and container platforms including:
- VMware NSX
- KVM
- Hyper-V
- Kubernetes-based AI platforms
- Collaborate with compute, storage, cloud, and DevOps teams to ensure end-to-end infrastructure performance.
- Support network automation initiatives using Python, Ansible, Terraform, and Infrastructure-as-Code (IaC) methodologies.
- Perform network performance testing, benchmarking, capacity planning, and optimization for evolving AI workloads.
- Provide technical leadership, troubleshooting expertise, and consulting guidance to enterprise customers and internal teams.
Required Skills & Experience
- Proven experience designing, deploying, and operating VXLAN EVPN-based data center networks.
- Strong hands-on experience with:
- Cisco Nexus 9000 platforms
- Cisco ACI
- Arista EOS
- Spine-leaf architectures
- Advanced knowledge of data center routing and switching technologies:
- BGP
- OSPF
- IS-IS
- Multicast
- Network virtualization
- Deep understanding of AI and HPC networking requirements, including:
- RoCEv2
- RDMA
- Data Center Bridging (DCB)
- Priority Flow Control (PFC)
- Congestion management and latency optimization
- Experience supporting GPU-accelerated infrastructure and distributed computing environments.
- Strong troubleshooting skills across network control plane, data plane, and performance issues.
- Experience with network automation and orchestration tools:
- Python
- Ansible
- Terraform
- Understanding of hybrid cloud and multi-cloud networking architectures.
Preferred Qualifications & Certifications
- CCIE Data Center certification preferred.
- CCNP Data Center or equivalent advanced networking certification.
- Cisco certifications such as:
- Cisco Certified Specialist – Data Center Core
- Cisco Certified Specialist – Enterprise Core
- Experience with NVIDIA networking technologies, including:
- NVIDIA Spectrum switches
- NVIDIA Cumulus Linux
- Arista certifications such as Arista ACE are highly desirable.
- Experience designing and supporting AI infrastructure, hyperscale data centers, or HPC environments.