Experience supporting GPU-based networking environments in production.
Ability to troubleshoot GPU cluster connectivity and performance issues.
Experience supporting deployment and logical turn-up activities.
Working familiarity with Arista EOS and/or NVIDIA/Cumulus networking platforms; exposure to RoCE/InfiniBand environments a plus.
Solid troubleshooting skills across L2/L3 protocols (VLANs, BGP, OSPF) and monitoring tooling.
Experience following ITIL-aligned incident, problem, and change management processes.
External Communities Job Description
seeking experienced GPU Infrastructure and Networking Consultants to support the rollout of a new GPU-as-a-Service platform.
Resources will review existing GPU architecture, provide deployment guidance, support production operations, and assist with long-term infrastructure strategy.
Candidates should possess deep expertise in GPU networking, NVIDIA-based environments, data center architectures, AI/HPC infrastructure, and operational troubleshooting.