{"schemaVersion":"jobsearcher.job.v1","id":"034201543036bfb36a89c8ea","url":"https://jobsearcher.com/jobs/034201543036bfb36a89c8ea","canonicalUrl":"https://jobsearcher.com/jobs/034201543036bfb36a89c8ea","title":"GPU Network Engineer","description":"GPU Network EngineerBased in Sunnyvale, CA, we are an innovative energy infrastructure company that develops cutting-edge data centers to drive the expansion of sustainable energy assets.Due to growth, we are looking for a hands-on GPU Network Engineer to design, build, and operate the high-speed fabrics that connect our GPU clusters.Must be local to Sunnyvale, CA or willing to relocate for this position. (4 days working on-site)What You Will Be Doing:Design and deploy east-west network fabrics for GPU-to-GPU, rack-to-rack, and cluster-to-cluster communication at scale.Build and operate InfiniBand (NDR/XDR) and RoCEv2 interconnect fabrics as the primary transport for GPU cluster workloads.Implement CLOS/ECMP architectures optimized for high-bandwidth, low-latency, lossless data movement across GPU clusters.Must Have Skills:2+ years in data center or HPC networking, with direct experience in GPU or AI cluster environments.Hands-on experience with GPU cluster networking: GPU-to-GPU, rack-to-rack, and cluster-to-cluster fabric design and operations.Deep working knowledge of InfiniBand (NDR/XDR) and RoCEv2 as primary GPU cluster interconnects -- this is the core of the role.Experience with multi-host networking for bare metal, KVM, and Kubernetes.Proficiency with Netbox, Netconf, and IaC tools (Ansible, Terraform).If hired, you will be rewarded with an offer that includes:200k-250k salaryPerformance BonusRSUsHealth Benefits401kPTOOpportunity to work with cutting-edge AI and HPC infrastructureCollaborative, fast-paced environmentFor this position you must be currently authorized to work in the United States. We do not sponsor for this position.","company":"Digipower X","rawCompany":"digipower x","city":"San Jose","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-19T13:12:41.282Z","occupations":[{"code":"15-1241.00","title":"Computer Network Architects","slug":"computer-network-architects"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541513","title":"Computer Facilities Management Services","slug":"computer-facilities-management-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"GPU Network Engineer","description":"GPU Network EngineerBased in Sunnyvale, CA, we are an innovative energy infrastructure company that develops cutting-edge data centers to drive the expansion of sustainable energy assets.Due to growth, we are looking for a hands-on GPU Network Engineer to design, build, and operate the high-speed fabrics that connect our GPU clusters.Must be local to Sunnyvale, CA or willing to relocate for this position. (4 days working on-site)What You Will Be Doing:Design and deploy east-west network fabrics for GPU-to-GPU, rack-to-rack, and cluster-to-cluster communication at scale.Build and operate InfiniBand (NDR/XDR) and RoCEv2 interconnect fabrics as the primary transport for GPU cluster workloads.Implement CLOS/ECMP architectures optimized for high-bandwidth, low-latency, lossless data movement across GPU clusters.Must Have Skills:2+ years in data center or HPC networking, with direct experience in GPU or AI cluster environments.Hands-on experience with GPU cluster networking: GPU-to-GPU, rack-to-rack, and cluster-to-cluster fabric design and operations.Deep working knowledge of InfiniBand (NDR/XDR) and RoCEv2 as primary GPU cluster interconnects -- this is the core of the role.Experience with multi-host networking for bare metal, KVM, and Kubernetes.Proficiency with Netbox, Netconf, and IaC tools (Ansible, Terraform).If hired, you will be rewarded with an offer that includes:200k-250k salaryPerformance BonusRSUsHealth Benefits401kPTOOpportunity to work with cutting-edge AI and HPC infrastructureCollaborative, fast-paced environmentFor this position you must be currently authorized to work in the United States. We do not sponsor for this position.","datePosted":"2026-09-19T13:12:41.282Z","dateModified":"2026-09-19T13:12:41.282Z","hiringOrganization":{"@type":"Organization","name":"Digipower X","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"San Jose","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"034201543036bfb36a89c8ea"},"url":"https://jobsearcher.com/jobs/034201543036bfb36a89c8ea"}}