{"schemaVersion":"jobsearcher.job.v1","id":"d58fd480d06aac136d4a26d3","url":"https://jobsearcher.com/jobs/d58fd480d06aac136d4a26d3","canonicalUrl":"https://jobsearcher.com/jobs/d58fd480d06aac136d4a26d3","title":"Staff HPC Network Architect","description":"Overview\nAs a Network Architect at Lambda, you will design high-performance data center networks powering AI workloads. You will define topology for large GPU clusters and multi-tenant environments, selecting cutting-edge networking tech to meet ultra-low latency needs. You’ll shape standards, roadmaps, and end-to-end data flow with cross-functional teams. This role offers impact at scale, enabling reliable, low-latency AI compute across hybrid environments. Join a mission-driven team building the world's best AI cloud.\n\nCompensation / Benefitsgenerous cash & equity compensationhealth, dental, and vision coveragewellness and commuter stipends401k plan with 2% company matchflexible paid time offhybrid work arrangement\nResponsibilitiesArchitect and optimize high-performance networks for cloud platforms with a focus on low latency and high bandwidthDefine topology and patterns for large-scale GPU clusters, storage backends, and multi-tenant environmentsEvaluate and select next-generation networking technologies (InfiniBand NDR/XDR, RoCE, 400G-1.6T Ethernet, etc.) to meet AI workloadsDevelop and maintain network architecture standards, reference designs, and scalability roadmaps for multi-site/hybrid environmentsCollaborate with compute and storage teams to ensure seamless data flow and fault toleranceLead network automation initiatives for provisioning, telemetry, and operational visibilityMentor engineers and cross-functional teams on advanced network concepts and best practicesProvide expert guidance on SDN, APIs, automation, and network telemetryDesign for resilience with redundancy, failover, and high availabilityCommunicate effectively to influence technical decisions across diverse teamsDemonstrate ownership and a proactive, can-do attitude in ambiguous settings\nKey requirements7+ years of experience architecting high-performance data center networks for HPC, AI/ML, or large-scale cloud infrastructureDeep expertise with InfiniBand (HDR/NDR) and advanced Ethernet fabrics (RoCE, RDMA)Strong understanding of data center switching architectures, congestion control (PFC, ECN), QoS, and VXLAN/EVPNExperience with low-latency, high-throughput data paths including GPU-to-GPU and storage traffic optimizationExpertise in BGP-based fabric design (eBGP underlays, MP-BGP EVPN, ECMP, route policy, convergence)Experience with SDN, APIs, automation, and network telemetryStrong optical networking knowledge (transceivers, fiber types, WDM) and fault-tolerant designExcellent communication and leadership skills; capable of cross-team influenceOwnership mindset; comfortable operating in ambiguityexcellent communicationleadershipownershipInfiniBand HDR/NDRRoCERDMA","company":"Lambda Labs","rawCompany":"lambda labs","city":"San Jose","state":"CA","isRemote":false,"isActive":true,"createdAt":"2026-09-16T03:15:25.427Z","occupations":[{"code":"15-1241.00","title":"Computer Network Architects","slug":"computer-network-architects"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541513","title":"Computer Facilities Management Services","slug":"computer-facilities-management-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff HPC Network Architect","description":"Overview\nAs a Network Architect at Lambda, you will design high-performance data center networks powering AI workloads. You will define topology for large GPU clusters and multi-tenant environments, selecting cutting-edge networking tech to meet ultra-low latency needs. You’ll shape standards, roadmaps, and end-to-end data flow with cross-functional teams. This role offers impact at scale, enabling reliable, low-latency AI compute across hybrid environments. Join a mission-driven team building the world's best AI cloud.\n\nCompensation / Benefitsgenerous cash & equity compensationhealth, dental, and vision coveragewellness and commuter stipends401k plan with 2% company matchflexible paid time offhybrid work arrangement\nResponsibilitiesArchitect and optimize high-performance networks for cloud platforms with a focus on low latency and high bandwidthDefine topology and patterns for large-scale GPU clusters, storage backends, and multi-tenant environmentsEvaluate and select next-generation networking technologies (InfiniBand NDR/XDR, RoCE, 400G-1.6T Ethernet, etc.) to meet AI workloadsDevelop and maintain network architecture standards, reference designs, and scalability roadmaps for multi-site/hybrid environmentsCollaborate with compute and storage teams to ensure seamless data flow and fault toleranceLead network automation initiatives for provisioning, telemetry, and operational visibilityMentor engineers and cross-functional teams on advanced network concepts and best practicesProvide expert guidance on SDN, APIs, automation, and network telemetryDesign for resilience with redundancy, failover, and high availabilityCommunicate effectively to influence technical decisions across diverse teamsDemonstrate ownership and a proactive, can-do attitude in ambiguous settings\nKey requirements7+ years of experience architecting high-performance data center networks for HPC, AI/ML, or large-scale cloud infrastructureDeep expertise with InfiniBand (HDR/NDR) and advanced Ethernet fabrics (RoCE, RDMA)Strong understanding of data center switching architectures, congestion control (PFC, ECN), QoS, and VXLAN/EVPNExperience with low-latency, high-throughput data paths including GPU-to-GPU and storage traffic optimizationExpertise in BGP-based fabric design (eBGP underlays, MP-BGP EVPN, ECMP, route policy, convergence)Experience with SDN, APIs, automation, and network telemetryStrong optical networking knowledge (transceivers, fiber types, WDM) and fault-tolerant designExcellent communication and leadership skills; capable of cross-team influenceOwnership mindset; comfortable operating in ambiguityexcellent communicationleadershipownershipInfiniBand HDR/NDRRoCERDMA","datePosted":"2026-09-16T03:15:25.427Z","dateModified":"2026-09-16T03:15:25.427Z","hiringOrganization":{"@type":"Organization","name":"Lambda Labs","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"San Jose","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"d58fd480d06aac136d4a26d3"},"url":"https://jobsearcher.com/jobs/d58fd480d06aac136d4a26d3"}}