{"schemaVersion":"jobsearcher.job.v1","id":"e96e67a42f9b1a0b60715c8e","url":"https://jobsearcher.com/jobs/e96e67a42f9b1a0b60715c8e","canonicalUrl":"https://jobsearcher.com/jobs/e96e67a42f9b1a0b60715c8e","title":"Staff HPC Systems Architect","description":"Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.\n\nIf you'd like to build the world's best AI cloud, join us.\n\n*Note: This position requires presence in our San Jose, San Francisco, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.\n\nWhat You’ll Do\n\nArchitect and define scalable compute platforms optimized for AI/ML, simulation, and high-throughput workloads.\n\nDevelop compute system standards and design patterns to ensure consistency, performance, and maintainability across infrastructure.\n\nEvaluate emerging CPU, GPU, and accelerator technologies, owning architectural tradeoff decisions that impact compute density, power, cooling, and total cost.\n\nCollaborate with product and engineering teams to map workload requirements to compute platform capabilities across bare metal and cloud deployments.\n\nExperience converting ambiguous business or customer needs into measurable platform requirements, technical specifications, acceptance criteria, and architecture decisions.\n\nDefine compute platform roadmaps and architectural reference designs that guide hardware selection, firmware baselines, rack-level, and cluster design.\n\nAct as a technical lead during new platform introductions, guiding validation and performance characterization efforts.\n\nMentor systems engineers and cross-functional stakeholders on compute performance tuning, sizing, and architectural decisions.\n\nYou\n\nProven experience (7+ years) architecting large-scale 10k-100k+ GPU HPC or cloud compute platforms.\n\nDeep knowledge of CPU/GPU architectures, memory hierarchies, and accelerator topologies.\n\nExperience designing systems around high-bandwidth, low-latency fabrics (NVLink, InfiniBand, and RoCE).\n\nStrong understanding of system performance tuning, resource scheduling, thermal and power optimization, and compute lifecycle management.\n\nComfortable working across hardware and software boundaries, especially at the intersection of compute architecture, OS behavior, and orchestration layers.\n\nSkilled at balancing architectural tradeoffs for density, power efficiency, cooling, and performance.\n\nStrong analytical and communication skills, with a track record of influencing technical strategy across teams.\n\nStrong ownership and can do attitude, self-starter who feels comfortable working in ambiguity.\n\nNice to Have\n\nHands-on experience with AI/ML workloads and their compute performance characteristics.\n\nFamiliarity with orchestration tools used in HPC. (Slurm, Kubernetes, etc)\n\nExperience with virtualization technologies, specifically GPU virtualization.\n\nExposure to hardware validation, vendor collaboration, and long-term OEM roadmap alignment.\n\nBackground in compute telemetry, real-time performance profiling, or large-scale A/B infrastructure testing.\n\nSalary Range Information\n\nThe annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.\n\nAbout Lambda\n\nFounded in 2012, with 500+ employees, and growing fast\n\nOur investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove\n\nWe have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG\n\nOur values are publicly available: https://lambda.ai/careers\n\nWe offer generous cash & equity compensation\n\nHealth, dental, and vision coverage for you and your dependents\n\nWellness and commuter stipends for select roles\n\n401k Plan with 2% company match (USA employees)\n\nFlexible paid time off plan that we all actually use\n\nEqual Opportunity Employer\n\nLambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.\n\nCompensation Range: $314K - $465K","company":"Lambda","rawCompany":"lambda","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-16T11:33:25.719Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff HPC Systems Architect","description":"Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.\n\nIf you'd like to build the world's best AI cloud, join us.\n\n*Note: This position requires presence in our San Jose, San Francisco, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.\n\nWhat You’ll Do\n\nArchitect and define scalable compute platforms optimized for AI/ML, simulation, and high-throughput workloads.\n\nDevelop compute system standards and design patterns to ensure consistency, performance, and maintainability across infrastructure.\n\nEvaluate emerging CPU, GPU, and accelerator technologies, owning architectural tradeoff decisions that impact compute density, power, cooling, and total cost.\n\nCollaborate with product and engineering teams to map workload requirements to compute platform capabilities across bare metal and cloud deployments.\n\nExperience converting ambiguous business or customer needs into measurable platform requirements, technical specifications, acceptance criteria, and architecture decisions.\n\nDefine compute platform roadmaps and architectural reference designs that guide hardware selection, firmware baselines, rack-level, and cluster design.\n\nAct as a technical lead during new platform introductions, guiding validation and performance characterization efforts.\n\nMentor systems engineers and cross-functional stakeholders on compute performance tuning, sizing, and architectural decisions.\n\nYou\n\nProven experience (7+ years) architecting large-scale 10k-100k+ GPU HPC or cloud compute platforms.\n\nDeep knowledge of CPU/GPU architectures, memory hierarchies, and accelerator topologies.\n\nExperience designing systems around high-bandwidth, low-latency fabrics (NVLink, InfiniBand, and RoCE).\n\nStrong understanding of system performance tuning, resource scheduling, thermal and power optimization, and compute lifecycle management.\n\nComfortable working across hardware and software boundaries, especially at the intersection of compute architecture, OS behavior, and orchestration layers.\n\nSkilled at balancing architectural tradeoffs for density, power efficiency, cooling, and performance.\n\nStrong analytical and communication skills, with a track record of influencing technical strategy across teams.\n\nStrong ownership and can do attitude, self-starter who feels comfortable working in ambiguity.\n\nNice to Have\n\nHands-on experience with AI/ML workloads and their compute performance characteristics.\n\nFamiliarity with orchestration tools used in HPC. (Slurm, Kubernetes, etc)\n\nExperience with virtualization technologies, specifically GPU virtualization.\n\nExposure to hardware validation, vendor collaboration, and long-term OEM roadmap alignment.\n\nBackground in compute telemetry, real-time performance profiling, or large-scale A/B infrastructure testing.\n\nSalary Range Information\n\nThe annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.\n\nAbout Lambda\n\nFounded in 2012, with 500+ employees, and growing fast\n\nOur investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove\n\nWe have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG\n\nOur values are publicly available: https://lambda.ai/careers\n\nWe offer generous cash & equity compensation\n\nHealth, dental, and vision coverage for you and your dependents\n\nWellness and commuter stipends for select roles\n\n401k Plan with 2% company match (USA employees)\n\nFlexible paid time off plan that we all actually use\n\nEqual Opportunity Employer\n\nLambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.\n\nCompensation Range: $314K - $465K","datePosted":"2026-09-16T11:33:25.719Z","dateModified":"2026-09-16T11:33:25.719Z","hiringOrganization":{"@type":"Organization","name":"Lambda","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"e96e67a42f9b1a0b60715c8e"},"url":"https://jobsearcher.com/jobs/e96e67a42f9b1a0b60715c8e"}}