{"schemaVersion":"jobsearcher.job.v1","id":"738625d548022b9bd1ae07fa","url":"https://jobsearcher.com/jobs/738625d548022b9bd1ae07fa","canonicalUrl":"https://jobsearcher.com/jobs/738625d548022b9bd1ae07fa","title":"GPU Cluster Sysadmin","description":"Stanford Research Computing (https://) is a collaboration between University IT and the Vice Provost and Dean of Research. They operate HPC environments for researchers, provide one-time consultations on projects (software and pipelines, data management, physical building design and fit-out), and provide contract support for individual Labs, Departments, and Schools. They have four open positions: three hybrid and one (marked) onsite.\r\nPrincipal Storage Architect & Team Lead (Technical Manager) Lead the storage team and set direction for large storage environments Oak (file storage used by multiple clusters), Fir (fast scratch for Sherlock), and Elm (object storage on top of tape). Knowledge of Lustre, Infiniband, and PB-scale storage is important.\r\nStorage Architect or Storage Sysadmin Maintain and expand Oak, a 20+ Pebibyte Lustre storage environment used by the largest HPC clusters. Depending on experience, may also have responsibility for Elm (object storage on top of tape). Knowledge of Lustre, Infiniband, and PB-scale storage is important.\r\nGPU Cluster Sysadmin With Marlowe—an 1SU NVIDIA DGX H100 SuperPOD with DDN Intelliflash and DDN NFS storage—work on latest AI/ML/Deep Learning/LLM software and frameworks, keep the environment up-to-date, and work with NVIDIA/DDN when trouble occurs.\r\nHPC Hardware & Infra Sysadmin [ONSITE] System administrator to help run hardware/infrastructure portions of Sherlock, their largest HPC cluster. Sherlock includes Intel and AMD x86_64 servers, three Infiniband fabrics, and an Ethernet backbone. Responsible for maintaining, troubleshooting, and improving the hardware.\r\nRelocation incentive is provided if you do not live in the Bay Area\r\nfree transit passes depending on where you live\r\nparking costs apply if you drive on on-site days\r\nsome on-call around the holidays\r\n403(b) match\r\ngood healthcare\r\n30+ days off per year (holidays + vacation)\r\nBenefits are documented publicly at https://.\r\nQuestions can be asked by replying here or emailing the person (info in profile).\r\nJ-18808-Ljbffr","company":"Meanderx","rawCompany":"meanderx","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-09T01:09:55.733Z","occupations":[{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541513","title":"Computer Facilities Management Services","slug":"computer-facilities-management-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"GPU Cluster Sysadmin","description":"Stanford Research Computing (https://) is a collaboration between University IT and the Vice Provost and Dean of Research. They operate HPC environments for researchers, provide one-time consultations on projects (software and pipelines, data management, physical building design and fit-out), and provide contract support for individual Labs, Departments, and Schools. They have four open positions: three hybrid and one (marked) onsite.\r\nPrincipal Storage Architect & Team Lead (Technical Manager) Lead the storage team and set direction for large storage environments Oak (file storage used by multiple clusters), Fir (fast scratch for Sherlock), and Elm (object storage on top of tape). Knowledge of Lustre, Infiniband, and PB-scale storage is important.\r\nStorage Architect or Storage Sysadmin Maintain and expand Oak, a 20+ Pebibyte Lustre storage environment used by the largest HPC clusters. Depending on experience, may also have responsibility for Elm (object storage on top of tape). Knowledge of Lustre, Infiniband, and PB-scale storage is important.\r\nGPU Cluster Sysadmin With Marlowe—an 1SU NVIDIA DGX H100 SuperPOD with DDN Intelliflash and DDN NFS storage—work on latest AI/ML/Deep Learning/LLM software and frameworks, keep the environment up-to-date, and work with NVIDIA/DDN when trouble occurs.\r\nHPC Hardware & Infra Sysadmin [ONSITE] System administrator to help run hardware/infrastructure portions of Sherlock, their largest HPC cluster. Sherlock includes Intel and AMD x86_64 servers, three Infiniband fabrics, and an Ethernet backbone. Responsible for maintaining, troubleshooting, and improving the hardware.\r\nRelocation incentive is provided if you do not live in the Bay Area\r\nfree transit passes depending on where you live\r\nparking costs apply if you drive on on-site days\r\nsome on-call around the holidays\r\n403(b) match\r\ngood healthcare\r\n30+ days off per year (holidays + vacation)\r\nBenefits are documented publicly at https://.\r\nQuestions can be asked by replying here or emailing the person (info in profile).\r\nJ-18808-Ljbffr","datePosted":"2026-08-09T01:09:55.733Z","dateModified":"2026-08-09T01:09:55.733Z","hiringOrganization":{"@type":"Organization","name":"Meanderx","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"738625d548022b9bd1ae07fa"},"url":"https://jobsearcher.com/jobs/738625d548022b9bd1ae07fa"}}