{"schemaVersion":"jobsearcher.job.v1","id":"e71284433c0e494e2c627c8a","url":"https://jobsearcher.com/jobs/e71284433c0e494e2c627c8a","canonicalUrl":"https://jobsearcher.com/jobs/e71284433c0e494e2c627c8a","title":"Data Center Operations Systems Engineer III","description":"Data Center Operations Systems Engineer III at lambda. About the role Join our team to maintain and expand the physical infrastructure that powers advanced AI. This role involves hands-on work in our data centers, ensuring the continuous operation and deployment of high-performance computing systems. You will contribute to building the world's best AI cloud.\r\nKey facts Location: Vernon, CA\r\nData Center Engagement: Full-time, on-site, shift work\r\nCompensation: $115K - $153K annually\r\nTeam: Data Center Business\r\nWhat you'll do Install and configure new server, storage, and network hardware, ensuring correct racking, labeling, and cabling.\r\nDiagnose and resolve hardware and software issues within advanced GPU and networking environments.\r\nMaintain accurate records of data center layouts and network configurations using DCIM software.\r\nCoordinate with supply chain and manufacturing teams for timely system deployments and large-scale project execution.\r\nManage inventory of parts, tracking equipment from delivery through deployment and handoff.\r\nCollaborate with hardware support to resolve complex incidents, document solutions, and share knowledge across operations.\r\nWork with the RMA team to process faulty parts returns and order replacements.\r\nAdhere to installation standards for placement, labeling, and cabling to ensure consistency across data centers.\r\nRequirements Proven experience with critical data center infrastructure, including power distribution, airflow, environmental monitoring, capacity planning, DCIM software, and structured cabling.\r\nFamiliarity with carrier DIA circuit testing and activation, along with fiber testing and troubleshooting.\r\nBasic understanding of cable optics and their various applications.\r\nSolid grasp of single and three-phase power principles.\r\nKnowledge of PDU balancing and its importance.\r\nExperience with multiple cable media types and their uses.\r\nUnderstanding of cold and hot aisle containment strategies.\r\nStrong knowledge of server hardware and the boot process.\r\nAbility to create, collaborate on, and refine complex maintenance procedures.\r\nExperience aligning operational capabilities with company goals by working with product management and support teams.\r\nSkill in translating business priorities into technical and operational requirements.\r\nCapability to support cross-functional projects where infrastructure is key.\r\nProactive approach and willingness to train junior staff on best practices.\r\nWillingness to travel 25%-30% for new data center setups as needed.\r\nNice to have Over 5 years of experience with critical data center infrastructure systems, such as power distribution, airflow management, environmental monitoring, capacity planning, DCIM software, structured cabling, and cable management.\r\nExperience with or knowledge of network topology, configurations, and 400gb Infiniband architectures.\r\nExperience with or knowledge of DDP or SCM cluster storage systems.\r\nOver 5 years of experience working with and reporting from ticketing systems like JIRA and Zendesk.\r\nAdvanced Linux administration skills.\r\nExperience with High Performance Compute GPU systems (air or water cooled), especially Nvidia NVL72.\r\nSkills & tools DCIM software\r\nTicketing systems (JIRA, Zendesk)\r\nLinux\r\nInfiniband\r\nDDP/SCM cluster storage\r\nNvidia NVL72\r\nPractical notes This role requires 5 days of on-site presence in our Los Angeles, CA Data Center. We offer generous cash and equity compensation, health, dental, and vision coverage for you and your dependents, wellness and commuter stipends for select roles, a 401k plan with a 2% company match (for USA employees), and a flexible paid time off plan.\r\nJ-18808-Ljbffr","company":"Lambda","rawCompany":"lambda","city":"Louisiana","state":"MO","isRemote":false,"isActive":false,"createdAt":"2026-08-12T00:51:01.787Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"15-1231.00","title":"Computer Network Support Specialists","slug":"computer-network-support-specialists"}],"industries":[{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541513","title":"Computer Facilities Management Services","slug":"computer-facilities-management-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Center Operations Systems Engineer III","description":"Data Center Operations Systems Engineer III at lambda. About the role Join our team to maintain and expand the physical infrastructure that powers advanced AI. This role involves hands-on work in our data centers, ensuring the continuous operation and deployment of high-performance computing systems. You will contribute to building the world's best AI cloud.\r\nKey facts Location: Vernon, CA\r\nData Center Engagement: Full-time, on-site, shift work\r\nCompensation: $115K - $153K annually\r\nTeam: Data Center Business\r\nWhat you'll do Install and configure new server, storage, and network hardware, ensuring correct racking, labeling, and cabling.\r\nDiagnose and resolve hardware and software issues within advanced GPU and networking environments.\r\nMaintain accurate records of data center layouts and network configurations using DCIM software.\r\nCoordinate with supply chain and manufacturing teams for timely system deployments and large-scale project execution.\r\nManage inventory of parts, tracking equipment from delivery through deployment and handoff.\r\nCollaborate with hardware support to resolve complex incidents, document solutions, and share knowledge across operations.\r\nWork with the RMA team to process faulty parts returns and order replacements.\r\nAdhere to installation standards for placement, labeling, and cabling to ensure consistency across data centers.\r\nRequirements Proven experience with critical data center infrastructure, including power distribution, airflow, environmental monitoring, capacity planning, DCIM software, and structured cabling.\r\nFamiliarity with carrier DIA circuit testing and activation, along with fiber testing and troubleshooting.\r\nBasic understanding of cable optics and their various applications.\r\nSolid grasp of single and three-phase power principles.\r\nKnowledge of PDU balancing and its importance.\r\nExperience with multiple cable media types and their uses.\r\nUnderstanding of cold and hot aisle containment strategies.\r\nStrong knowledge of server hardware and the boot process.\r\nAbility to create, collaborate on, and refine complex maintenance procedures.\r\nExperience aligning operational capabilities with company goals by working with product management and support teams.\r\nSkill in translating business priorities into technical and operational requirements.\r\nCapability to support cross-functional projects where infrastructure is key.\r\nProactive approach and willingness to train junior staff on best practices.\r\nWillingness to travel 25%-30% for new data center setups as needed.\r\nNice to have Over 5 years of experience with critical data center infrastructure systems, such as power distribution, airflow management, environmental monitoring, capacity planning, DCIM software, structured cabling, and cable management.\r\nExperience with or knowledge of network topology, configurations, and 400gb Infiniband architectures.\r\nExperience with or knowledge of DDP or SCM cluster storage systems.\r\nOver 5 years of experience working with and reporting from ticketing systems like JIRA and Zendesk.\r\nAdvanced Linux administration skills.\r\nExperience with High Performance Compute GPU systems (air or water cooled), especially Nvidia NVL72.\r\nSkills & tools DCIM software\r\nTicketing systems (JIRA, Zendesk)\r\nLinux\r\nInfiniband\r\nDDP/SCM cluster storage\r\nNvidia NVL72\r\nPractical notes This role requires 5 days of on-site presence in our Los Angeles, CA Data Center. We offer generous cash and equity compensation, health, dental, and vision coverage for you and your dependents, wellness and commuter stipends for select roles, a 401k plan with a 2% company match (for USA employees), and a flexible paid time off plan.\r\nJ-18808-Ljbffr","datePosted":"2026-08-12T00:51:01.787Z","dateModified":"2026-08-12T00:51:01.787Z","hiringOrganization":{"@type":"Organization","name":"Lambda","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Louisiana","addressRegion":"MO","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"e71284433c0e494e2c627c8a"},"url":"https://jobsearcher.com/jobs/e71284433c0e494e2c627c8a"}}