Software Engineering Manager II, Site Reliability Engineering, Data Cloud SRE
Overview
As an SRE lead, you drive the reliability and performance of key Google services by shaping and mentoring a software/systems engineering team. You own end-to-end uptime, automate incident response, and scale systems to meet user demand. You work cross-functionally to reduce outages and improve automation, all while upholding a culture of learning, collaboration, and fearless experimentation.
Compensation / Benefitsbonus target 20%equitybenefitscareer growth opportunities
ResponsibilitiesLead a team of software/systems engineers on uptime-focused projectsOwn end-to-end availability and performance of core servicesBuild automation to prevent recurring issues and automate responsesManage on-call rotations across continents using a follow-the-sun modelDesign, develop, and deliver software to improve availability, scalability, latency, and efficiency of Google services
Key requirementsBachelor’s degree in Computer Science, related field, or equivalent practical experience8 years of software development experience in one or more programming languages3 years of experience managing people or teams3 years of experience designing, analyzing, and troubleshooting distributed systemsExperience in Artificial Intelligence or Machine Learningleadership and mentorshipcollaboration and cross-functional teamworkproblem solving and analytical thinkingdistributed systems design and troubleshootingsite reliability engineering practicesautomation and tooling development