JOBSEARCHER

Staff Cloud Reliability Engineer

Overview In this role you will strengthen cloud reliability by developing and integrating tooling for cloud-based systems. You will containerize applications, drive automation, and broaden observability across platforms. You’ll scope projects, support development and operations teams, and participate in on-call incident response. This role offers hands-on work with modern tech and a chance to shape reliability at scale in a fast-growing ad-tech environment. Compensation / Benefitsfully paid health insurancepaid parental leaveunlimited PTO ResponsibilitiesDevelop, configure, and deploy tools for cloud-based systems and servicesContainerize new and legacy applicationsMaintain awareness of new and exciting technologiesProvide LOE/scoping for projectsSupport development and operations teamsEnhance, modify or debug developer code as neededEnsure comprehensive observability across all aspects of application and system performanceParticipate in on-call rotation for timely response to incidents and maintenance Key requirements8+ years experience in a DevOps or SRE-related role3+ years Linux administration in a professional setting3+ years experience with one or more large cloud providers (AWS, Google)3+ years experience with serverless architecture (AWS Lambda, Google Cloud Functions)3+ years using Docker and or Kubernetes3+ years experience using Terraform - writing vanilla Terraform to instrumenting modulesAbility to create CI/CD pipelines GitHub ActionsProficiency in at least one of Python or GoLangExperience with SQL and Google BigQueryproblem solvingteam collaborationcommunicationLinux administrationAWS and/or Google Cloud Platformserverless architectures (Lambda, Functions)