JOBSEARCHER

Senior Site Reliability Engineer

I'm partnering with a fast-growing Healthcare AI & Automation company that's looking to add a Site Reliability Engineer (SRE) as well as a Sr DevOps to their team. This is a highly visible, hands-on role where you'll help scale and support production systems that power mission-critical solutions for enterprise customers.What You'll Be DoingOwn the reliability, availability, and performance of production systems.Build automation, tooling, and operational workflows using Python.Partner closely with Backend, Data, and ML Engineering teams to improve platform scalability and resilience.Lead incident response efforts, root cause analysis, and post-mortem reviews.Establish and improve observability, monitoring, alerting, and SLOs.Help define and implement SRE best practices as the company continues to scale.Support and optimize distributed cloud infrastructure in a high-growth environment.What We're Looking For7+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or related roles.Strong hands-on coding experience with Python.Deep understanding of Incident ManagementObservability & MonitoringPerformance TuningReliability EngineeringSLOs/SLIsAutomationExperience working in fast-paced environments with evolving requirements.Strong collaboration and communication skills.Tech StackAWS (ECS, Lambda)TerraformDatadogGrafanaPostgreSQLModern CI/CD pipelinesPythonCompensation & BenefitsBase Salary: Up to $180KStaff-Level Compensation: Up to $220KEquityUnlimited PTOComprehensive Benefits PackageFlexible Hybrid Schedule (Onsite Monday & Wednesday)This is an excellent opportunity for an engineer who enjoys ownership, solving complex reliability challenges, and helping build scalable infrastructure at a company where their impact will be felt immediately.