JOBSEARCHER

Platform Reliability Engineer

RemotehunterDenver, COL6 LeadSeptember 17th, 2026
1. About Our Client:The organization is a technology consulting and software development company operating in the United States. It focuses on delivering cloud, artificial intelligence, data, and enterprise solutions. The company addresses the challenge of integrating advanced technologies into business operations, supporting clients with scalable and innovative software solutions.2. About the Opportunity:The Platform Reliability Engineer role is responsible for ensuring the availability, performance, and operational excellence of large-scale distributed systems in production environments. This position bridges development and operations by applying software engineering principles to infrastructure challenges, aiming to improve platform reliability while reducing operational effort. The role plays a critical part in maintaining and enhancing complex services to make reliability a fundamental engineering deliverable.3. Responsibilities:Ensure availability and performance of distributed systems in productionApply software engineering practices to infrastructure and operations problemsAutomate and operate complex services to reduce operational toilCollaborate across teams to improve platform reliabilityLead incident response and conduct post-incident reviews4. Requirements:Bachelor’s degree in Computer Science, Engineering, or related fieldMinimum of 5 years experience in SRE, DevOps, or production engineering supporting large-scale distributed systemsProficient in programming languages such as Python, Go, or Java for automation and toolingExtensive hands-on experience with Linux systems administration at scaleExperience operating Kubernetes and container-based workloads in productionKnowledge of observability tools like Prometheus, Grafana, OpenTelemetry, ELK/EFKExperience designing and managing CI/CD pipelines for infrastructure and applicationsUnderstanding of distributed system design, including consistency, partitioning, and failure modesStrong communication and documentation skillsPreferred:Experience with SLOs and error budgetsExposure to chaos engineering toolsHands-on experience with major cloud platforms (AWS, Azure, or GCP)Experience in capacity planning, performance engineering, or load testingFamiliarity with service mesh technologies5. Pay Range and Compensation Package:Salary range: $100,000–$150,000 annuallyEqual Opportunity Statement: Our client is an equal opportunity employer. They celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, or national origin.Note:RemoteHunter is not the Employer of Record (EOR) for this role. Our purpose in this opportunity is to connect exceptional candidates with leading employers. We help job seekers worldwide discover roles that match their goals and guide them to complete their full application directly through the hiring company’s career page or ATS.