{"schemaVersion":"jobsearcher.job.v1","id":"ca38bfcc2a76614469e1be7c","url":"https://jobsearcher.com/jobs/ca38bfcc2a76614469e1be7c","canonicalUrl":"https://jobsearcher.com/jobs/ca38bfcc2a76614469e1be7c","title":"Site Reliability Engineer","description":"We are seeking a hands‑on Site Reliability Engineer to join our team. You will bridge the gap between software development and infrastructure operations, treating operational challenges as engineering problems. By leveraging automation, designing resilient distributed systems, and championing observability, you will ensure that our customers can secure and manage their access control systems without friction or failure.\n\nKey Responsibilities\n\nInfrastructure & Automation: Design, build, and maintain scalable, secure multi‑tenant cloud infrastructure using Infrastructure as Code (IaC) principles.\n\nUptime & Reliability: Own the availability, latency, performance, and capacity planning of the pdk.io platform and its supporting backend microservices.\n\nObservability: Develop and manage robust monitoring, logging, and alerting systems to gain deep visibility into cloud infrastructure, API health, and IoT endpoint performance.\n\nIncident Response: Participate in a collaborative on‑call rotation. Lead rapid incident response mitigation and drive rigorous, blameless post‑mortems to ensure long‑term system resilience.\n\nCI/CD Pipeline Management: Optimize and secure automated deployment pipelines to enable developers to ship code to production safely and efficiently.\n\nCross‑Functional Collaboration: Partner closely with backend developers and hardware engineering teams to define Service Level Indicators (SLIs), Service Level Objectives (SLOs), and manage error budgets.\n\nContribute to technical documentation and knowledge sharing\n\nTooling\n\nCI automation with GitHub Actions and Argo Workflows\n\nIaC with OpenTofu & Terragrunt at the core\n\nObservability stack powered by Prometheus and Grafana\n\nRequired Qualifications\n\nBachelor’s or Master’s degree in Computer Science, Information Systems, or related fields or equivalent experience\n\n3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role supporting production cloud environments\n\nCloud Architecture: Deep hands‑on experience with major cloud providers (AWS or GCP) and a strong command of containerization and orchestration technologies (Docker and Kubernetes).\n\nCoding & Scripting: Strong programming proficiency in languages such as Python, Go, TypeScript, or Bash for automation, internal tooling, and system integrations.\n\nSystems & Networking: Solid fundamentals in Linux/Unix administration, networking protocols (TCP/IP, DNS, HTTP/S, load balancing), and cloud security best practices.\n\nMindset: A passionate problem‑solver who prioritizes automation over manual operations and thrives in high‑ownership environments.\n\nMust pass drug and criminal background check\n\nWork well in an onsite team environment\n\nPreferred Qualifications\n\nExperience with multi‑region deployments, failover strategies, and data consistency\n\nMessaging Systems: Experience managing high‑throughput message queues or data streaming platforms, specifically RabbitMQ or Apache Kafka.\n\nExperience operating production systems at scale\n\nData Infrastructure: Familiarity with modern data stack environments, such as Snowflake or relational databases in a self‑hosted environment.\n\nCertifications: Relevant industry certifications such as Certified Kubernetes Administrator (CKA) or AWS Certified DevOps Engineer Professional.\n\nFamiliarity with regulatory requirements (SOC2, GDPR, etc)\n\nNice to Have\n\nExperience in physical security or access control systems\n\nFamiliarity with GCP ecosystem and tooling\n\nExperience working in a scaling startup environment\n\nCompetitive salary starting at $125,000 depending on experience\n\nComprehensive medical, dental, and vision coverage\n\n401(k) with company match\n\n3‑5 weeks PTO annually based on tenure\n\nPaid company holidays\n\nWork Location\nThis position is an in‑office, non‑remote role, working from ProdataKey Headquarters located in Draper, Utah.\n\n#J-18808-Ljbffr","company":"Prodatakey","rawCompany":"prodatakey","city":"Draper","state":"UT","isRemote":false,"isActive":false,"createdAt":"2026-07-16T03:37:13.220Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Site Reliability Engineer","description":"We are seeking a hands‑on Site Reliability Engineer to join our team. You will bridge the gap between software development and infrastructure operations, treating operational challenges as engineering problems. By leveraging automation, designing resilient distributed systems, and championing observability, you will ensure that our customers can secure and manage their access control systems without friction or failure.\n\nKey Responsibilities\n\nInfrastructure & Automation: Design, build, and maintain scalable, secure multi‑tenant cloud infrastructure using Infrastructure as Code (IaC) principles.\n\nUptime & Reliability: Own the availability, latency, performance, and capacity planning of the pdk.io platform and its supporting backend microservices.\n\nObservability: Develop and manage robust monitoring, logging, and alerting systems to gain deep visibility into cloud infrastructure, API health, and IoT endpoint performance.\n\nIncident Response: Participate in a collaborative on‑call rotation. Lead rapid incident response mitigation and drive rigorous, blameless post‑mortems to ensure long‑term system resilience.\n\nCI/CD Pipeline Management: Optimize and secure automated deployment pipelines to enable developers to ship code to production safely and efficiently.\n\nCross‑Functional Collaboration: Partner closely with backend developers and hardware engineering teams to define Service Level Indicators (SLIs), Service Level Objectives (SLOs), and manage error budgets.\n\nContribute to technical documentation and knowledge sharing\n\nTooling\n\nCI automation with GitHub Actions and Argo Workflows\n\nIaC with OpenTofu & Terragrunt at the core\n\nObservability stack powered by Prometheus and Grafana\n\nRequired Qualifications\n\nBachelor’s or Master’s degree in Computer Science, Information Systems, or related fields or equivalent experience\n\n3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role supporting production cloud environments\n\nCloud Architecture: Deep hands‑on experience with major cloud providers (AWS or GCP) and a strong command of containerization and orchestration technologies (Docker and Kubernetes).\n\nCoding & Scripting: Strong programming proficiency in languages such as Python, Go, TypeScript, or Bash for automation, internal tooling, and system integrations.\n\nSystems & Networking: Solid fundamentals in Linux/Unix administration, networking protocols (TCP/IP, DNS, HTTP/S, load balancing), and cloud security best practices.\n\nMindset: A passionate problem‑solver who prioritizes automation over manual operations and thrives in high‑ownership environments.\n\nMust pass drug and criminal background check\n\nWork well in an onsite team environment\n\nPreferred Qualifications\n\nExperience with multi‑region deployments, failover strategies, and data consistency\n\nMessaging Systems: Experience managing high‑throughput message queues or data streaming platforms, specifically RabbitMQ or Apache Kafka.\n\nExperience operating production systems at scale\n\nData Infrastructure: Familiarity with modern data stack environments, such as Snowflake or relational databases in a self‑hosted environment.\n\nCertifications: Relevant industry certifications such as Certified Kubernetes Administrator (CKA) or AWS Certified DevOps Engineer Professional.\n\nFamiliarity with regulatory requirements (SOC2, GDPR, etc)\n\nNice to Have\n\nExperience in physical security or access control systems\n\nFamiliarity with GCP ecosystem and tooling\n\nExperience working in a scaling startup environment\n\nCompetitive salary starting at $125,000 depending on experience\n\nComprehensive medical, dental, and vision coverage\n\n401(k) with company match\n\n3‑5 weeks PTO annually based on tenure\n\nPaid company holidays\n\nWork Location\nThis position is an in‑office, non‑remote role, working from ProdataKey Headquarters located in Draper, Utah.\n\n#J-18808-Ljbffr","datePosted":"2026-07-16T03:37:13.220Z","dateModified":"2026-07-16T03:37:13.220Z","hiringOrganization":{"@type":"Organization","name":"Prodatakey","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Draper","addressRegion":"UT","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"ca38bfcc2a76614469e1be7c"},"url":"https://jobsearcher.com/jobs/ca38bfcc2a76614469e1be7c"}}