{"schemaVersion":"jobsearcher.job.v1","id":"2ce2fd00f488bdebb057a0ce","url":"https://jobsearcher.com/jobs/2ce2fd00f488bdebb057a0ce","canonicalUrl":"https://jobsearcher.com/jobs/2ce2fd00f488bdebb057a0ce","title":"Distributed Systems Architect (Python)","description":"Job role: SRE Architect [AIOps & Dynatrace]\nLocation : Atlanta, GA [ hybrid ]\nDuration : Long Term\nSkills Required: SRE, AWS, AIOps, Dynatrace, Observability\nWe are looking for a SRE Architect who combines strong operational expertise with hands on development skills. This role requires deep experience managing large scale distributed systems, improving reliability, eliminating toil through automation, and guiding domain teams in maturing their SRE practices.\nThe ideal candidate is equally comfortable debugging production issues, writing automation or REST APIs, interpreting code repositories, and implementing resiliency patterns such as circuit breakers.\nThis is a pure Individual Contributor role with high technical depth and ownership.\nKey Responsibilities\nPerform SRE operations for distributed systems, ensuring high availability, reliability, and operational excellence.\nAI in SRE\nPartner with application/domain teams to strengthen their SRE maturity and operational readiness.\nWrite automation, scripts, and REST APIs to integrate with external systems and eliminate repetitive tasks.\nOnboard services to Dynatrace/observability platforms; define dashboards, alerts, SLIs, SLOs.\nArchitect and implement resiliency patterns including failover strategies, circuit breakers, graceful degradation.\nDrive cost optimization (FinOps) initiatives across cloud workloads.\nSupport AWS (or other cloud platforms) operations and engineering needs.\nWork with ROSA/container platforms for deployment, scaling, and reliability.\nRecommend improvements in technology, architecture, and domain-specific reliability areas.\nManage and support large scale systems operating at scale.\nReduce toil by identifying repetitive tasks and automating them.\nContribute code, read/interpret service repositories, and assist teams with engineering tasks as needed.\nRequired Skills & Experience\nStrong background in SRE operations for distributed systems.\nProficiency in development/coding (Python, Go, shell scripting, or similar).\nAbility to read/interpret codebases and build REST APIs.\nExperience with Dynatrace/observability onboarding and ecosystem.\nDeep knowledge of resiliency engineering and failover strategies.\nStrong understanding of FinOps principles and cloud cost optimization.\nHands-on experience with AWS or any other cloud provider.\nExperience with ROSA or Kubernetes-based container platforms.\nProven automation skills to eliminate operational toil.\nExperience managing large-scale systems in production.\nCapability to suggest architecture and domain improvements.\nStrong analytical, troubleshooting, and collaboration skills.\nFor applications and inquiries, contact: hirings@openkyber.com","company":"Openkyber","rawCompany":"openkyber","city":"Alaska","state":"MI","isRemote":false,"isActive":false,"createdAt":"2026-08-09T14:25:34.029Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Distributed Systems Architect (Python)","description":"Job role: SRE Architect [AIOps & Dynatrace]\nLocation : Atlanta, GA [ hybrid ]\nDuration : Long Term\nSkills Required: SRE, AWS, AIOps, Dynatrace, Observability\nWe are looking for a SRE Architect who combines strong operational expertise with hands on development skills. This role requires deep experience managing large scale distributed systems, improving reliability, eliminating toil through automation, and guiding domain teams in maturing their SRE practices.\nThe ideal candidate is equally comfortable debugging production issues, writing automation or REST APIs, interpreting code repositories, and implementing resiliency patterns such as circuit breakers.\nThis is a pure Individual Contributor role with high technical depth and ownership.\nKey Responsibilities\nPerform SRE operations for distributed systems, ensuring high availability, reliability, and operational excellence.\nAI in SRE\nPartner with application/domain teams to strengthen their SRE maturity and operational readiness.\nWrite automation, scripts, and REST APIs to integrate with external systems and eliminate repetitive tasks.\nOnboard services to Dynatrace/observability platforms; define dashboards, alerts, SLIs, SLOs.\nArchitect and implement resiliency patterns including failover strategies, circuit breakers, graceful degradation.\nDrive cost optimization (FinOps) initiatives across cloud workloads.\nSupport AWS (or other cloud platforms) operations and engineering needs.\nWork with ROSA/container platforms for deployment, scaling, and reliability.\nRecommend improvements in technology, architecture, and domain-specific reliability areas.\nManage and support large scale systems operating at scale.\nReduce toil by identifying repetitive tasks and automating them.\nContribute code, read/interpret service repositories, and assist teams with engineering tasks as needed.\nRequired Skills & Experience\nStrong background in SRE operations for distributed systems.\nProficiency in development/coding (Python, Go, shell scripting, or similar).\nAbility to read/interpret codebases and build REST APIs.\nExperience with Dynatrace/observability onboarding and ecosystem.\nDeep knowledge of resiliency engineering and failover strategies.\nStrong understanding of FinOps principles and cloud cost optimization.\nHands-on experience with AWS or any other cloud provider.\nExperience with ROSA or Kubernetes-based container platforms.\nProven automation skills to eliminate operational toil.\nExperience managing large-scale systems in production.\nCapability to suggest architecture and domain improvements.\nStrong analytical, troubleshooting, and collaboration skills.\nFor applications and inquiries, contact: hirings@openkyber.com","datePosted":"2026-08-09T14:25:34.029Z","dateModified":"2026-08-09T14:25:34.029Z","hiringOrganization":{"@type":"Organization","name":"Openkyber","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Alaska","addressRegion":"MI","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"2ce2fd00f488bdebb057a0ce"},"url":"https://jobsearcher.com/jobs/2ce2fd00f488bdebb057a0ce"}}