{"schemaVersion":"jobsearcher.job.v1","id":"61cace8d32cd99119f8da264","url":"https://jobsearcher.com/jobs/61cace8d32cd99119f8da264","canonicalUrl":"https://jobsearcher.com/jobs/61cace8d32cd99119f8da264","title":"DevOps, SRE & Application Infrastructure Architect","description":"Please Find The JD BelowPosition: Sr. Principal DevOps, SRE & Application Infrastructure ArchitectLocation: Sunnyvale, CA (Onsite)Duration: 6 to 12+ Months ContractSRE & Application Infrastructure Architect (Minimum 4 to 5 Years)Job DescriptionNote: Must have the experience DevOps, SRE Application Infrastructure Architect (Minimum 4 to 5 Years)Skills: Digital : DevOps~Digital : Site Reliability Engineering (SRE)Key ResponsibilitiesInfrastructure & GitOpsK8s & Containerization: Design| deploy| and optimize secure Docker/Kubernetes (AKS) environments using Helm and ArgoCD.Networking & Edge: Manage cloud Ingress| Load Balancers| and end-to-end certificate management (SSL/mTLS).CI/CD & Automation: Automate tasks with Shell/Python; build GitOps pipelines and manage schema migrations via Flyway.SRE & Observability Reliability: Own end-to-end production availability and performance; define and track SLAs/SLOs/SLIs and error budgets.Telemetry: Build observability stacks using OpenTelemetry | Prometheus| Grafana| and Splunk.Incident Management: Lead P0/P1 incident response| deep-dive distributed system debugging| RCAs| and on-call rotations.Application & Database Operations Polyglot DB Management: Design and operate high-availability cloud database infrastructure (Oracle| Postgres| Cassandra| Couchbase| Redis| CockroachDB).Data Replication & DR: Manage Oracle GoldenGate replication| patching| purging| and execute robust P0 Disaster Recovery/failover strategies.App Support: Perform deep-dive troubleshooting within Java application layers| gRPC| REST/HTTP/JSON| and caching/messaging systems (SNS/SQS| Elasticsearch| Solr).Required Technical SkillsOrchestration & DevOps: Kubernetes (AKS)| Docker| Helm| ArgoCD | Flyway. Scripting & OS: Strong Linux/Unix internals| Shell scripting| and Python. Observability: OpenTelemetry | Prometheus| Grafana| Splunk. Database & Replication: SQL/PL-SQL (procedures| triggers| tuning) | GoldenGate | NoSQL (Cassandra| Couchbase| Redis)| and Transactional DBs (Oracle| Postgres).App Troubleshooting: Java application debugging| gRPC| REST| and cloud-native caching/queues.","company":"Compugra Systems","rawCompany":"compugra systems","city":"Sunnyvale","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-07-20T10:42:55.192Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"DevOps, SRE & Application Infrastructure Architect","description":"Please Find The JD BelowPosition: Sr. Principal DevOps, SRE & Application Infrastructure ArchitectLocation: Sunnyvale, CA (Onsite)Duration: 6 to 12+ Months ContractSRE & Application Infrastructure Architect (Minimum 4 to 5 Years)Job DescriptionNote: Must have the experience DevOps, SRE Application Infrastructure Architect (Minimum 4 to 5 Years)Skills: Digital : DevOps~Digital : Site Reliability Engineering (SRE)Key ResponsibilitiesInfrastructure & GitOpsK8s & Containerization: Design| deploy| and optimize secure Docker/Kubernetes (AKS) environments using Helm and ArgoCD.Networking & Edge: Manage cloud Ingress| Load Balancers| and end-to-end certificate management (SSL/mTLS).CI/CD & Automation: Automate tasks with Shell/Python; build GitOps pipelines and manage schema migrations via Flyway.SRE & Observability Reliability: Own end-to-end production availability and performance; define and track SLAs/SLOs/SLIs and error budgets.Telemetry: Build observability stacks using OpenTelemetry | Prometheus| Grafana| and Splunk.Incident Management: Lead P0/P1 incident response| deep-dive distributed system debugging| RCAs| and on-call rotations.Application & Database Operations Polyglot DB Management: Design and operate high-availability cloud database infrastructure (Oracle| Postgres| Cassandra| Couchbase| Redis| CockroachDB).Data Replication & DR: Manage Oracle GoldenGate replication| patching| purging| and execute robust P0 Disaster Recovery/failover strategies.App Support: Perform deep-dive troubleshooting within Java application layers| gRPC| REST/HTTP/JSON| and caching/messaging systems (SNS/SQS| Elasticsearch| Solr).Required Technical SkillsOrchestration & DevOps: Kubernetes (AKS)| Docker| Helm| ArgoCD | Flyway. Scripting & OS: Strong Linux/Unix internals| Shell scripting| and Python. Observability: OpenTelemetry | Prometheus| Grafana| Splunk. Database & Replication: SQL/PL-SQL (procedures| triggers| tuning) | GoldenGate | NoSQL (Cassandra| Couchbase| Redis)| and Transactional DBs (Oracle| Postgres).App Troubleshooting: Java application debugging| gRPC| REST| and cloud-native caching/queues.","datePosted":"2026-07-20T10:42:55.192Z","dateModified":"2026-07-20T10:42:55.192Z","hiringOrganization":{"@type":"Organization","name":"Compugra Systems","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Sunnyvale","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"61cace8d32cd99119f8da264"},"url":"https://jobsearcher.com/jobs/61cace8d32cd99119f8da264"}}