{"schemaVersion":"jobsearcher.job.v1","id":"099deab29cd60fcc2ccf9ed8","url":"https://jobsearcher.com/jobs/099deab29cd60fcc2ccf9ed8","canonicalUrl":"https://jobsearcher.com/jobs/099deab29cd60fcc2ccf9ed8","title":"Observability Operation Engineer","description":"Job Title : Observability Operations EngineerJob Location : Phoenix, Arizona, United StatesSkill Cluster : - AIMS-ObservabilityOpenSource.Primary Skill : Observability, Splunk Enterprise Administration, KubernetesExperience : 10 yearsCurrent Status : ActiveJob Duration : 6-12 MonthsJob Description :P1C3STS We are seeking a highly skilled Senior Observability Operations Engineer to manage and enhance our enterprise observability platform. The ideal candidate will have deep expertise in Dynatrace, Splunk, OpenSearch/Elasticsearch, Kubernetes, Linux, and cloud-native observability solutions. Experience leveraging AI/ML and Generative AI to improve observability, automate operations, and accelerate incident resolution is highly desirable.The role is responsible for ensuring high availability, scalability, operational excellence, and continuous improvement of enterprise monitoring and logging platforms supporting mission-critical applications.Required Technical SkillsObservability PlatformsDynatrace AdministrationSplunk Enterprise AdministrationOpenSearch AdministrationElasticsearch AdministrationGrafanaPrometheusKibanaJaegerOpenTelemetryKafka (preferred)Administer and optimize enterprise observability platforms including Dynatrace, Splunk, and OpenSearch/Elasticsearch.Manage large-scale OpenSearch/Elasticsearch clusters, including indexing strategies, performance tuning, shard optimization, backups, and capacity planning.Configure Dynatrace OneAgent, ActiveGate, Synthetic Monitoring, Real User Monitoring (RUM), Digital Experience Monitoring (DEM), Davis AI, and Application Performance Monitoring (APM).Administer Splunk Enterprise, Universal Forwarders, Indexers, Search Heads, Cluster Manager, Deployment Server, and Splunk ITSI.Develop dashboards, alerts, reports, and executive operational metrics.Support Linux-based infrastructure and Kubernetes environments (Docker/OpenShift/Rancher preferred).Implement observability best practices using OpenTelemetry, distributed tracing, metrics, logs, and events.Perform root cause analysis for production incidents using observability platforms.Collaborate with Platform Engineering, SRE, DevOps, Infrastructure, and Application teams.Automate operational tasks using Python, Shell scripting, REST APIs, Terraform, or Ansible..Improve platform reliability through automation, self-healing, and AI-assisted operationsDynatrace AdministrationSplunk, Open Search AdminstrationELK, Prometheus, GrefanaKibanaKafkaDocker, KubernetesAny Cloud experience","company":"Kaizen Technologies","rawCompany":"kaizen technologies","city":"Phoenix","state":"AZ","isRemote":false,"isActive":false,"createdAt":"2026-09-22T07:44:29.927Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"15-1299.05","title":"Information Security Engineers","slug":"information-security-engineers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Observability Operation Engineer","description":"Job Title : Observability Operations EngineerJob Location : Phoenix, Arizona, United StatesSkill Cluster : - AIMS-ObservabilityOpenSource.Primary Skill : Observability, Splunk Enterprise Administration, KubernetesExperience : 10 yearsCurrent Status : ActiveJob Duration : 6-12 MonthsJob Description :P1C3STS We are seeking a highly skilled Senior Observability Operations Engineer to manage and enhance our enterprise observability platform. The ideal candidate will have deep expertise in Dynatrace, Splunk, OpenSearch/Elasticsearch, Kubernetes, Linux, and cloud-native observability solutions. Experience leveraging AI/ML and Generative AI to improve observability, automate operations, and accelerate incident resolution is highly desirable.The role is responsible for ensuring high availability, scalability, operational excellence, and continuous improvement of enterprise monitoring and logging platforms supporting mission-critical applications.Required Technical SkillsObservability PlatformsDynatrace AdministrationSplunk Enterprise AdministrationOpenSearch AdministrationElasticsearch AdministrationGrafanaPrometheusKibanaJaegerOpenTelemetryKafka (preferred)Administer and optimize enterprise observability platforms including Dynatrace, Splunk, and OpenSearch/Elasticsearch.Manage large-scale OpenSearch/Elasticsearch clusters, including indexing strategies, performance tuning, shard optimization, backups, and capacity planning.Configure Dynatrace OneAgent, ActiveGate, Synthetic Monitoring, Real User Monitoring (RUM), Digital Experience Monitoring (DEM), Davis AI, and Application Performance Monitoring (APM).Administer Splunk Enterprise, Universal Forwarders, Indexers, Search Heads, Cluster Manager, Deployment Server, and Splunk ITSI.Develop dashboards, alerts, reports, and executive operational metrics.Support Linux-based infrastructure and Kubernetes environments (Docker/OpenShift/Rancher preferred).Implement observability best practices using OpenTelemetry, distributed tracing, metrics, logs, and events.Perform root cause analysis for production incidents using observability platforms.Collaborate with Platform Engineering, SRE, DevOps, Infrastructure, and Application teams.Automate operational tasks using Python, Shell scripting, REST APIs, Terraform, or Ansible..Improve platform reliability through automation, self-healing, and AI-assisted operationsDynatrace AdministrationSplunk, Open Search AdminstrationELK, Prometheus, GrefanaKibanaKafkaDocker, KubernetesAny Cloud experience","datePosted":"2026-09-22T07:44:29.927Z","dateModified":"2026-09-22T07:44:29.927Z","hiringOrganization":{"@type":"Organization","name":"Kaizen Technologies","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Phoenix","addressRegion":"AZ","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"099deab29cd60fcc2ccf9ed8"},"url":"https://jobsearcher.com/jobs/099deab29cd60fcc2ccf9ed8"}}