{"schemaVersion":"jobsearcher.job.v1","id":"5b3d42f3026effd485762865","url":"https://jobsearcher.com/jobs/5b3d42f3026effd485762865","canonicalUrl":"https://jobsearcher.com/jobs/5b3d42f3026effd485762865","title":"Manager - Software Engineering DevSecOps - Incident","description":"NO SPONSORSHIP - NO OPTManager, Software Engineering DevSecOpsSALARY: $160k - $190kLOCATION: Chicago, ILHybrid 3 days onsite and 2 days remoteLooking for a manager to lead a team 6-10 - L1, L2, support engineers. 24/7 support. Jenkins CI/CD, Kubernetes k8s apache kafka splunk datadog Dynatrace application servers devops middleware storage network security vault certs secrets sla governance platform infrastructure servicenow or equivalent ITSM toolingIncident Management & Environment Support:Lead L1 and L2 support engineers in all incident response activities including triage, investigation, coordination, resolution, closure, and post-incident reporting.Oversee technical analysis of environment incidents across application deployments, middleware, and platform layers while coordinating response activities with internal engineering, platform, and application development teams.Serve as Tier 3 escalation point for complex incidents beyond L2 capability — triaging, directing, and driving resolution across Platform (k8s, Kafka, TFE), S&I (deployment, middleware, storage, network), Security (Vault, certs, secrets), and App Dev teams.Own the full incident lifecycle — from first alert through to RCA documentation and permanent fix or accepted workaround.Drive post-incident reviews for all P1 and P2 incidents, ensuring root cause is identified, documented, and actioned — not filed.Qualifications:Technical Skills:Deployment & Pipeline tooling: Harness (continuous delivery pipelines, deployment verification, rollback automation), Jenkins (CI/CD pipeline management, job configuration, build troubleshooting), GitHub (branching strategies, pull request workflows, pipeline integration).Container & orchestration platforms: Kubernetes (k8s) pod lifecycle management, namespace operations, log retrieval, resource troubleshooting, and coordination with Platform teams on cluster-level issues.Messaging & streaming platforms: Apache Kafka topic management, consumer group monitoring, lag analysis, and escalation to Platform for broker-level issues.Secrets & configuration management: HashiCorp Vault — secrets retrieval, token/lease troubleshooting, policy review, and escalation to Security teams for certificate and secrets rotation.Monitoring & observability: Proficiency in at least two production monitoring toolsets (e.g. Splunk, Dynatrace, Datadog, AppDynamics, Prometheus/Grafana) alert triage, dashboard interpretation, log analysis, and tuning requests.Middleware platforms: Working knowledge of middleware infrastructure including application servers, messaging brokers, storage integrations, and network-layer dependencies sufficient to triage, gather diagnostics, and route correctly to L3.Incident and ticketing platforms: ServiceNow or equivalent ITSM tooling incident creation, SLA tracking, problem record management, and reporting.MTTR and operational metrics: Ability to build and maintain operational dashboards and reports covering MTTR, SLA compliance, alert-to-incident ratio, repeat incident rate, and deployment success rate.Education and/or Experience:Minimum 5 years of hands-on environment operations, production support, or infrastructure operations experience, including interdisciplinary experience across four or more of the following: application deployment pipelines, container platform operations, middleware support, incident management, monitoring and observability, configuration management, release engineering, platform operations, or scripting and automation.Technical experience and comprehensive knowledge of production environment failure modes — including deployment failures, configuration drift, platform instability, and integration breakdowns — and the methodologies used to diagnose and resolve them.Demonstrated experience defining and enforcing SLA frameworks in a tiered support model (L1/L2/L3 or equivalent).Familiarity with financial services or other regulated-industry production environments is a strong advantage — understanding of change governance, audit requirements, and production access controls.Industry knowledge of current and emerging practices in environment operations, platform reliability, and support automation.Shift work and on-call availability required — including 24×7 on-call response capacity and availability during planned and emergency maintenance windows.Previous people management or team lead experience required; formal people management experience strongly preferred.","company":"Request Technology","rawCompany":"request technology","city":"Chicago","state":"IL","isRemote":false,"isActive":false,"createdAt":"2026-08-06T10:23:40.934Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"11-3021.00","title":"Computer and Information Systems Managers","slug":"computer-and-information-systems-managers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541519","title":"Other Computer Related Services","slug":"other-computer-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Manager - Software Engineering DevSecOps - Incident","description":"NO SPONSORSHIP - NO OPTManager, Software Engineering DevSecOpsSALARY: $160k - $190kLOCATION: Chicago, ILHybrid 3 days onsite and 2 days remoteLooking for a manager to lead a team 6-10 - L1, L2, support engineers. 24/7 support. Jenkins CI/CD, Kubernetes k8s apache kafka splunk datadog Dynatrace application servers devops middleware storage network security vault certs secrets sla governance platform infrastructure servicenow or equivalent ITSM toolingIncident Management & Environment Support:Lead L1 and L2 support engineers in all incident response activities including triage, investigation, coordination, resolution, closure, and post-incident reporting.Oversee technical analysis of environment incidents across application deployments, middleware, and platform layers while coordinating response activities with internal engineering, platform, and application development teams.Serve as Tier 3 escalation point for complex incidents beyond L2 capability — triaging, directing, and driving resolution across Platform (k8s, Kafka, TFE), S&I (deployment, middleware, storage, network), Security (Vault, certs, secrets), and App Dev teams.Own the full incident lifecycle — from first alert through to RCA documentation and permanent fix or accepted workaround.Drive post-incident reviews for all P1 and P2 incidents, ensuring root cause is identified, documented, and actioned — not filed.Qualifications:Technical Skills:Deployment & Pipeline tooling: Harness (continuous delivery pipelines, deployment verification, rollback automation), Jenkins (CI/CD pipeline management, job configuration, build troubleshooting), GitHub (branching strategies, pull request workflows, pipeline integration).Container & orchestration platforms: Kubernetes (k8s) pod lifecycle management, namespace operations, log retrieval, resource troubleshooting, and coordination with Platform teams on cluster-level issues.Messaging & streaming platforms: Apache Kafka topic management, consumer group monitoring, lag analysis, and escalation to Platform for broker-level issues.Secrets & configuration management: HashiCorp Vault — secrets retrieval, token/lease troubleshooting, policy review, and escalation to Security teams for certificate and secrets rotation.Monitoring & observability: Proficiency in at least two production monitoring toolsets (e.g. Splunk, Dynatrace, Datadog, AppDynamics, Prometheus/Grafana) alert triage, dashboard interpretation, log analysis, and tuning requests.Middleware platforms: Working knowledge of middleware infrastructure including application servers, messaging brokers, storage integrations, and network-layer dependencies sufficient to triage, gather diagnostics, and route correctly to L3.Incident and ticketing platforms: ServiceNow or equivalent ITSM tooling incident creation, SLA tracking, problem record management, and reporting.MTTR and operational metrics: Ability to build and maintain operational dashboards and reports covering MTTR, SLA compliance, alert-to-incident ratio, repeat incident rate, and deployment success rate.Education and/or Experience:Minimum 5 years of hands-on environment operations, production support, or infrastructure operations experience, including interdisciplinary experience across four or more of the following: application deployment pipelines, container platform operations, middleware support, incident management, monitoring and observability, configuration management, release engineering, platform operations, or scripting and automation.Technical experience and comprehensive knowledge of production environment failure modes — including deployment failures, configuration drift, platform instability, and integration breakdowns — and the methodologies used to diagnose and resolve them.Demonstrated experience defining and enforcing SLA frameworks in a tiered support model (L1/L2/L3 or equivalent).Familiarity with financial services or other regulated-industry production environments is a strong advantage — understanding of change governance, audit requirements, and production access controls.Industry knowledge of current and emerging practices in environment operations, platform reliability, and support automation.Shift work and on-call availability required — including 24×7 on-call response capacity and availability during planned and emergency maintenance windows.Previous people management or team lead experience required; formal people management experience strongly preferred.","datePosted":"2026-08-06T10:23:40.934Z","dateModified":"2026-08-06T10:23:40.934Z","hiringOrganization":{"@type":"Organization","name":"Request Technology","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Chicago","addressRegion":"IL","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"5b3d42f3026effd485762865"},"url":"https://jobsearcher.com/jobs/5b3d42f3026effd485762865"}}