JOBSEARCHER

Mainframe Support

ARCHIVED

We can't find an active application page for this role right now. It may reopen or be listed elsewhere. Use Next Steps to search for an active apply link and similar live jobs.

Title: Mainframe SupportLocation: Columbus, Ohio, United States (Onsite)Type: C2C/W2Role SummaryThe Mainframe SRE is responsible for ensuring the reliability, availability, performance, and scalability of enterprise mainframe platforms. This role blends traditional mainframe engineering with modern SRE principles, focusing on automation, observability, incident management, and continuous improvement. The lead will guide a team of engineers while partnering closely with application, infrastructure, and operations teams.Key ResponsibilitiesLead the Mainframe SRE support team, providing technical direction, mentoring, and performance guidanceOwn the reliability, availability, and resilience of mainframe environments (z/OS and related subsystems)Define and implement SRE practices such as SLIs, SLOs, SLAs, error budgets, and reliability metricsDrive automation to reduce manual operations, improve recovery time, and enhance system stabilityOversee monitoring, alerting, and observability for mainframe systems using modern and legacy toolsLead incident management, root cause analysis (RCA), and post-incident reviewsPartner with application development teams to improve reliability, performance, and deployment practicesPlan and execute capacity management, performance tuning, and workload optimizationEnsure compliance with security, regulatory, and audit requirementsLead disaster recovery (DR) planning, testing, and high-availability strategiesChampion continuous improvement, DevOps, and SRE culture within mainframe operationsRequired Qualifications10+ years of experience in mainframe systems engineering or operationsStrong hands-on expertise with IBM z/OSExperience with core mainframe components such as:CICS, IMS, DB2JES2/JES3MQ, SMF, SDSFSolid understanding of mainframe performance tuning and capacity planningExperience leading production support and managing major incidentsStrong scripting and automation skills (REXX, JCL, CLIST, Python, or equivalent)Familiarity with monitoring and scheduling tools (e.g., OMEGAMON, CA/BMC tools, Control-M)Preferred QualificationsExperience applying SRE principles in a mainframe or hybrid (mainframe + distributed) environmentExposure to DevOps, CI/CD, and automation frameworksKnowledge of Linux on Z and cloud integration patternsExperience with resilience engineering, chaos testing, or fault injection conceptsPrior people-lead or technical-lead experience