{"schemaVersion":"jobsearcher.job.v1","id":"5a8dbd93bfd3181cb777c2b6","url":"https://jobsearcher.com/jobs/5a8dbd93bfd3181cb777c2b6","canonicalUrl":"https://jobsearcher.com/jobs/5a8dbd93bfd3181cb777c2b6","title":"Kubernetes System Engineer","description":"RedLine Performance Solutions (RedLine) has been in the HPC solutions engineering services business for over 26 years and is consistently determined to keep the \"bar of excellence\" quite high for new hires. This enables RedLine to accomplish what other firms cannot and promotes a high level of staff retention. We offer services ranging from full life cycle HPC systems engineering to remote managed services to HPC program analysis. We are looking for an Kubernetes System Engineer to join us.\n\nRedLine is looking for a Kubernetes System Engineer to join us. The successful candidate will be responsible for the architecture, operation, and maintenance of a critical High-Performance Computing (HPC) and Kubernetes infrastructure. This role requires a deep understanding of cloud-native technologies, robust security practices, and large-scale system administration to maintain a secure and reliable platform.\n\nAn active DoD Secret or Top Secret security clearance is a requirement to apply, as are current Linux+ and Security+ (or equivalent) certifications. This position on-site at the customer location in Aberdeen, Maryland. Reloation may be considered. This full-time position offers a full benefits package including paid time off, 401k match, and health care benefits.\n\nJob Responsibilities:\nKubernetes platform architecture and operations\nDesign, deploy, and operate highly available RKE2 Kubernetes clusters, including multi-control-plane environments with stable etcd quorum\nManage Kubernetes versioning upgrades and compatibility, along with cluster certificate authorities and trust chains\nOversee complete lifecycle of Kubernetes nodes (cordon, drain, replacement) and operate container runtimes like containerd\nTune kubelet behavior, manage resource pressure, and ensure consistent node configuration across all environments\nNetworking, security, and identity\nDesign and operate Kubernetes networking (CNI), implement network policies for workload isolation, and manage ingress controllers and DNS configurations\nImplement and enforce security best practices, including RBAC, admission controls, pod security standards, secrets management, and audit logging\nPerform routine systems administration and apply necessary STIGs and OS maintenance to ensure compliance for CUI-level operations\nIntegrate Kubernetes with enterprise identity services (LDAP/FreeIPA) and implement SSO with support for CAC/MFA\nData, CI/CD, and Reliability\nDesign and operate Kubernetes storage solutions using CSI drivers (Lustre, Weka), manage persistent volumes, and integrate S3 object storage\nOperate and maintain CI/CD infrastructure, including GitLab and container registries (Harbor, Artifactory), to support developer workflows\nImplement comprehensive monitoring, logging, and alerting. Lead incident response, perform capacity planning, and maintain operational runbooks\nArchitect for high availability, define RPO/RTO, and implement robust backup, restore, and failover procedures for all stateful services\nIntegrate Kubernetes workloads with HPC schedulers like Slurm/PBS and enable seamless, secure job submission and identity mapping between platforms.\n\nRequired Skills:\nProven experience in systems administration, particularly in Linux-based environments\nExtensive hands-on experience designing, building, and operating production Kubernetes clusters\nDeep understanding of Kubernetes networking, security principles (RBAC, Network Policy, Pod Security Standards), and storage (CSI)\nStrong knowledge of container runtimes (containers) and the full node lifecycle\nExperience integrating applications and platforms with identity management systems like LDAP or FreeIPA\nFamiliarity with operating CI/CD pipelines and associated tools (e.g., GitLab, Artifactory, Harbor).\n\nPreferred Skills:\nSpecific experience with RKE2 is highly desirable\nExperience working in secure, compliance-driven environments (e.g., CUI, DoD)\nKnowledge of integrating Kubernetes with HPC schedulers (Slurm, PBS) and high-performance storage (Lustre, Weka)\nProficiency with observability stacks for monitoring, logging, and alerting\nExperience with Infrastructure as Code (IaC) and configuration management tools\nDemonstrated ability to design and test high-availability and disaster recovery plans.\nTo learn more about what makes RedLine a great place to work, please visit our website at https://redlineperf.com/careers/\n\nTotal Rewards:\nCompetitive salary band: $140,000 – $180,000/year dependent on experience and relocation needs\nMedical, dental & vision coverage with substantial company contribution\nCompany funded Healthcare Reimbursement Account (HRA)\nPaid time off (PTO) + 11 paid holidays\nCompany Match 100% immediately vested retirement savings (401k)\nEmployee wellness programs & gym discounts\nEmployee assistance & concierge services\nProfessional development resources","company":"Redline Performance Solutions","rawCompany":"redline performance solutions","city":"Aberdeen","state":"MD","isRemote":false,"isActive":false,"createdAt":"2026-04-12T19:28:35.097Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"15-1211.00","title":"Computer Systems Analysts","slug":"computer-systems-analysts"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541519","title":"Other Computer Related Services","slug":"other-computer-related-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Kubernetes System Engineer","description":"RedLine Performance Solutions (RedLine) has been in the HPC solutions engineering services business for over 26 years and is consistently determined to keep the \"bar of excellence\" quite high for new hires. This enables RedLine to accomplish what other firms cannot and promotes a high level of staff retention. We offer services ranging from full life cycle HPC systems engineering to remote managed services to HPC program analysis. We are looking for an Kubernetes System Engineer to join us.\n\nRedLine is looking for a Kubernetes System Engineer to join us. The successful candidate will be responsible for the architecture, operation, and maintenance of a critical High-Performance Computing (HPC) and Kubernetes infrastructure. This role requires a deep understanding of cloud-native technologies, robust security practices, and large-scale system administration to maintain a secure and reliable platform.\n\nAn active DoD Secret or Top Secret security clearance is a requirement to apply, as are current Linux+ and Security+ (or equivalent) certifications. This position on-site at the customer location in Aberdeen, Maryland. Reloation may be considered. This full-time position offers a full benefits package including paid time off, 401k match, and health care benefits.\n\nJob Responsibilities:\nKubernetes platform architecture and operations\nDesign, deploy, and operate highly available RKE2 Kubernetes clusters, including multi-control-plane environments with stable etcd quorum\nManage Kubernetes versioning upgrades and compatibility, along with cluster certificate authorities and trust chains\nOversee complete lifecycle of Kubernetes nodes (cordon, drain, replacement) and operate container runtimes like containerd\nTune kubelet behavior, manage resource pressure, and ensure consistent node configuration across all environments\nNetworking, security, and identity\nDesign and operate Kubernetes networking (CNI), implement network policies for workload isolation, and manage ingress controllers and DNS configurations\nImplement and enforce security best practices, including RBAC, admission controls, pod security standards, secrets management, and audit logging\nPerform routine systems administration and apply necessary STIGs and OS maintenance to ensure compliance for CUI-level operations\nIntegrate Kubernetes with enterprise identity services (LDAP/FreeIPA) and implement SSO with support for CAC/MFA\nData, CI/CD, and Reliability\nDesign and operate Kubernetes storage solutions using CSI drivers (Lustre, Weka), manage persistent volumes, and integrate S3 object storage\nOperate and maintain CI/CD infrastructure, including GitLab and container registries (Harbor, Artifactory), to support developer workflows\nImplement comprehensive monitoring, logging, and alerting. Lead incident response, perform capacity planning, and maintain operational runbooks\nArchitect for high availability, define RPO/RTO, and implement robust backup, restore, and failover procedures for all stateful services\nIntegrate Kubernetes workloads with HPC schedulers like Slurm/PBS and enable seamless, secure job submission and identity mapping between platforms.\n\nRequired Skills:\nProven experience in systems administration, particularly in Linux-based environments\nExtensive hands-on experience designing, building, and operating production Kubernetes clusters\nDeep understanding of Kubernetes networking, security principles (RBAC, Network Policy, Pod Security Standards), and storage (CSI)\nStrong knowledge of container runtimes (containers) and the full node lifecycle\nExperience integrating applications and platforms with identity management systems like LDAP or FreeIPA\nFamiliarity with operating CI/CD pipelines and associated tools (e.g., GitLab, Artifactory, Harbor).\n\nPreferred Skills:\nSpecific experience with RKE2 is highly desirable\nExperience working in secure, compliance-driven environments (e.g., CUI, DoD)\nKnowledge of integrating Kubernetes with HPC schedulers (Slurm, PBS) and high-performance storage (Lustre, Weka)\nProficiency with observability stacks for monitoring, logging, and alerting\nExperience with Infrastructure as Code (IaC) and configuration management tools\nDemonstrated ability to design and test high-availability and disaster recovery plans.\nTo learn more about what makes RedLine a great place to work, please visit our website at https://redlineperf.com/careers/\n\nTotal Rewards:\nCompetitive salary band: $140,000 – $180,000/year dependent on experience and relocation needs\nMedical, dental & vision coverage with substantial company contribution\nCompany funded Healthcare Reimbursement Account (HRA)\nPaid time off (PTO) + 11 paid holidays\nCompany Match 100% immediately vested retirement savings (401k)\nEmployee wellness programs & gym discounts\nEmployee assistance & concierge services\nProfessional development resources","datePosted":"2026-04-12T19:28:35.097Z","dateModified":"2026-04-12T19:28:35.097Z","hiringOrganization":{"@type":"Organization","name":"Redline Performance Solutions","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Aberdeen","addressRegion":"MD","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"5a8dbd93bfd3181cb777c2b6"},"url":"https://jobsearcher.com/jobs/5a8dbd93bfd3181cb777c2b6"}}