Systems Engineer - TS/SCI
Overview
In this role you will design, deploy, and optimize HPC and GPU clusters for IC/DoD-adjacent environments. You will work closely with cross-functional teams to ensure scalable, power-efficient Linux-based GPU infrastructure across on-premises and cloud, while advancing performance and reliability. You will stay ahead of GPU and Linux trends and contribute to secure, well-documented solutions. This is a hands-on, impact-focused position in a stateful, security-cleared setting with ongoing optimization and validation work.
ResponsibilitiesInstall and maintain GPU/HPC hardware on‑prem and in the cloudAnalyze cluster performance to identify bottlenecks and improve efficiencyInstall and configure HPC/GPU schedulers and workflow tools (Slurm, PBS, Apache Airflow, Kubernetes)Work on GPU power management for Linux platformsDesign and execute GPU/Linux tests, benchmarking, and debugging; expand testing suiteCreate and maintain technical documentation and Linux best practicesStay informed on GPU industry trends and contribute to Linux-specific optimization approachesCollaborate with the team to deliver robust GPU solutions
Key requirementsBachelor's degree in Computer Science, Electrical Engineering, or related field4+ years of relevant systems engineering experienceExpertise in Linux OS integration and hardware architectureKnowledge of parallel computing, graphics algorithms, and real-time rendering in LinuxStrong problem-solving and team collaborationExcellent communication in a Linux contextProficiency with Python or BashProficiency with automation tools (Ansible, Puppet, Salt, Terraform)DoD 8570.11 IAT Level II certification or higher (e.g., Security+ CE, CCNA-Security, GICSP, GSEC, SSCP; CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP)Preferred: GPU virtualization, Docker/Kubernetes, Prometheus/Grafana, distributed resource scheduling, data center networking conceptsFamiliarity with DHCP, DNS, TCP/IP, VLANs, HSRP, SNMP; knowledge of firewall, IDS/IPS, and PBR conceptsExcellent communicationStrong problem-solvingTeam collaborationLinux system administrationGPU hardware and NVIDIA productsKubernetes, Docker