JOBSEARCHER

High Performance Computing Engineer

High Performance Computing EngineerSan Francisco Bay AreaWe are pioneering advanced computational solutions requiring massive computational power and sophisticated optimization. Our HPC team designs and implements cutting-edge solutions that push the boundaries of computational performance.We're seeking a High Performance Computing Engineer with deep expertise in parallel computing and heterogeneous architecture optimization. The ideal candidate will have extensive experience with CUDA/ROCm, MPI, and large-scale system optimization.Advanced expertise in using AI coding assistants for HPC code generation and optimizationDemonstrated ability to leverage AI tools for performance debugging and analysisExperience using AI systems for parallel algorithm implementation and optimizationStrong proficiency in using AI to accelerate development workflowsAbility to effectively combine AI-generated solutions with manual optimizationExperience using AI tools for code review and quality assuranceDesign and optimize high-performance computing applications for large-scale clustersImplement and tune parallel algorithms across distributed systemsDevelop efficient CUDA/ROCm implementations for computational kernelsCreate and maintain MPI-based distributed computing solutionsOptimize memory access patterns and communication protocolsProfile and analyze performance bottlenecks in distributed systemsImplement efficient linear algebra and scientific computing routinesManage and optimize job scheduling and resource allocationDesign and implement fault tolerance and recovery systemsCollaborate with research teams to optimize computational workloadsM.S. or Ph.D. in Computer Science, Engineering, or related field from a top-tier universityExpert knowledge of parallel programming models (MPI, OpenMP)Advanced expertise in CUDA or ROCm programmingStrong proficiency in C++ and parallel algorithm implementationDeep understanding of computer architecture and memory hierarchiesExpert-level knowledge of Linux/Unix environmentsExperience with large-scale cluster managementExperience with InfiniBand and high-speed interconnectsKnowledge of scientific computing libraries (BLAS, LAPACK)Expertise in vectorization and SIMD optimizationBackground in distributed algorithmsExperience with performance modeling and predictionFamiliarity with job scheduling systems (Slurm, PBS)Experience with container technologies for HPCACM-ICPC Regional or World Finals medalistUSACO (USA Computing Olympiad) Gold/Platinum awardTop-tier algorithmic competition achievementsAdvanced parallel programming techniquesProficiency in performance optimization toolsStrong debugging and profiling skillsExpertise in distributed computing conceptsKnowledge of network topology and optimizationExperience with scientific computing applicationsCUDA/ROCmMPIOpenMPLinux/UnixPerformance profiling toolsJob scheduling systemsVersion control systemsC/C++Access to cutting-edge HPC infrastructureCompetitive compensation packageProfessional development opportunitiesCollaboration with leading researchersHealth and retirement benefits