Senior Bioinformatics / Scientific Data Engineer
Overview
In this role you design, implement, and maintain data pipelines and computational workflows to support large-scale scientific datasets. You’ll build ETL processes and automation to ensure data is analysis-ready and reproducible, while validating quality and provenance. You will work with cross-functional teams to tailor solutions for analytical needs and optimize performance in high-volume environments. This position offers the chance to contribute to advanced research data ecosystems within a federal/public-trust context, shaping how data drives insights.
Compensation / BenefitsMedical, Rx, Dental & Vision Insurance401(k) Retirement PlanParental LeaveTuition Reimbursement, Personal Development, Certifications & Learning OpportunitiesEmployee Assistance ProgramDiscretionary variable incentive bonus
ResponsibilitiesDesign, develop, and maintain data pipelines and workflows for large-scale datasetsImplement ETL processes and automation to ensure readiness and reproducibility of dataPerform data validation, quality control, and document pipelines and provenanceCollaborate with technical and domain teams to troubleshoot data issues and tailor solutionsOptimize performance and resource usage for high-volume data processing
Key requirementsMaster in bioinformatics, computational science, or related fieldFive years of relevant experienceProficiency in Python, R, and workflow toolsExperience with databases and cloud/HPC environmentsExperience working with complex or scientific datasetsMust be able to obtain and maintain a Federal or DoD public trustCollaborativeProblem-solvingAttention to detailPythonRworkflow tools