Data Engineer
Data EngineerThe Data Engineer works within the ETL and Operations team to build high quality data pipelines driving analytic solutions from diverse, disparate sources of data.
This role requires a comprehensive understanding of data architecture, data engineering, and data analysis. The ideal candidate is a skilled data engineer with experience creating data products supporting analytic solutions. They are able to identify and implement solutions in a highly technical environment and work as part of a technical, cross-functional team. Strong problem-solving and troubleshooting skills are a must.
This role focuses on designing, developing, and maintaining datasets within AWS. Day-to-day responsibilities include:
Coordinating, building, and managing new data ingests
Architecting updates, fixes, and optimizations across suite of production jobs
Providing feedback on and enacting changes for improvements across the team including both technical and process updates
Learn and assess new technologies for implementation through proof-of-concept projects and testing
Mentoring junior developers
Partnering with data analysts and data scientists to design, build, and deploy aggregation processes
Top Skills:
Hadoop/Hive experience
Very Strong SQL
AWS experience (EMR, Lambda, Glue, Step Functions)
Experience working with large data sets (billions of records per day)
Write and review code
Python scripting (PySpark, control scripts)
Architecting/designing data pipelines
Query and data pipeline optimization
Strong communication and collaboration skills
Ability to help others and mentor junior resources
Nice to have:
Spark
Scala
Git
Shell scripting
Linux