{"schemaVersion":"jobsearcher.job.v1","id":"8ea2aaa066c8ed20d2bc850b","url":"https://jobsearcher.com/jobs/8ea2aaa066c8ed20d2bc850b","canonicalUrl":"https://jobsearcher.com/jobs/8ea2aaa066c8ed20d2bc850b","title":"Data Engineer","description":"Position Purpose:\nWe are seeking a Data Engineer to support the development of cloud-native data solutions that improve operational efficiency, support regulatory reporting, and drive actionable insight for our commercial and government healthcare clients. This role is responsible for building, maintaining, and optimizing scalable data pipelines and validation workflows across complex datasets—enabling downstream analytics, application logic, and system modernization initiatives.\nPrimary Duties & Responsibilities:\nDesign, develop, and maintain ETL/ELT pipelines using structured and semi-structured data from relational databases, flat files, APIs, and cloud data sources.\nCollaborate with backend and architecture teams to define data transformation flows aligned with dashboard and reporting application needs.\nDesign and optimize data schemas to ensure performance, integrity, and compatibility with reporting requirements.\nDevelop and maintain efficient, testable, and reusable data processing scripts using Python, SQL, and cloud-native tools.\nCollaborate with DevOps, analysts, and application developers to align pipelines with system architecture, storage strategy, and reporting needs.\nImplement data quality and validation checks and document data lineage and pipeline logic for audit and reuse.\nTroubleshoot performance issues in data jobs and support data pipeline operations across environments (DEV, VAL, PROD).\nContribute CI/CD workflows and automation strategies to promote rapid iteration and secure deployment of data services.\nAssist in developing or maintaining data documentation, including metadata, data dictionaries, and technical user guides.\nStay up to date with emerging technologies, techniques, and trends to inform product development and decision-making.\nMinimum Qualifications:\nBachelor's degree in computer science, engineering, statistics, or related field.\n4+ years of experience in a data engineering, data pipeline, or ETL/ELT development role.\nStrong proficiency in SQL, data transformation logic, and performance tuning for large datasets.\nProficiency with Python and libraries such as Pandas, PySpark, or Numpy.\nExperience with modern version control and CI/CD practices (e.g., GitHub Actions, Jenkins).\nUnderstanding of distributed computing solutions for data processing (e.g. AWS Glue, AWS EMR, Apache Hadoop, Apache Spark)\nExperience with data pipeline orchestration frameworks or serverless tools (e.g., AWS StepFunctions, AWS Glue Workflows, and/or AWS Lambda).\nPreferred Qualifications:\nExperience developing solutions in AWS cloud environments (S3, Lambda, Aurora, SNS, CloudFormation/CDK).\nExperience with Snowflake, Amazon RDS, or other cloud-native data warehouses.\nFamiliarity with modern data transformation tools (dbt, Dataform, or equivalent) and best practices for modular, tested SQL transformations within an ELT architecture.\nExperience supporting healthcare data systems or CMS data environments\nAWS certifications are a plus.","company":"Magpie Health Analytics","rawCompany":"magpie health analytics","city":"Remote","state":"OR","isRemote":false,"isActive":false,"createdAt":"2026-08-05T15:12:52.939Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer","description":"Position Purpose:\nWe are seeking a Data Engineer to support the development of cloud-native data solutions that improve operational efficiency, support regulatory reporting, and drive actionable insight for our commercial and government healthcare clients. This role is responsible for building, maintaining, and optimizing scalable data pipelines and validation workflows across complex datasets—enabling downstream analytics, application logic, and system modernization initiatives.\nPrimary Duties & Responsibilities:\nDesign, develop, and maintain ETL/ELT pipelines using structured and semi-structured data from relational databases, flat files, APIs, and cloud data sources.\nCollaborate with backend and architecture teams to define data transformation flows aligned with dashboard and reporting application needs.\nDesign and optimize data schemas to ensure performance, integrity, and compatibility with reporting requirements.\nDevelop and maintain efficient, testable, and reusable data processing scripts using Python, SQL, and cloud-native tools.\nCollaborate with DevOps, analysts, and application developers to align pipelines with system architecture, storage strategy, and reporting needs.\nImplement data quality and validation checks and document data lineage and pipeline logic for audit and reuse.\nTroubleshoot performance issues in data jobs and support data pipeline operations across environments (DEV, VAL, PROD).\nContribute CI/CD workflows and automation strategies to promote rapid iteration and secure deployment of data services.\nAssist in developing or maintaining data documentation, including metadata, data dictionaries, and technical user guides.\nStay up to date with emerging technologies, techniques, and trends to inform product development and decision-making.\nMinimum Qualifications:\nBachelor's degree in computer science, engineering, statistics, or related field.\n4+ years of experience in a data engineering, data pipeline, or ETL/ELT development role.\nStrong proficiency in SQL, data transformation logic, and performance tuning for large datasets.\nProficiency with Python and libraries such as Pandas, PySpark, or Numpy.\nExperience with modern version control and CI/CD practices (e.g., GitHub Actions, Jenkins).\nUnderstanding of distributed computing solutions for data processing (e.g. AWS Glue, AWS EMR, Apache Hadoop, Apache Spark)\nExperience with data pipeline orchestration frameworks or serverless tools (e.g., AWS StepFunctions, AWS Glue Workflows, and/or AWS Lambda).\nPreferred Qualifications:\nExperience developing solutions in AWS cloud environments (S3, Lambda, Aurora, SNS, CloudFormation/CDK).\nExperience with Snowflake, Amazon RDS, or other cloud-native data warehouses.\nFamiliarity with modern data transformation tools (dbt, Dataform, or equivalent) and best practices for modular, tested SQL transformations within an ELT architecture.\nExperience supporting healthcare data systems or CMS data environments\nAWS certifications are a plus.","datePosted":"2026-08-05T15:12:52.939Z","dateModified":"2026-08-05T15:12:52.939Z","hiringOrganization":{"@type":"Organization","name":"Magpie Health Analytics","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Remote","addressRegion":"OR","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"8ea2aaa066c8ed20d2bc850b"},"url":"https://jobsearcher.com/jobs/8ea2aaa066c8ed20d2bc850b"}}