Data Engineer
About the Role
We're looking for a Data Engineer to join our team supporting the BioSurveillance Integration Platform (BIP) — a mission-critical program delivering data-driven insight for biosurveillance, force health protection, and operational decision-making. You'll design and build the data pipelines that power mission analytics, dashboards, and AI/ML-enabled workflows, working alongside data architects, platform engineers, data scientists, and customer stakeholders to deliver trusted, production-ready data products.
This role follows a phased work arrangement: fully remote during Phase 1 (stand-up and early build), transitioning to hybrid during Phase 2 to support integration, testing, and delivery milestones.
What You'll Do
Design, develop, test, and maintain data pipelines supporting mission analytics, reporting, and AI/ML workflows
Ingest structured, semi-structured, and unstructured data and prepare it for use in enterprise platforms, dashboards, and models
Build and support ETL/ELT processes, transformation logic, schema mapping, and data quality monitoring
Improve pipeline performance, resolve data defects, and document data flows
Implement automated data quality checks, alerting, and observability on production pipelines
Collaborate closely with architects, engineers, data scientists, and mission analysts to deliver usable data products
Produce interface documentation, data dictionaries, pipeline runbooks, and release documentation
What You'll Need
Bachelor's degree, or equivalent relevant experience
Demonstrated experience designing and maintaining data pipelines in a production environment
Proficiency in SQL and a primary data engineering language (Python, Scala, or Java)
Active Secret clearance, with ability to maintain it throughout the period of performance
U.S. Citizenship
Nice to Have
Bachelor's or Master's degree in Computer Science, Data Engineering, or related field
Experience with Palantir Foundry, Maven Smart System, or other federal-scale data platforms
Experience with modern data stack tooling (Apache Spark, Airflow, dbt, Kafka, Databricks, Snowflake)
Cloud data engineering certification (AWS, Azure, or GCP)
Prior experience on federal/DoD programs operating under the Risk Management Framework (RMF)
Familiarity with healthcare or biosurveillance data (HL7 v2, HL7 FHIR, DoD medical reporting formats)
Tools & Platforms
Palantir Foundry, Maven Smart System, Apache Spark, Airflow, dbt, Kafka, AWS/Azure/GCP, Jira, Confluence, Git, Microsoft Office 365
Pay: $50,000.00 - $70,000.00 per year
Benefits:
401(k)
401(k) matching
Health insurance
Life insurance
Referral program
Work Location: Hybrid remote in Omaha, NE 68127