{"schemaVersion":"jobsearcher.job.v1","id":"465af8d97683819e94ab05c7","url":"https://jobsearcher.com/jobs/465af8d97683819e94ab05c7","canonicalUrl":"https://jobsearcher.com/jobs/465af8d97683819e94ab05c7","title":"Sensitive-Data Engineer","description":"Active Top Secret/SCI Clearance with Polygraph (REQUIRED)\r\nAre you passionate about harnessing data to solve some of the nation's most critical challenges? Do you thrive on innovation, collaboration, and building resilient solutions in complex environments?\r\nJoin a high-impact team at the forefront of national security, where your work directly supports mission success. We're seeking a Data Engineer with a rare mix of curiosity, craftsmanship, and commitment to excellence. In this role, you'll design and optimize secure, scalable data pipelines while working alongside elite engineers, mission partners, and data experts to unlock actionable insights from diverse datasets.\r\nRequirements\r\nEngineer robust, secure, and scalable data pipelines using Apache Spark , Apache Hudi , AWS EMR , and Kubernetes\r\nMaintain data provenance and access controls to ensure full lineage and auditability of mission-critical datasets\r\nClean, transform, and condition data using tools such as dbt , Apache NiFi , or Pandas\r\nBuild and orchestrate repeatable ETL workflows using Apache Airflow , Dagster , or Prefect\r\nDevelop API connectors for ingesting structured and unstructured data sources\r\nCollaborate with data stewards, architects, and mission teams to align on data standards , quality, and integrity\r\nProvide advanced database administration for Oracle , PostgreSQL , MongoDB , Elasticsearch , and others\r\nIngest and analyze streaming data using tools like Apache Kafka , AWS Kinesis , or Apache Flink\r\nPerform real-time and batch processing on large datasets in secure cloud environments (e.g., AWS GovCloud , C2S )\r\nImplement and monitor data quality and validation checks using tools such as Great Expectations or Deequ\r\nWork across agile teams using DevSecOps practices to build resilient full-stack solutions with Python , Java , or Scala\r\nRequired Skills\r\nExperience building and maintaining data pipelines using Apache Spark , Airflow , NiFi , or dbt\r\nProficiency in Python , SQL , and one or more of: Java , Scala\r\nStrong understanding of cloud services (especially AWS and GovCloud ), including S3 , EC2 , Lambda , EMR , Glue , Redshift , or Snowflake\r\nHands-on experience with streaming frameworks such as Apache Kafka , Kafka Connect , or Flink\r\nFamiliarity with data lakehouse formats (e.g., Apache Hudi, Delta Lake, or Iceberg)\r\nExperience with NoSQL and RDBMS technologies such as MongoDB , DynamoDB , PostgreSQL , or MySQL\r\nAbility to implement and maintain data validation frameworks (e.g., Great Expectations , Deequ )\r\nComfortable working in Linux/Unix environments, using bash scripting , Git, and CI/CD tools\r\nKnowledge of containerization and orchestration tools like Docker and Kubernetes\r\nCollaborative mindset with experience working in Agile/Scrum environments using Jira , Confluence , and Git-based workflows\r\nJ-18808-Ljbffr","company":"Equilibrium Technologies","rawCompany":"equilibrium technologies","city":"Chantilly","state":"VA","isRemote":false,"isActive":false,"createdAt":"2026-07-16T01:29:35.942Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Sensitive-Data Engineer","description":"Active Top Secret/SCI Clearance with Polygraph (REQUIRED)\r\nAre you passionate about harnessing data to solve some of the nation's most critical challenges? Do you thrive on innovation, collaboration, and building resilient solutions in complex environments?\r\nJoin a high-impact team at the forefront of national security, where your work directly supports mission success. We're seeking a Data Engineer with a rare mix of curiosity, craftsmanship, and commitment to excellence. In this role, you'll design and optimize secure, scalable data pipelines while working alongside elite engineers, mission partners, and data experts to unlock actionable insights from diverse datasets.\r\nRequirements\r\nEngineer robust, secure, and scalable data pipelines using Apache Spark , Apache Hudi , AWS EMR , and Kubernetes\r\nMaintain data provenance and access controls to ensure full lineage and auditability of mission-critical datasets\r\nClean, transform, and condition data using tools such as dbt , Apache NiFi , or Pandas\r\nBuild and orchestrate repeatable ETL workflows using Apache Airflow , Dagster , or Prefect\r\nDevelop API connectors for ingesting structured and unstructured data sources\r\nCollaborate with data stewards, architects, and mission teams to align on data standards , quality, and integrity\r\nProvide advanced database administration for Oracle , PostgreSQL , MongoDB , Elasticsearch , and others\r\nIngest and analyze streaming data using tools like Apache Kafka , AWS Kinesis , or Apache Flink\r\nPerform real-time and batch processing on large datasets in secure cloud environments (e.g., AWS GovCloud , C2S )\r\nImplement and monitor data quality and validation checks using tools such as Great Expectations or Deequ\r\nWork across agile teams using DevSecOps practices to build resilient full-stack solutions with Python , Java , or Scala\r\nRequired Skills\r\nExperience building and maintaining data pipelines using Apache Spark , Airflow , NiFi , or dbt\r\nProficiency in Python , SQL , and one or more of: Java , Scala\r\nStrong understanding of cloud services (especially AWS and GovCloud ), including S3 , EC2 , Lambda , EMR , Glue , Redshift , or Snowflake\r\nHands-on experience with streaming frameworks such as Apache Kafka , Kafka Connect , or Flink\r\nFamiliarity with data lakehouse formats (e.g., Apache Hudi, Delta Lake, or Iceberg)\r\nExperience with NoSQL and RDBMS technologies such as MongoDB , DynamoDB , PostgreSQL , or MySQL\r\nAbility to implement and maintain data validation frameworks (e.g., Great Expectations , Deequ )\r\nComfortable working in Linux/Unix environments, using bash scripting , Git, and CI/CD tools\r\nKnowledge of containerization and orchestration tools like Docker and Kubernetes\r\nCollaborative mindset with experience working in Agile/Scrum environments using Jira , Confluence , and Git-based workflows\r\nJ-18808-Ljbffr","datePosted":"2026-07-16T01:29:35.942Z","dateModified":"2026-07-16T01:29:35.942Z","hiringOrganization":{"@type":"Organization","name":"Equilibrium Technologies","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Chantilly","addressRegion":"VA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"465af8d97683819e94ab05c7"},"url":"https://jobsearcher.com/jobs/465af8d97683819e94ab05c7"}}