{"schemaVersion":"jobsearcher.job.v1","id":"e71d4c4d81d0c63f869cf196","url":"https://jobsearcher.com/jobs/e71d4c4d81d0c63f869cf196","canonicalUrl":"https://jobsearcher.com/jobs/e71d4c4d81d0c63f869cf196","title":"Database Engineer (Remote)","description":"Overview:\n\nGovCIO is currently hiring for Data Engineer to design and build enriched datasets using medallion architecture (bronze/silver/gold) in Databricks, sourcing data from the VA Corporate Data Warehouse (CDW) and other internal and external systems. The datasets you build will feed directly into Power BI semantic models used by stakeholders across the organization. This role works closely with clients to gather requirements, track work in JIRA, and deliver high-quality, well-documented data pipelines. This is a fully remote position.\n\nResponsibilities:\n\nOur Data Engineering team supports multiple offices and programs across the Department of Veterans Affairs by building enriched, analytics-ready datasets that power clinical and operational decision-making across the Veterans Health Administration (VHA). We work at the intersection of enterprise data warehousing and modern cloud analytics, turning complex healthcare data into reliable, reusable data products.\n\nWhat You'll Do\n\nDesign, build, and maintain enriched datasets following medallion architecture (bronze, silver, gold layers) in Databricks\nDevelop, orchestrate, and monitor Databricks jobs and pipelines\nWrite efficient, well-structured SQL to transform and model healthcare data\nPrepare and optimize gold-layer datasets for consumption by Power BI semantic models\nManage code and version control through GitHub, following team branching and review practices\nLog, track, and update work items in JIRA\nPartner directly with clients and stakeholders to understand requirements, clarify data definitions, and validate outputs\nTroubleshoot data quality issues and ensure accuracy and consistency across pipelines\nWrite ad hoc queries to support client requests, investigations, and one-off analyses\nPerform data validation and quality assurance checks on pipelines and enriched datasets to ensure accuracy, completeness, and consistency\nDocument data lineage, transformation logic, and business rules for enriched datasets\nCollaborate with Power BI developers and data scientists on the broader team to ensure datasets meet downstream reporting and analytical needs\nQualifications:\n\nRequired Skills and Experience\n\nBachelors degree and 12+ yrs of experience (or commensurate experience)\nHands-on experience building data pipelines and transformations in Databricks (PySpark and/or SQL)\nStrong SQL skills, including complex joins, aggregations, and performance tuning\nExperience working with Power BI, including preparing data for semantic models\nFamiliarity with GitHub for version control and collaborative development\nExperience using JIRA or similar tools to track and manage work\nStrong communication skills and comfort working directly with clients/stakeholders\nUnderstanding of medallion (bronze/silver/gold) data architecture principles\n\nPreferred Skills and Experience\n\nPrior experience working within the VA or VHA\nDeep understanding of the VA Corporate Data Warehouse (CDW)\nFamiliarity with VistA/CPRS\nFamiliarity with the Federal EHR (Oracle Health)\nExperience with healthcare data domains (e.g., appointments, consults/referrals, orders, TIU notes, visits)\n\nWhat We're Looking For\n\nSomeone who can work independently on complex data problems, communicate clearly with non-technical clients, and take ownership of datasets from raw source through to a polished, analytics-ready product. Healthcare and federal data experience is a strong plus given the sensitivity and complexity of the data we work with.\n\nPosted Salary Range: USD $130,000.00 - USD $160,000.00 /Yr.","company":"GovCIO","rawCompany":"govcio","isRemote":true,"isActive":false,"createdAt":"2026-09-16T10:35:16.439Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Database Engineer (Remote)","description":"Overview:\n\nGovCIO is currently hiring for Data Engineer to design and build enriched datasets using medallion architecture (bronze/silver/gold) in Databricks, sourcing data from the VA Corporate Data Warehouse (CDW) and other internal and external systems. The datasets you build will feed directly into Power BI semantic models used by stakeholders across the organization. This role works closely with clients to gather requirements, track work in JIRA, and deliver high-quality, well-documented data pipelines. This is a fully remote position.\n\nResponsibilities:\n\nOur Data Engineering team supports multiple offices and programs across the Department of Veterans Affairs by building enriched, analytics-ready datasets that power clinical and operational decision-making across the Veterans Health Administration (VHA). We work at the intersection of enterprise data warehousing and modern cloud analytics, turning complex healthcare data into reliable, reusable data products.\n\nWhat You'll Do\n\nDesign, build, and maintain enriched datasets following medallion architecture (bronze, silver, gold layers) in Databricks\nDevelop, orchestrate, and monitor Databricks jobs and pipelines\nWrite efficient, well-structured SQL to transform and model healthcare data\nPrepare and optimize gold-layer datasets for consumption by Power BI semantic models\nManage code and version control through GitHub, following team branching and review practices\nLog, track, and update work items in JIRA\nPartner directly with clients and stakeholders to understand requirements, clarify data definitions, and validate outputs\nTroubleshoot data quality issues and ensure accuracy and consistency across pipelines\nWrite ad hoc queries to support client requests, investigations, and one-off analyses\nPerform data validation and quality assurance checks on pipelines and enriched datasets to ensure accuracy, completeness, and consistency\nDocument data lineage, transformation logic, and business rules for enriched datasets\nCollaborate with Power BI developers and data scientists on the broader team to ensure datasets meet downstream reporting and analytical needs\nQualifications:\n\nRequired Skills and Experience\n\nBachelors degree and 12+ yrs of experience (or commensurate experience)\nHands-on experience building data pipelines and transformations in Databricks (PySpark and/or SQL)\nStrong SQL skills, including complex joins, aggregations, and performance tuning\nExperience working with Power BI, including preparing data for semantic models\nFamiliarity with GitHub for version control and collaborative development\nExperience using JIRA or similar tools to track and manage work\nStrong communication skills and comfort working directly with clients/stakeholders\nUnderstanding of medallion (bronze/silver/gold) data architecture principles\n\nPreferred Skills and Experience\n\nPrior experience working within the VA or VHA\nDeep understanding of the VA Corporate Data Warehouse (CDW)\nFamiliarity with VistA/CPRS\nFamiliarity with the Federal EHR (Oracle Health)\nExperience with healthcare data domains (e.g., appointments, consults/referrals, orders, TIU notes, visits)\n\nWhat We're Looking For\n\nSomeone who can work independently on complex data problems, communicate clearly with non-technical clients, and take ownership of datasets from raw source through to a polished, analytics-ready product. Healthcare and federal data experience is a strong plus given the sensitivity and complexity of the data we work with.\n\nPosted Salary Range: USD $130,000.00 - USD $160,000.00 /Yr.","datePosted":"2026-09-16T10:35:16.439Z","dateModified":"2026-09-16T10:35:16.439Z","hiringOrganization":{"@type":"Organization","name":"GovCIO","sameAs":"https://jobsearcher.com"},"jobLocationType":"TELECOMMUTE","applicantLocationRequirements":{"@type":"Country","name":"US"},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"e71d4c4d81d0c63f869cf196"},"url":"https://jobsearcher.com/jobs/e71d4c4d81d0c63f869cf196"}}