{"schemaVersion":"jobsearcher.job.v1","id":"5940634f1c02fa7b81013a2b","url":"https://jobsearcher.com/jobs/5940634f1c02fa7b81013a2b","canonicalUrl":"https://jobsearcher.com/jobs/5940634f1c02fa7b81013a2b","title":"Lead Data Engineer","description":"Role : Lead Data EngineerLocation : Maclean VA (100% Onsite)ContractNOTE : NO GCData Platform, Cloud Migration, ETL/ELTSecondary : AI/ML, PythonRole SummaryLead Data Engineer responsible for designing, migrating, and optimizing enterprise-scale data platforms on Microsoft Azure, with strong hands-on ownership of ETL/ELT pipelines, cloud data warehousing, and data integration across heterogeneous sources.Key ResponsibilitiesLead data migration initiatives to Azure, including on prem to cloud and legacy ETL modernization.Design, build, and maintain scalable ETL/ELT pipelines using Azure Data Factory, Databricks (PySpark), and Informatica (PowerCenter / IICS).Implement end to end data ingestion, transformation, and loading into Azure Data Lake Gen2, Azure SQL, Azure Synapse, and Snowflake.Migrate and refactor existing ETL workflows (Informatica / SSIS) to cloud native Azure architectures.Optimize data pipelines for performance, reliability, and cost efficiency.Ensure data quality, governance, security, and compliance during migration and ongoing operations.Collaborate with architects, business stakeholders, and downstream analytics teams.Required SkillsAzure Migration: Hands on experience with Azure Data Lake (Gen1/Gen2), Azure Data Factory, Azure Databricks, Azure Synapse, Azure SQLETL/ELT: Strong expertise in Informatica PowerCenter, Informatica IICS, ETL design, CDC, SCDs, and batch/stream processingProgramming: Python, PySpark, SQL, PL/SQLDatabases & Warehousing: Oracle, SQL Server,Cloud DevOps: Azure DevOps CI/CD pipelines, Git-based version controlData Modeling: Star schemas, dimensional modelingNice to HaveAI/ML experiencePower BI / semantic modelling exposure","company":"Arkhya Tech","rawCompany":"arkhya tech","city":"McLean","state":"VA","isRemote":false,"isActive":false,"createdAt":"2026-07-27T11:30:43.799Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Lead Data Engineer","description":"Role : Lead Data EngineerLocation : Maclean VA (100% Onsite)ContractNOTE : NO GCData Platform, Cloud Migration, ETL/ELTSecondary : AI/ML, PythonRole SummaryLead Data Engineer responsible for designing, migrating, and optimizing enterprise-scale data platforms on Microsoft Azure, with strong hands-on ownership of ETL/ELT pipelines, cloud data warehousing, and data integration across heterogeneous sources.Key ResponsibilitiesLead data migration initiatives to Azure, including on prem to cloud and legacy ETL modernization.Design, build, and maintain scalable ETL/ELT pipelines using Azure Data Factory, Databricks (PySpark), and Informatica (PowerCenter / IICS).Implement end to end data ingestion, transformation, and loading into Azure Data Lake Gen2, Azure SQL, Azure Synapse, and Snowflake.Migrate and refactor existing ETL workflows (Informatica / SSIS) to cloud native Azure architectures.Optimize data pipelines for performance, reliability, and cost efficiency.Ensure data quality, governance, security, and compliance during migration and ongoing operations.Collaborate with architects, business stakeholders, and downstream analytics teams.Required SkillsAzure Migration: Hands on experience with Azure Data Lake (Gen1/Gen2), Azure Data Factory, Azure Databricks, Azure Synapse, Azure SQLETL/ELT: Strong expertise in Informatica PowerCenter, Informatica IICS, ETL design, CDC, SCDs, and batch/stream processingProgramming: Python, PySpark, SQL, PL/SQLDatabases & Warehousing: Oracle, SQL Server,Cloud DevOps: Azure DevOps CI/CD pipelines, Git-based version controlData Modeling: Star schemas, dimensional modelingNice to HaveAI/ML experiencePower BI / semantic modelling exposure","datePosted":"2026-07-27T11:30:43.799Z","dateModified":"2026-07-27T11:30:43.799Z","hiringOrganization":{"@type":"Organization","name":"Arkhya Tech","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"McLean","addressRegion":"VA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"5940634f1c02fa7b81013a2b"},"url":"https://jobsearcher.com/jobs/5940634f1c02fa7b81013a2b"}}