{"schemaVersion":"jobsearcher.job.v1","id":"37f0fca32404ca876215a044","url":"https://jobsearcher.com/jobs/37f0fca32404ca876215a044","canonicalUrl":"https://jobsearcher.com/jobs/37f0fca32404ca876215a044","title":"Lead Data Engineer","description":"HiLabs is looking for highly motivated and technically strong Lead Data Engineers with deep expertise in Big Data platforms and a passion for building scalable, data-intensive systems. The ideal candidates will have strong hands-on experience in Spark, PySpark, distributed systems, and modern data ecosystem tools, and will enjoy owning complex data platforms from design to production.The individuals who will join the HiLabs engineering team should be continually striving to advance engineering excellence, platform scalability, and data innovation. The mission is to power the next generation of healthcare intelligence platforms through innovation, collaboration, and transparency. You will be a leader and a doer who thrives in a fast-paced, onsite, product-driven environment.ResponsibilitiesLead end-to-end design and development of Big Data and backend platformsArchitect and build scalable, secure, and high-performance distributed data systemsDesign and implement robust Spark and PySpark-based data pipelinesDevelop modular, reusable, and scalable components aligned with business needsManage the complete software development lifecycle from design through deploymentCollaborate closely with product, analytics, and DevOps teamsMentor engineers and drive high coding standards and best practicesPerform code reviews, testing, debugging, and performance optimizationEnsure data platform reliability, scalability, and operational excellenceDrive continuous improvement in data engineering architecture and toolingImplement CI/CD processes and automation for data workflowsEnsure adherence to security, compliance, and high-quality engineering standardsDesired ProfileBachelor’s or Master’s degree in Computer Science, Mathematics, or a related quantitative discipline from a reputed institutionGraduates from Tier-1 institutes such as IIT, NIT, or BITS are strongly preferredUS Master’s degree is a strong advantage; recent graduates from 2023, 2024, or 2025 with strong relevant experience are encouraged to apply6 to 10 years of hands-on experience building and scaling Big Data applications and distributed data platformsStrong expertise in Spark and PySparkExperience with big data ecosystem tools such as Apache Solr, Hive, HBaseExperience with Elasticsearch and MongoDBHands-on experience with workflow orchestration tools such as Airflow or OozieStrong experience with relational databases such as MySQL, SQL Server, or OracleSolid understanding of distributed systems and large-scale architectureExperience with version control tools such as Git or BitbucketExperience with CI/CD tools such as Maven, Jenkins, and JIRAExperience working in Agile software delivery environmentsStrong debugging, performance tuning, and optimization skillsExperience with AWS and or Azure cloud platforms is a plusExperience in healthcare data or analytics domain is preferredStrong communication, collaboration, and leadership skillsAbility to lead by example in design, coding, and problem solvingLocation and Work ModelOnsite role based in Bethesda, MarylandOpen to candidates from DC, Maryland, Baltimore, Virginia, or anywhere in the USCandidates must be open to relocation to BethesdaThe HiLabs StoryHiLabs was born in the halls of Yale University when a cardiologist and an AI expert teamed up to tackle data quality challenges in the US healthcare system. Over the years, HiLabs has built one of the most advanced healthcare data platforms, combining AI, data science, and deep domain expertise to deliver real-world impact across the US healthcare ecosystem.HiLabs TeamMultidisciplinary industry leadersHealthcare domain expertsAI, ML, and data science specialistsProfessionals from the world’s top universities and institutes including Harvard, Yale, Carnegie Mellon, Duke, Georgia Tech, Indian Institute of Management, and Indian Institute of Technology.What We OfferCompetitive base salary, attractive incentive policies, comprehensive benefits including medical coverage for you and your family, 401k, PTOs, employee stock options, relocation support, and an autonomous collaborative working environment. You will work alongside highly qualified professionals from top medical schools, business schools, and engineering institutes, with strong mentorship and long-term growth.","company":"Hilabs","rawCompany":"hilabs","city":"Bethesda","state":"MD","isRemote":false,"isActive":false,"createdAt":"2026-04-12T18:26:00.931Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Lead Data Engineer","description":"HiLabs is looking for highly motivated and technically strong Lead Data Engineers with deep expertise in Big Data platforms and a passion for building scalable, data-intensive systems. The ideal candidates will have strong hands-on experience in Spark, PySpark, distributed systems, and modern data ecosystem tools, and will enjoy owning complex data platforms from design to production.The individuals who will join the HiLabs engineering team should be continually striving to advance engineering excellence, platform scalability, and data innovation. The mission is to power the next generation of healthcare intelligence platforms through innovation, collaboration, and transparency. You will be a leader and a doer who thrives in a fast-paced, onsite, product-driven environment.ResponsibilitiesLead end-to-end design and development of Big Data and backend platformsArchitect and build scalable, secure, and high-performance distributed data systemsDesign and implement robust Spark and PySpark-based data pipelinesDevelop modular, reusable, and scalable components aligned with business needsManage the complete software development lifecycle from design through deploymentCollaborate closely with product, analytics, and DevOps teamsMentor engineers and drive high coding standards and best practicesPerform code reviews, testing, debugging, and performance optimizationEnsure data platform reliability, scalability, and operational excellenceDrive continuous improvement in data engineering architecture and toolingImplement CI/CD processes and automation for data workflowsEnsure adherence to security, compliance, and high-quality engineering standardsDesired ProfileBachelor’s or Master’s degree in Computer Science, Mathematics, or a related quantitative discipline from a reputed institutionGraduates from Tier-1 institutes such as IIT, NIT, or BITS are strongly preferredUS Master’s degree is a strong advantage; recent graduates from 2023, 2024, or 2025 with strong relevant experience are encouraged to apply6 to 10 years of hands-on experience building and scaling Big Data applications and distributed data platformsStrong expertise in Spark and PySparkExperience with big data ecosystem tools such as Apache Solr, Hive, HBaseExperience with Elasticsearch and MongoDBHands-on experience with workflow orchestration tools such as Airflow or OozieStrong experience with relational databases such as MySQL, SQL Server, or OracleSolid understanding of distributed systems and large-scale architectureExperience with version control tools such as Git or BitbucketExperience with CI/CD tools such as Maven, Jenkins, and JIRAExperience working in Agile software delivery environmentsStrong debugging, performance tuning, and optimization skillsExperience with AWS and or Azure cloud platforms is a plusExperience in healthcare data or analytics domain is preferredStrong communication, collaboration, and leadership skillsAbility to lead by example in design, coding, and problem solvingLocation and Work ModelOnsite role based in Bethesda, MarylandOpen to candidates from DC, Maryland, Baltimore, Virginia, or anywhere in the USCandidates must be open to relocation to BethesdaThe HiLabs StoryHiLabs was born in the halls of Yale University when a cardiologist and an AI expert teamed up to tackle data quality challenges in the US healthcare system. Over the years, HiLabs has built one of the most advanced healthcare data platforms, combining AI, data science, and deep domain expertise to deliver real-world impact across the US healthcare ecosystem.HiLabs TeamMultidisciplinary industry leadersHealthcare domain expertsAI, ML, and data science specialistsProfessionals from the world’s top universities and institutes including Harvard, Yale, Carnegie Mellon, Duke, Georgia Tech, Indian Institute of Management, and Indian Institute of Technology.What We OfferCompetitive base salary, attractive incentive policies, comprehensive benefits including medical coverage for you and your family, 401k, PTOs, employee stock options, relocation support, and an autonomous collaborative working environment. You will work alongside highly qualified professionals from top medical schools, business schools, and engineering institutes, with strong mentorship and long-term growth.","datePosted":"2026-04-12T18:26:00.931Z","dateModified":"2026-04-12T18:26:00.931Z","hiringOrganization":{"@type":"Organization","name":"Hilabs","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Bethesda","addressRegion":"MD","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"37f0fca32404ca876215a044"},"url":"https://jobsearcher.com/jobs/37f0fca32404ca876215a044"}}