{"schemaVersion":"jobsearcher.job.v1","id":"c9e5062feedc4a78f852ce05","url":"https://jobsearcher.com/jobs/c9e5062feedc4a78f852ce05","canonicalUrl":"https://jobsearcher.com/jobs/c9e5062feedc4a78f852ce05","title":"Data Engineer— Databricks, PySpark & Python","description":"We're looking for a Senior Data Engineers with strong Databricks, PySpark, and Python expertise to help modernize a large-scale data platform and build AI-ready infrastructure. This is a hands-on contract role for someone who thrives in fast-moving environments, can work through ambiguity, and is comfortable owning solutions from design through production. You'll help build scalable pipelines, shape data models, and contribute to the evolution of a modern cloud data platform supporting analytics, reporting, and future AI/ML use cases — as the platform evolves from a legacy analytics setup into a modern hybrid architecture built on Databricks and cloud object storage, with BigQuery continuing to support reporting and BI.What You'll DoDesign and implement scalable ETL and ELT pipelines using PySpark on DatabricksBuild ingestion frameworks for structured and semi-structured data from multiple sourcesDevelop high-performance data transformations that are maintainable, testable, and production-readyIntegrate pipelines and workflows across cloud storage and BigQuery-based reporting environmentsContribute to data modeling decisions, including schemas, transformations, and domain structuresSupport schema evolution and design choices that enable long-term platform scalabilityTranslate ambiguous business and technical requirements into clear, actionable engineering solutionsCollaborate closely with data engineers, DevOps, and business stakeholders to align priorities and unblock deliveryImprove reliability, observability, performance, and operational quality across pipelines and platform componentsDebug, optimize, and continuously improve existing data workflows and engineering practicesWhat You BringStrong hands-on experience with Databricks in production environmentsAdvanced proficiency in PythonDeep experience with Apache Spark and PySparkStrong experience building cloud-based data pipelines and large-scale data processing systemsExcellent SQL skills and strong database fundamentalsExperience designing and maintaining scalable ETL or ELT workflowsAbility to write clean, maintainable, and testable production-grade codeStrong problem-solving skills, with the ability to break down complex technical challenges independentlyComfort working in ambiguous environments and driving execution with limited oversightStrong communication skills for collaborating across technical and non-technical stakeholdersAvailability with meaningful overlap with US time zones","company":"Toptal","rawCompany":"toptal","city":"Northern Cambria","state":"PA","isRemote":false,"isActive":false,"createdAt":"2026-08-04T10:39:18.883Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Engineer— Databricks, PySpark & Python","description":"We're looking for a Senior Data Engineers with strong Databricks, PySpark, and Python expertise to help modernize a large-scale data platform and build AI-ready infrastructure. This is a hands-on contract role for someone who thrives in fast-moving environments, can work through ambiguity, and is comfortable owning solutions from design through production. You'll help build scalable pipelines, shape data models, and contribute to the evolution of a modern cloud data platform supporting analytics, reporting, and future AI/ML use cases — as the platform evolves from a legacy analytics setup into a modern hybrid architecture built on Databricks and cloud object storage, with BigQuery continuing to support reporting and BI.What You'll DoDesign and implement scalable ETL and ELT pipelines using PySpark on DatabricksBuild ingestion frameworks for structured and semi-structured data from multiple sourcesDevelop high-performance data transformations that are maintainable, testable, and production-readyIntegrate pipelines and workflows across cloud storage and BigQuery-based reporting environmentsContribute to data modeling decisions, including schemas, transformations, and domain structuresSupport schema evolution and design choices that enable long-term platform scalabilityTranslate ambiguous business and technical requirements into clear, actionable engineering solutionsCollaborate closely with data engineers, DevOps, and business stakeholders to align priorities and unblock deliveryImprove reliability, observability, performance, and operational quality across pipelines and platform componentsDebug, optimize, and continuously improve existing data workflows and engineering practicesWhat You BringStrong hands-on experience with Databricks in production environmentsAdvanced proficiency in PythonDeep experience with Apache Spark and PySparkStrong experience building cloud-based data pipelines and large-scale data processing systemsExcellent SQL skills and strong database fundamentalsExperience designing and maintaining scalable ETL or ELT workflowsAbility to write clean, maintainable, and testable production-grade codeStrong problem-solving skills, with the ability to break down complex technical challenges independentlyComfort working in ambiguous environments and driving execution with limited oversightStrong communication skills for collaborating across technical and non-technical stakeholdersAvailability with meaningful overlap with US time zones","datePosted":"2026-08-04T10:39:18.883Z","dateModified":"2026-08-04T10:39:18.883Z","hiringOrganization":{"@type":"Organization","name":"Toptal","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Northern Cambria","addressRegion":"PA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"c9e5062feedc4a78f852ce05"},"url":"https://jobsearcher.com/jobs/c9e5062feedc4a78f852ce05"}}