Databricks Data Engineer
Remote
Contract (7 months 25 days)
Published 17 hours ago
AWS certifications
data governance
aws cloud
data modeling
performance optimization
pyspark
Troubleshooting & Debugging
SQL
DevOps & CI/CD
ETL/ELT pipelines
We are seeking a Databricks Engineer to lead the design and implementation of a scalable Sales Data Platform as part of the OneData initiative on Databricks running on AWS.
The role is hands-on and architecture-driven, focused on data ingestion, transformation, modeling, and optimization using modern lakehouse patterns.
The architect will work closely with data engineers, source system teams, and downstream consumers to deliver high-quality, governed, and performance-optimized sales datasets.
Key Responsibilities:
Architecture & Design:
Define end-to-end lakehouse architecture on Databricks (AWS) for Sales data domains
Design medallion architecture (Bronze / Silver / Gold) aligned with OneData standards
Establish data modeling standards for Sales facts, dimensions, hierarchies, and aggregations
Define scalable ingestion patterns for batch and incremental loads
Drive performance, scalability, and cost optimization best practices
Data Engineering & Implementation:
Build and guide development of PySpark-based data pipelines in Databricks
Implement Delta Lake features:
ACID transactions
Schema evolution & enforcement
Time travel & versioning
Design and optimize large-scale joins, aggregations, and window functions
Implement CDC and incremental processing using watermarking and change detection
Ensure idempotent, restartable, and fault-tolerant pipelines
AWS & Platform Integration:
Architect solutions using AWS services:
Amazon S3 (data lake storage)
IAM (security & access control)
AWS Glue / Glue Catalog
CloudWatch (monitoring & logging)
Optimize Databricks cluster configurations (job vs all-purpose clusters)
Implement secrets management and secure connectivity patterns
Data Quality, Governance & Reliability:
Define and implement data quality checks and validations
Ensure data lineage and metadata capture
Implement error handling, auditing, and reconciliation frameworks
Support data governance and access control requirements
DevOps & Operational Excellence:
Implement CI/CD pipelines for Databricks notebooks and jobs
Enforce code versioning, reviews, and deployment standards
Design monitoring, alerting, and SLA tracking for pipelines
Support production stabilization and performance tuning
Collaboration & Leadership:
Act as technical lead / mentor for Databricks data engineers
Collaborate with:
Source system teams (Sales, CRM, ERP)
Data consumers (analytics, downstream apps)
Cloud/platform teams
Translate business requirements into robust technical designs
Required Skills & Experience:
Core Technical Skills:
8+ years in Data Engineering / Data Architecture roles
4+ years hands-on experience with Databricks
Strong expertise in PySpark & Spark SQL
Strong expertise in dbt
Deep experience with Delta Lake
Strong knowledge of AWS cloud services (S3, IAM, Glue, CloudWatch)
Data Engineering Expertise:
Sales data domain experience (orders, revenue, pricing, customers, products)
Strong understanding of:
Fact & dimension modeling
Slowly Changing Dimensions (SCD Type 1 / 2)
Large-scale data processing patterns
Experience handling high-volume, high-velocity datasets
Platform & Operational Skills:
Databricks job orchestration and scheduling
Cluster sizing and performance tuning
CI/CD for data platforms
Strong troubleshooting and debugging skills
Nice-to-Have:
Experience with enterprise OneData / Data Mesh programs
Exposure to real-time or near-real-time ingestion patterns
Experience integrating CRM / Sales systems (e.g., Salesforce, SAP Sales data)
AWS certifications or Databricks certifications
The pay range that the employer in good faith reasonably expects to pay for this position is $60.10/hour - $93.90/hour. Our benefits include medical, dental, vision and retirement benefits. Applications will be accepted on an ongoing basis.
Tundra Technical Solutions is among North America’s leading providers of Staffing and Consulting Services. Our success and our clients’ success are built on a foundation of service excellence. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. Qualified applicants with arrest or conviction records will be considered for employment in accordance with applicable law, including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Unincorporated LA County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: client provided property, including hardware (both of which may include data) entrusted to you from theft, loss or damage; return all portable client computer hardware in your possession (including the data contained therein) upon completion of the assignment, and; maintain the confidentiality of client proprietary, confidential, or non-public information. In addition, job duties require access to secure and protected client information technology systems and related data security obligations.