{"schemaVersion":"jobsearcher.job.v1","id":"ad124ba70e5d3aca5105484f","url":"https://jobsearcher.com/jobs/ad124ba70e5d3aca5105484f","canonicalUrl":"https://jobsearcher.com/jobs/ad124ba70e5d3aca5105484f","title":"AWS Data Engineer (Databricks, AWS, Python)","description":"We are seeking an AWS Data Engineer for an onsite contract (W-2) engagement in Houston, TX. You will design, build, and operate large-scale data pipelines on AWS and Databricks, turning raw source data into reliable, well-modelled datasets that power analytics and downstream applications.\n\nResponsibilities:\n- Build and maintain batch and streaming ETL/ELT pipelines using Databricks, PySpark, and Python\n- Develop and optimise Delta Lake tables and medallion (bronze/silver/gold) architectures\n- Engineer data solutions on AWS using S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions, and Kinesis\n- Model data for analytics and reporting; tune Spark jobs and SQL for performance and cost\n- Implement data quality checks, monitoring, alerting, and lineage across pipelines\n- Automate deployments with CI/CD and infrastructure-as-code; manage workflow orchestration\n- Partner with analytics, product, and business teams to deliver production-grade data products\n\nWork mode: This is a 100% onsite role in Houston, TX — no remote or hybrid option.\n\nEmployment: Contract on our W-2 only. Candidates must hold independent work authorisation (no third-party employer or sponsorship arrangement). Rate depends on experience.\n\nRequirements:\n\nMust have: AWS\nMust have: Databricks\nMust have: Python\nStrong PySpark and Spark performance tuning experience\nAdvanced SQL and dimensional/analytical data modelling\nDelta Lake, Unity Catalog, and medallion architecture experience\nAWS data services: S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions, Kinesis\nWorkflow orchestration (Airflow, Databricks Workflows, or similar)\nCI/CD, Git, and infrastructure-as-code (Terraform or CloudFormation)\nData quality, observability, and pipeline monitoring practices\nGood to have: Snowflake, Kafka, dbt, streaming/real-time pipelines\nBachelor's degree in Computer Science, Engineering, or a related field\nWork mode: onsite in Houston, TX (mandatory)\nEmployment: our W-2 only — independent work authorisation required, no third-party or sponsorship","company":"Josh Pros","rawCompany":"josh pros","city":"Houston","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-09-16T10:35:16.398Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AWS Data Engineer (Databricks, AWS, Python)","description":"We are seeking an AWS Data Engineer for an onsite contract (W-2) engagement in Houston, TX. You will design, build, and operate large-scale data pipelines on AWS and Databricks, turning raw source data into reliable, well-modelled datasets that power analytics and downstream applications.\n\nResponsibilities:\n- Build and maintain batch and streaming ETL/ELT pipelines using Databricks, PySpark, and Python\n- Develop and optimise Delta Lake tables and medallion (bronze/silver/gold) architectures\n- Engineer data solutions on AWS using S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions, and Kinesis\n- Model data for analytics and reporting; tune Spark jobs and SQL for performance and cost\n- Implement data quality checks, monitoring, alerting, and lineage across pipelines\n- Automate deployments with CI/CD and infrastructure-as-code; manage workflow orchestration\n- Partner with analytics, product, and business teams to deliver production-grade data products\n\nWork mode: This is a 100% onsite role in Houston, TX — no remote or hybrid option.\n\nEmployment: Contract on our W-2 only. Candidates must hold independent work authorisation (no third-party employer or sponsorship arrangement). Rate depends on experience.\n\nRequirements:\n\nMust have: AWS\nMust have: Databricks\nMust have: Python\nStrong PySpark and Spark performance tuning experience\nAdvanced SQL and dimensional/analytical data modelling\nDelta Lake, Unity Catalog, and medallion architecture experience\nAWS data services: S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions, Kinesis\nWorkflow orchestration (Airflow, Databricks Workflows, or similar)\nCI/CD, Git, and infrastructure-as-code (Terraform or CloudFormation)\nData quality, observability, and pipeline monitoring practices\nGood to have: Snowflake, Kafka, dbt, streaming/real-time pipelines\nBachelor's degree in Computer Science, Engineering, or a related field\nWork mode: onsite in Houston, TX (mandatory)\nEmployment: our W-2 only — independent work authorisation required, no third-party or sponsorship","datePosted":"2026-09-16T10:35:16.398Z","dateModified":"2026-09-16T10:35:16.398Z","hiringOrganization":{"@type":"Organization","name":"Josh Pros","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Houston","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"ad124ba70e5d3aca5105484f"},"url":"https://jobsearcher.com/jobs/ad124ba70e5d3aca5105484f"}}