{"schemaVersion":"jobsearcher.job.v1","id":"c47c217dd6b5d927f72c520e","url":"https://jobsearcher.com/jobs/c47c217dd6b5d927f72c520e","canonicalUrl":"https://jobsearcher.com/jobs/c47c217dd6b5d927f72c520e","title":"Databricks Data Architect","description":"Databricks Data ArchitectLocation: West Houston, TXWork Arrangement: Hybrid – 3 days onsite per weekContract: 12 monthsWe are seeking a Databricks Data Architect to design and lead modern cloud data solutions using Databricks and AWS. This is a hands-on architecture role focused on Lakehouse architecture, streaming, performance, governance, and cloud optimization.Key ResponsibilitiesDesign enterprise Databricks Lakehouse architectures using Medallion Architecture.Architect batch and real-time data solutions using PySpark, Structured Streaming, Kafka, Delta Lake, and AWS.Design solutions for late-arriving data, event-time processing, watermarking, CDC, CDF, and streaming recovery.Leverage modern Databricks capabilities including Lakeflow/LDP, Serverless Compute, Unity Catalog, Liquid Clustering, Z-Ordering, and Photon.Evaluate emerging capabilities including Zerobus, Lakebase-to-Lakehouse synchronization, and LTAP.Design and optimize AWS data platforms using S3, Glue, EMR, Kinesis/MSK, Redshift, Lambda, IAM, and CloudWatch.Drive performance and cost optimization across Databricks and AWS workloads.Establish standards for data governance, security, lineage, CI/CD, and infrastructure as code.Lead technical design discussions and mentor Data Engineers.Required Experience8+ years of experience in Data Engineering, Data Architecture, or Cloud Data Platforms.Strong hands-on experience with Databricks and AWS.Expert knowledge of Spark/PySpark and Structured Streaming.Strong experience with Kafka and/or Kinesis/MSK.Deep understanding of Delta Lake and Medallion Architecture.Experience with Unity Catalog, CDF/CDC, Liquid Clustering, and Z-Ordering.Strong understanding of Databricks performance and cost optimization.Excellent communication and stakeholder management skills.Preferred: Databricks certification, AWS certification, and experience with newer Databricks capabilities such as Lakeflow, Zerobus, Lakebase, Serverless Compute, and LTAP.Local Houston candidates preferred.","company":"Sgi","rawCompany":"sgi","city":"Houston","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-08-27T09:02:14.299Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Databricks Data Architect","description":"Databricks Data ArchitectLocation: West Houston, TXWork Arrangement: Hybrid – 3 days onsite per weekContract: 12 monthsWe are seeking a Databricks Data Architect to design and lead modern cloud data solutions using Databricks and AWS. This is a hands-on architecture role focused on Lakehouse architecture, streaming, performance, governance, and cloud optimization.Key ResponsibilitiesDesign enterprise Databricks Lakehouse architectures using Medallion Architecture.Architect batch and real-time data solutions using PySpark, Structured Streaming, Kafka, Delta Lake, and AWS.Design solutions for late-arriving data, event-time processing, watermarking, CDC, CDF, and streaming recovery.Leverage modern Databricks capabilities including Lakeflow/LDP, Serverless Compute, Unity Catalog, Liquid Clustering, Z-Ordering, and Photon.Evaluate emerging capabilities including Zerobus, Lakebase-to-Lakehouse synchronization, and LTAP.Design and optimize AWS data platforms using S3, Glue, EMR, Kinesis/MSK, Redshift, Lambda, IAM, and CloudWatch.Drive performance and cost optimization across Databricks and AWS workloads.Establish standards for data governance, security, lineage, CI/CD, and infrastructure as code.Lead technical design discussions and mentor Data Engineers.Required Experience8+ years of experience in Data Engineering, Data Architecture, or Cloud Data Platforms.Strong hands-on experience with Databricks and AWS.Expert knowledge of Spark/PySpark and Structured Streaming.Strong experience with Kafka and/or Kinesis/MSK.Deep understanding of Delta Lake and Medallion Architecture.Experience with Unity Catalog, CDF/CDC, Liquid Clustering, and Z-Ordering.Strong understanding of Databricks performance and cost optimization.Excellent communication and stakeholder management skills.Preferred: Databricks certification, AWS certification, and experience with newer Databricks capabilities such as Lakeflow, Zerobus, Lakebase, Serverless Compute, and LTAP.Local Houston candidates preferred.","datePosted":"2026-08-27T09:02:14.299Z","dateModified":"2026-08-27T09:02:14.299Z","hiringOrganization":{"@type":"Organization","name":"Sgi","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Houston","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"c47c217dd6b5d927f72c520e"},"url":"https://jobsearcher.com/jobs/c47c217dd6b5d927f72c520e"}}