{"schemaVersion":"jobsearcher.job.v1","id":"9f76f7a368f8ca855d79ab52","url":"https://jobsearcher.com/jobs/9f76f7a368f8ca855d79ab52","canonicalUrl":"https://jobsearcher.com/jobs/9f76f7a368f8ca855d79ab52","title":"GCP Data Engineer","description":"Senior Data Engineer (Spark Streaming & GCP)\r\nJob Summary\r\nWe are seeking a highly skilled Senior Data Engineer with strong expertise in Apache Spark, Streaming Technologies, and Google Cloud Platform (GCP) to design, build, and optimize scalable data pipelines supporting analytics, reporting, and machine learning workloads. The ideal candidate will have extensive experience developing both batch and real-time data processing solutions using Spark, Kafka, and cloud-native data services.\r\nRequired Experience\r\n10+ years of overall IT experience\r\n5+ years of recent hands-on GCP experience\r\nStrong experience building enterprise-scale data platforms and streaming architectures\r\nRequired Skills\r\nStrong programming skills in Python and SQL\r\nHands-on expertise with Apache Spark (PySpark, Spark SQL, DataFrames, Spark Streaming)\r\nExperience with Kafka, Pub/Sub, Flink, or other streaming technologies\r\nStrong knowledge of BigQuery, GCS, Data Lakes, and Data Warehousing concepts\r\nExperience designing and developing ETL/ELT pipelines\r\nData modeling and performance optimization experience\r\nExperience with Airflow for workflow orchestration\r\nKnowledge of Snowflake, Redshift, or other cloud data warehouses\r\nExperience implementing data quality, monitoring, and alerting solutions\r\nPreferred Skills\r\nScala or Java development experience\r\nDatabricks experience\r\nDocker and Kubernetes\r\nCI/CD and DevOps practices\r\nExperience supporting ML/Data Science workloads\r\nResponsibilities\r\nDesign and develop scalable batch and real-time data pipelines using Spark and Kafka/PubSub\r\nBuild and optimize BigQuery-based data platforms and lakehouse architectures\r\nDevelop ETL/ELT frameworks for data ingestion, transformation, and delivery\r\nOptimize Spark jobs, SQL queries, and data workflows for performance and cost efficiency\r\nImplement data quality, monitoring, validation, and alerting mechanisms\r\nCollaborate with Data Scientists, Analysts, and Business Teams to deliver reliable data solutions\r\nSupport production deployments and troubleshoot complex data engineering issues\r\nJ-18808-Ljbffr","company":"Socket","rawCompany":"socket","city":"Sunnyvale","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-08T01:53:07.260Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"GCP Data Engineer","description":"Senior Data Engineer (Spark Streaming & GCP)\r\nJob Summary\r\nWe are seeking a highly skilled Senior Data Engineer with strong expertise in Apache Spark, Streaming Technologies, and Google Cloud Platform (GCP) to design, build, and optimize scalable data pipelines supporting analytics, reporting, and machine learning workloads. The ideal candidate will have extensive experience developing both batch and real-time data processing solutions using Spark, Kafka, and cloud-native data services.\r\nRequired Experience\r\n10+ years of overall IT experience\r\n5+ years of recent hands-on GCP experience\r\nStrong experience building enterprise-scale data platforms and streaming architectures\r\nRequired Skills\r\nStrong programming skills in Python and SQL\r\nHands-on expertise with Apache Spark (PySpark, Spark SQL, DataFrames, Spark Streaming)\r\nExperience with Kafka, Pub/Sub, Flink, or other streaming technologies\r\nStrong knowledge of BigQuery, GCS, Data Lakes, and Data Warehousing concepts\r\nExperience designing and developing ETL/ELT pipelines\r\nData modeling and performance optimization experience\r\nExperience with Airflow for workflow orchestration\r\nKnowledge of Snowflake, Redshift, or other cloud data warehouses\r\nExperience implementing data quality, monitoring, and alerting solutions\r\nPreferred Skills\r\nScala or Java development experience\r\nDatabricks experience\r\nDocker and Kubernetes\r\nCI/CD and DevOps practices\r\nExperience supporting ML/Data Science workloads\r\nResponsibilities\r\nDesign and develop scalable batch and real-time data pipelines using Spark and Kafka/PubSub\r\nBuild and optimize BigQuery-based data platforms and lakehouse architectures\r\nDevelop ETL/ELT frameworks for data ingestion, transformation, and delivery\r\nOptimize Spark jobs, SQL queries, and data workflows for performance and cost efficiency\r\nImplement data quality, monitoring, validation, and alerting mechanisms\r\nCollaborate with Data Scientists, Analysts, and Business Teams to deliver reliable data solutions\r\nSupport production deployments and troubleshoot complex data engineering issues\r\nJ-18808-Ljbffr","datePosted":"2026-08-08T01:53:07.260Z","dateModified":"2026-08-08T01:53:07.260Z","hiringOrganization":{"@type":"Organization","name":"Socket","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Sunnyvale","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"9f76f7a368f8ca855d79ab52"},"url":"https://jobsearcher.com/jobs/9f76f7a368f8ca855d79ab52"}}