{"schemaVersion":"jobsearcher.job.v1","id":"7ede4c59eb461e6b8a68c2ef","url":"https://jobsearcher.com/jobs/7ede4c59eb461e6b8a68c2ef","canonicalUrl":"https://jobsearcher.com/jobs/7ede4c59eb461e6b8a68c2ef","title":"Java Spark Engineer","description":"Primary Responsibilities \r • Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java) \r • Lead design of batch and streaming ETL/ELT systems handling large data volumes \r • Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction \r • Set coding standards and lead code/design reviews across the team \r • Drive technical decisions on data architecture, storage formats, and pipeline orchestration \r • Mentor mid-level and junior engineers; act as a technical escalation point \r • Partner with product, analytics, and platform teams to translate requirements into scalable systems \r • Own production reliability — on-call ownership, incident response, root-cause analysis for pipeline failures \r • Evaluate and introduce new tools/frameworks where they improve the system \r • Contribute to capacity planning and cost optimization for cluster infrastructure \r Required Qualifications \r • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field \r • 7+ years of professional Java development experience \r • 5+ years hands-on experience with Apache Spark in production environments \r • Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management \r • Proven track record designing systems processing terabyte+ scale data \r • Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg) \r • Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark \r • Proficiency with Kafka \r • Strong grasp of CI/CD, containerization, and infrastructure-as-code practices \r Preferred Qualifications \r • Experience with Flink or other stream-processing frameworks \r • Familiarity with data governance, lineage, and quality frameworks \r • Experience with workflow orchestration at scale  \r • Background in system design for multi-tenant or multi-region data platforms \r • Prior experience leading a team or acting as a technical lead","company":"2t Consulting","rawCompany":"2t consulting","city":"Colonia","state":"NJ","isRemote":false,"isActive":true,"createdAt":"2026-08-13T23:21:08.252Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Java Spark Engineer","description":"Primary Responsibilities \r • Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java) \r • Lead design of batch and streaming ETL/ELT systems handling large data volumes \r • Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction \r • Set coding standards and lead code/design reviews across the team \r • Drive technical decisions on data architecture, storage formats, and pipeline orchestration \r • Mentor mid-level and junior engineers; act as a technical escalation point \r • Partner with product, analytics, and platform teams to translate requirements into scalable systems \r • Own production reliability — on-call ownership, incident response, root-cause analysis for pipeline failures \r • Evaluate and introduce new tools/frameworks where they improve the system \r • Contribute to capacity planning and cost optimization for cluster infrastructure \r Required Qualifications \r • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field \r • 7+ years of professional Java development experience \r • 5+ years hands-on experience with Apache Spark in production environments \r • Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management \r • Proven track record designing systems processing terabyte+ scale data \r • Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg) \r • Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark \r • Proficiency with Kafka \r • Strong grasp of CI/CD, containerization, and infrastructure-as-code practices \r Preferred Qualifications \r • Experience with Flink or other stream-processing frameworks \r • Familiarity with data governance, lineage, and quality frameworks \r • Experience with workflow orchestration at scale  \r • Background in system design for multi-tenant or multi-region data platforms \r • Prior experience leading a team or acting as a technical lead","datePosted":"2026-08-13T23:21:08.252Z","dateModified":"2026-08-13T23:21:08.252Z","hiringOrganization":{"@type":"Organization","name":"2t Consulting","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Colonia","addressRegion":"NJ","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"7ede4c59eb461e6b8a68c2ef"},"url":"https://jobsearcher.com/jobs/7ede4c59eb461e6b8a68c2ef"}}