{"schemaVersion":"jobsearcher.job.v1","id":"1ff171d3b93e5a9f8f43c611","url":"https://jobsearcher.com/jobs/1ff171d3b93e5a9f8f43c611","canonicalUrl":"https://jobsearcher.com/jobs/1ff171d3b93e5a9f8f43c611","title":"Java Spark Engineer","description":"Job DescriptionPrimary Responsibilities• Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)• Lead design of batch and streaming ETL/ELT systems handling large data volumes• Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction• Set coding standards and lead code/design reviews across the team• Drive technical decisions on data architecture, storage formats, and pipeline orchestration• Mentor mid-level and junior engineers; act as a technical escalation point• Partner with product, analytics, and platform teams to translate requirements into scalable systems• Own production reliability — on-call ownership, incident response, root-cause analysis for pipeline failures• Evaluate and introduce new tools/frameworks where they improve the system• Contribute to capacity planning and cost optimization for cluster infrastructureRequired Qualifications• Bachelor’s or Master’s degree in Computer Science, Engineering, or related field• 7+ years of professional Java development experience• 5+ years hands-on experience with Apache Spark in production environments• Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management• Proven track record designing systems processing terabyte+ scale data• Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg)• Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark• Proficiency with Kafka• Strong grasp of CI/CD, containerization, and infrastructure-as-code practicesPreferred Qualifications• Experience with Flink or other stream-processing frameworks• Familiarity with data governance, lineage, and quality frameworks• Experience with workflow orchestration at scale • Background in system design for multi-tenant or multi-region data platforms• Prior experience leading a team or acting as a technical lead","company":"Veriipro","rawCompany":"veriipro","city":"Berkeley Heights","state":"NJ","isRemote":false,"isActive":false,"createdAt":"2026-08-05T23:37:52.996Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Java Spark Engineer","description":"Job DescriptionPrimary Responsibilities• Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)• Lead design of batch and streaming ETL/ELT systems handling large data volumes• Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction• Set coding standards and lead code/design reviews across the team• Drive technical decisions on data architecture, storage formats, and pipeline orchestration• Mentor mid-level and junior engineers; act as a technical escalation point• Partner with product, analytics, and platform teams to translate requirements into scalable systems• Own production reliability — on-call ownership, incident response, root-cause analysis for pipeline failures• Evaluate and introduce new tools/frameworks where they improve the system• Contribute to capacity planning and cost optimization for cluster infrastructureRequired Qualifications• Bachelor’s or Master’s degree in Computer Science, Engineering, or related field• 7+ years of professional Java development experience• 5+ years hands-on experience with Apache Spark in production environments• Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management• Proven track record designing systems processing terabyte+ scale data• Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg)• Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark• Proficiency with Kafka• Strong grasp of CI/CD, containerization, and infrastructure-as-code practicesPreferred Qualifications• Experience with Flink or other stream-processing frameworks• Familiarity with data governance, lineage, and quality frameworks• Experience with workflow orchestration at scale • Background in system design for multi-tenant or multi-region data platforms• Prior experience leading a team or acting as a technical lead","datePosted":"2026-08-05T23:37:52.996Z","dateModified":"2026-08-05T23:37:52.996Z","hiringOrganization":{"@type":"Organization","name":"Veriipro","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Berkeley Heights","addressRegion":"NJ","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"1ff171d3b93e5a9f8f43c611"},"url":"https://jobsearcher.com/jobs/1ff171d3b93e5a9f8f43c611"}}