{"schemaVersion":"jobsearcher.job.v1","id":"d972527d03d3f32a44b0a523","url":"https://jobsearcher.com/jobs/d972527d03d3f32a44b0a523","canonicalUrl":"https://jobsearcher.com/jobs/d972527d03d3f32a44b0a523","title":"Big Data Engineer / Sr. Engineer","description":"Job Description\r\nJob Title: Big Data Engineer / Sr. Engineer\r\nLocation: Jersey City, NJ\r\nFull Time Permanent\r\nIf you are an exceptional developer with an aptitude to learn and implement using new technologies, and who loves to push the boundaries to solve complex business problems innovatively, then we would like to talk with you.\r\nRESPONSIBILITIES\r\nEvaluating, developing, maintaining and testing big data solutions for advanced analytics projects.\r\nManaging big data pre-processing & reporting workflows, including collecting, parsing, managing, analyzing and visualizing large sets of data to turn information into business insights.\r\nTesting various machine learning models on Big Data, and deploying learned models for ongoing scoring and prediction. Appreciation of the mechanics of complex machine learning algorithms would be a strong advantage.\r\nQUALIFICATIONS & EXPERIENCE\r\n3+ years of demonstrable experience designing technological solutions to complex data problems, developing & testing modular, reusable, efficient and scalable code to implement those solutions.\r\nExpert-level proficiency in at least one of Java, C++ or Python (preferred). Scala knowledge is a strong advantage.\r\nStrong understanding and experience in distributed computing frameworks, particularly Apache Hadoop 2.0 (YARN, MR & HDFS) and associated technologies – one or more of Hive, Sqoop, Avro, Flume, Oozie, Zookeeper, etc.\r\nHands-on experience with Apache Spark and its components (Streaming, SQL, MLLib) is a strong advantage.\r\nOperating knowledge of cloud computing platforms (AWS, especially EMR, EC2, S3, SWF services and the AWS CLI).\r\nExperience working within a Linux computing environment, and use of command line tools including knowledge of shell/Python scripting for automating common tasks.\r\nAbility to work in a team in an agile setting, familiarity with JIRA and clear understanding of how Git works.\r\nIn addition, the ideal candidate would have great problem-solving skills, and the ability & confidence to hack their way out of tight corners.\r\nMUST HAVE (hands-on) experience\r\nJava, Python or C++ expertise.\r\nLinux environment and shell scripting.\r\nDistributed computing frameworks (Hadoop or Spark).\r\nCloud computing platforms (AWS).\r\nDESIRABLE (would be a plus)\r\nStatistical or machine learning DSL like R.\r\nDistributed and low latency (streaming) application architecture.\r\nRow store distributed DBMSs such as Cassandra.\r\nFamiliarity with API design.\r\nEDUCATION\r\nB.E/B.Tech in Computer Science or related technical degree.\r\nAdditional Information\r\nAll your information will be kept confidential according to EEO guidelines.\r\nJ-18808-Ljbffr","company":"Implify","rawCompany":"implify","city":"Jersey City","state":"NJ","isRemote":false,"isActive":false,"createdAt":"2026-07-16T02:02:27.203Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Big Data Engineer / Sr. Engineer","description":"Job Description\r\nJob Title: Big Data Engineer / Sr. Engineer\r\nLocation: Jersey City, NJ\r\nFull Time Permanent\r\nIf you are an exceptional developer with an aptitude to learn and implement using new technologies, and who loves to push the boundaries to solve complex business problems innovatively, then we would like to talk with you.\r\nRESPONSIBILITIES\r\nEvaluating, developing, maintaining and testing big data solutions for advanced analytics projects.\r\nManaging big data pre-processing & reporting workflows, including collecting, parsing, managing, analyzing and visualizing large sets of data to turn information into business insights.\r\nTesting various machine learning models on Big Data, and deploying learned models for ongoing scoring and prediction. Appreciation of the mechanics of complex machine learning algorithms would be a strong advantage.\r\nQUALIFICATIONS & EXPERIENCE\r\n3+ years of demonstrable experience designing technological solutions to complex data problems, developing & testing modular, reusable, efficient and scalable code to implement those solutions.\r\nExpert-level proficiency in at least one of Java, C++ or Python (preferred). Scala knowledge is a strong advantage.\r\nStrong understanding and experience in distributed computing frameworks, particularly Apache Hadoop 2.0 (YARN, MR & HDFS) and associated technologies – one or more of Hive, Sqoop, Avro, Flume, Oozie, Zookeeper, etc.\r\nHands-on experience with Apache Spark and its components (Streaming, SQL, MLLib) is a strong advantage.\r\nOperating knowledge of cloud computing platforms (AWS, especially EMR, EC2, S3, SWF services and the AWS CLI).\r\nExperience working within a Linux computing environment, and use of command line tools including knowledge of shell/Python scripting for automating common tasks.\r\nAbility to work in a team in an agile setting, familiarity with JIRA and clear understanding of how Git works.\r\nIn addition, the ideal candidate would have great problem-solving skills, and the ability & confidence to hack their way out of tight corners.\r\nMUST HAVE (hands-on) experience\r\nJava, Python or C++ expertise.\r\nLinux environment and shell scripting.\r\nDistributed computing frameworks (Hadoop or Spark).\r\nCloud computing platforms (AWS).\r\nDESIRABLE (would be a plus)\r\nStatistical or machine learning DSL like R.\r\nDistributed and low latency (streaming) application architecture.\r\nRow store distributed DBMSs such as Cassandra.\r\nFamiliarity with API design.\r\nEDUCATION\r\nB.E/B.Tech in Computer Science or related technical degree.\r\nAdditional Information\r\nAll your information will be kept confidential according to EEO guidelines.\r\nJ-18808-Ljbffr","datePosted":"2026-07-16T02:02:27.203Z","dateModified":"2026-07-16T02:02:27.203Z","hiringOrganization":{"@type":"Organization","name":"Implify","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Jersey City","addressRegion":"NJ","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"d972527d03d3f32a44b0a523"},"url":"https://jobsearcher.com/jobs/d972527d03d3f32a44b0a523"}}