{"schemaVersion":"jobsearcher.job.v1","id":"e124bb9be56a459aa3cea85a","url":"https://jobsearcher.com/jobs/e124bb9be56a459aa3cea85a","canonicalUrl":"https://jobsearcher.com/jobs/e124bb9be56a459aa3cea85a","title":"Sr. Big Data Engineer","description":"Overview\nIn this role you will design and implement big data solutions that enable advanced analytics. You will work with a cross-functional team to pre-process, analyze and visualize large data sets, turning information into business insights. You’ll evaluate and deploy machine learning models on big data, tackling complex problems with scalable code. This is an opportunity to shape analytics capabilities in a fast-paced, collaborative environment.\n\nResponsibilitiesEvaluate, develop, maintain and test big data solutions for advanced analytics projectsDesign and implement big data pre-processing and reporting workflows (collecting, parsing, managing, analyzing and visualizing data)Test machine learning models on big data and deploy them for scoring and predictionCollaborate in an agile team to deliver modular, reusable and scalable codeWork with cloud platforms and distributed computing tools to process large datasetsParticipate in automation and scripting to streamline tasks\nKey requirementsExpert-level proficiency in at least one of Java, C++ or PythonScala knowledge is a strong advantageExperience with distributed computing frameworks (Hadoop 2.0, YARN, MR, HDFS) and related technologies (Hive, Sqoop, Avro, Flume, Oozie, Zookeeper)Hands-on experience with Apache Spark (Streaming, SQL, MLLib) is a strong advantageCloud computing platforms experience (AWS: EMR, EC2, S3, SWF; AWS CLI)Linux environment proficiency and shell/Python scriptingAgile experience and familiarity with Jira; understanding of Git workflowsStrong problem-solving skills and ability to navigate tight cornersTeamwork and collaborationProblem-solving mindsetAdaptability in a fast-paced environmentJavaPythonC++","company":"Sonsoft","rawCompany":"sonsoft","city":"Brooklyn","state":"NY","isRemote":false,"isActive":false,"createdAt":"2026-09-15T04:03:26.791Z","occupations":[{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Sr. Big Data Engineer","description":"Overview\nIn this role you will design and implement big data solutions that enable advanced analytics. You will work with a cross-functional team to pre-process, analyze and visualize large data sets, turning information into business insights. You’ll evaluate and deploy machine learning models on big data, tackling complex problems with scalable code. This is an opportunity to shape analytics capabilities in a fast-paced, collaborative environment.\n\nResponsibilitiesEvaluate, develop, maintain and test big data solutions for advanced analytics projectsDesign and implement big data pre-processing and reporting workflows (collecting, parsing, managing, analyzing and visualizing data)Test machine learning models on big data and deploy them for scoring and predictionCollaborate in an agile team to deliver modular, reusable and scalable codeWork with cloud platforms and distributed computing tools to process large datasetsParticipate in automation and scripting to streamline tasks\nKey requirementsExpert-level proficiency in at least one of Java, C++ or PythonScala knowledge is a strong advantageExperience with distributed computing frameworks (Hadoop 2.0, YARN, MR, HDFS) and related technologies (Hive, Sqoop, Avro, Flume, Oozie, Zookeeper)Hands-on experience with Apache Spark (Streaming, SQL, MLLib) is a strong advantageCloud computing platforms experience (AWS: EMR, EC2, S3, SWF; AWS CLI)Linux environment proficiency and shell/Python scriptingAgile experience and familiarity with Jira; understanding of Git workflowsStrong problem-solving skills and ability to navigate tight cornersTeamwork and collaborationProblem-solving mindsetAdaptability in a fast-paced environmentJavaPythonC++","datePosted":"2026-09-15T04:03:26.791Z","dateModified":"2026-09-15T04:03:26.791Z","hiringOrganization":{"@type":"Organization","name":"Sonsoft","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Brooklyn","addressRegion":"NY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"e124bb9be56a459aa3cea85a"},"url":"https://jobsearcher.com/jobs/e124bb9be56a459aa3cea85a"}}