{"schemaVersion":"jobsearcher.job.v1","id":"1443ed2dff57e0d2850b58bf","url":"https://jobsearcher.com/jobs/1443ed2dff57e0d2850b58bf","canonicalUrl":"https://jobsearcher.com/jobs/1443ed2dff57e0d2850b58bf","title":"Sr. Data Engineer","description":"Are you an experienced Data Engineer looking for the chance to make a real impact? We're hiring! We help organizations create robust technical architecture, develop complex systems and build out their infrastructures. Check out our opening here.\nRole and Responsibilities:\nDesign and Develop scalable industry standard ETL pipelines using big data technologies.\nLoading from disparate data sets by leveraging various big data technologies EMR, Hive, Spark. Presto etc.\nIngest, Extract, Move, Transform, Cleanse, and Load massive structured and unstructured data in Hadoop environment both in batch and real-time.\nBuild and maintain robust data pipelines for multiple use cases, and ensure high quality data i.e., accurate, complete and timely.\nHigh-speed querying using in-memory technologies such as Spark.\nDesign and implement data modeling\nTranslate complex functional and technical requirements into detailed design\nCode and query Optimization.\nSkills & Qualifications\n3-5 years of experience in Hadoop (Hive and Impala) & Spark (Spark/Scala)\nProficient in a scripting language and object-oriented language – Python/Scala preferred\nProficient with a public cloud (AWS, Azure, GCP)\nExpert in SQL development\nProficient in using Big Data stack (Hadoop, Hive, Spark, Kafka, Kerberos, OOZIE, impala, etc.)\nAdept in multi-threading and concurrency concepts.\nJob Type: Full-time\nPay: $100,000.00 - $151,000.00 per year\nSchedule:\nMonday to Friday\nAbility to commute/relocate:\nSan Francisco Bay Area, CA: Reliably commute or planning to relocate before starting work (Required)\nExperience:\nScripting: 2 years (Required)\nHadoop: 2 years (Preferred)\nPython: 2 years (Required)\nWork Location: In person","company":"Invictus Data","rawCompany":"invictus data","city":"Alameda","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-10T13:01:26.494Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Sr. Data Engineer","description":"Are you an experienced Data Engineer looking for the chance to make a real impact? We're hiring! We help organizations create robust technical architecture, develop complex systems and build out their infrastructures. Check out our opening here.\nRole and Responsibilities:\nDesign and Develop scalable industry standard ETL pipelines using big data technologies.\nLoading from disparate data sets by leveraging various big data technologies EMR, Hive, Spark. Presto etc.\nIngest, Extract, Move, Transform, Cleanse, and Load massive structured and unstructured data in Hadoop environment both in batch and real-time.\nBuild and maintain robust data pipelines for multiple use cases, and ensure high quality data i.e., accurate, complete and timely.\nHigh-speed querying using in-memory technologies such as Spark.\nDesign and implement data modeling\nTranslate complex functional and technical requirements into detailed design\nCode and query Optimization.\nSkills & Qualifications\n3-5 years of experience in Hadoop (Hive and Impala) & Spark (Spark/Scala)\nProficient in a scripting language and object-oriented language – Python/Scala preferred\nProficient with a public cloud (AWS, Azure, GCP)\nExpert in SQL development\nProficient in using Big Data stack (Hadoop, Hive, Spark, Kafka, Kerberos, OOZIE, impala, etc.)\nAdept in multi-threading and concurrency concepts.\nJob Type: Full-time\nPay: $100,000.00 - $151,000.00 per year\nSchedule:\nMonday to Friday\nAbility to commute/relocate:\nSan Francisco Bay Area, CA: Reliably commute or planning to relocate before starting work (Required)\nExperience:\nScripting: 2 years (Required)\nHadoop: 2 years (Preferred)\nPython: 2 years (Required)\nWork Location: In person","datePosted":"2026-08-10T13:01:26.494Z","dateModified":"2026-08-10T13:01:26.494Z","hiringOrganization":{"@type":"Organization","name":"Invictus Data","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Alameda","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"1443ed2dff57e0d2850b58bf"},"url":"https://jobsearcher.com/jobs/1443ed2dff57e0d2850b58bf"}}