{"schemaVersion":"jobsearcher.job.v1","id":"ca1d810bb5feee8a3ed51cbc","url":"https://jobsearcher.com/jobs/ca1d810bb5feee8a3ed51cbc","canonicalUrl":"https://jobsearcher.com/jobs/ca1d810bb5feee8a3ed51cbc","title":"Sieve Forward Deployed Engineer","description":"Sieve — Forward Deployed Engineer Type: Full-time | On-site | San Francisco, CA Compensation: $150,000 – $250,000 + competitive equity Experience: 1 – 3 years Hiring count: 4 (hiring multiple) Visa sponsorship: Yes — H-1B, OPT Tech stack: Python, PyTorch (or similar ML frameworks), large-scale data pipelines\nAbout Sieve Sieve is an AI research lab focused exclusively on video data. Video makes up ~80% of internet traffic and is the dominant medium across creativity, communication, gaming, AR/VR, and robotics — but progress in video modeling has been bottlenecked by access to high-quality training data.\nSieve combines exabyte-scale video infrastructure, novel video understanding techniques, and dozens of diverse data sources to build datasets that push the frontier of video modeling — with precision, quality, and speed that has earned the trust of frontier AI labs, Fortune 100 companies, and fast-growing generative AI startups. Beyond video, the team works on audio and multimodal data processing for AI training and evaluation.\nSeed-stage, founded 2022, San Francisco. Website: sievedata.com\nAbout This Role You'll own end-to-end dataset projects for customers — from untangling ambiguous requirements through shipping production systems that find, generate, filter, transform, evaluate, and package high-quality datasets at scale. This is a high-agency role working directly with customers and internal teams, combining research prototypes with reliable production pipelines. You'll ship fast, move between technical domains within each project, and own customer outcomes directly.\nWhat You'll Own Work directly with customers to translate ambiguous dataset needs into concrete technical systems and delivery timelines\nBuild custom algorithms, models, and large-scale data pipelines spanning computer vision, audio processing, text processing, and metadata analysis\nMove between research prototypes and production systems, using models and APIs creatively to solve customer problems\nBreak down customer-level goals into the models, heuristics, infrastructure, and QA steps needed to deliver\nOptimize performance through pre/post-processing, parallelism, inference optimization, fine-tuning, and evaluation loops\nMust-Have Strong Python developer with hands-on experience building custom algorithms, model workflows, or large-scale data pipelines\nComfortable working directly with customers or external teams to translate ambiguous needs into technical systems\nDeep intuition for dataset quality, filtering, labeling, evaluation, and edge cases\nAble to move quickly between research prototypes and reliable production systems without creating brittle code\n1–3 years of experience shipping technical work in a startup or high-velocity environment\nNice-to-Have Experience building custom algorithms or ML workflows for production video, audio, or multimodal data\nHands-on work with large-scale data pipelines at scale\nBackground with PyTorch or similar ML frameworks in production\nActive contributor to open source projects\nEarly hire experience at a startup\nBenefits & Perks 401(k)\nFull health insurance\nBreakfast, lunch, and dinner covered\nChoice of snacks\nUbers covered home\nCompetitive equity\n\n#J-18808-Ljbffr","company":"David Joseph","rawCompany":"david joseph","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-07-16T03:44:47.682Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Sieve Forward Deployed Engineer","description":"Sieve — Forward Deployed Engineer Type: Full-time | On-site | San Francisco, CA Compensation: $150,000 – $250,000 + competitive equity Experience: 1 – 3 years Hiring count: 4 (hiring multiple) Visa sponsorship: Yes — H-1B, OPT Tech stack: Python, PyTorch (or similar ML frameworks), large-scale data pipelines\nAbout Sieve Sieve is an AI research lab focused exclusively on video data. Video makes up ~80% of internet traffic and is the dominant medium across creativity, communication, gaming, AR/VR, and robotics — but progress in video modeling has been bottlenecked by access to high-quality training data.\nSieve combines exabyte-scale video infrastructure, novel video understanding techniques, and dozens of diverse data sources to build datasets that push the frontier of video modeling — with precision, quality, and speed that has earned the trust of frontier AI labs, Fortune 100 companies, and fast-growing generative AI startups. Beyond video, the team works on audio and multimodal data processing for AI training and evaluation.\nSeed-stage, founded 2022, San Francisco. Website: sievedata.com\nAbout This Role You'll own end-to-end dataset projects for customers — from untangling ambiguous requirements through shipping production systems that find, generate, filter, transform, evaluate, and package high-quality datasets at scale. This is a high-agency role working directly with customers and internal teams, combining research prototypes with reliable production pipelines. You'll ship fast, move between technical domains within each project, and own customer outcomes directly.\nWhat You'll Own Work directly with customers to translate ambiguous dataset needs into concrete technical systems and delivery timelines\nBuild custom algorithms, models, and large-scale data pipelines spanning computer vision, audio processing, text processing, and metadata analysis\nMove between research prototypes and production systems, using models and APIs creatively to solve customer problems\nBreak down customer-level goals into the models, heuristics, infrastructure, and QA steps needed to deliver\nOptimize performance through pre/post-processing, parallelism, inference optimization, fine-tuning, and evaluation loops\nMust-Have Strong Python developer with hands-on experience building custom algorithms, model workflows, or large-scale data pipelines\nComfortable working directly with customers or external teams to translate ambiguous needs into technical systems\nDeep intuition for dataset quality, filtering, labeling, evaluation, and edge cases\nAble to move quickly between research prototypes and reliable production systems without creating brittle code\n1–3 years of experience shipping technical work in a startup or high-velocity environment\nNice-to-Have Experience building custom algorithms or ML workflows for production video, audio, or multimodal data\nHands-on work with large-scale data pipelines at scale\nBackground with PyTorch or similar ML frameworks in production\nActive contributor to open source projects\nEarly hire experience at a startup\nBenefits & Perks 401(k)\nFull health insurance\nBreakfast, lunch, and dinner covered\nChoice of snacks\nUbers covered home\nCompetitive equity\n\n#J-18808-Ljbffr","datePosted":"2026-07-16T03:44:47.682Z","dateModified":"2026-07-16T03:44:47.682Z","hiringOrganization":{"@type":"Organization","name":"David Joseph","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"ca1d810bb5feee8a3ed51cbc"},"url":"https://jobsearcher.com/jobs/ca1d810bb5feee8a3ed51cbc"}}