{"schemaVersion":"jobsearcher.job.v1","id":"df88649ef7ecaafd95ed451c","url":"https://jobsearcher.com/jobs/df88649ef7ecaafd95ed451c","canonicalUrl":"https://jobsearcher.com/jobs/df88649ef7ecaafd95ed451c","title":"Software Engineer, ML Infrastructure, Content Retrieval Platform, Level 4","description":"Overview\nIn this role you will help design and optimize the ML infrastructure that supports content retrieval and recommendations at scale. You’ll collaborate across the Content Retrieval Platform and Content ML teams to build high‑performance inference, data pipelines, and training infrastructure in the cloud. You’ll deploy cutting‑edge models to production while upholding security and reliability standards. This is an opportunity to shape scalable ML systems used by hundreds of millions of users and contribute to Snap’s privacy‑first culture.\n\nCompensation / Benefitspaid parental leavecomprehensive medical coverageemotional and mental health support programsequity in RSUscompensation packages tied to long-term success\nResponsibilitiesDesign and optimize infrastructure for ML workloads at scale to improve reliability and efficiencyBuild and enhance feature generation and serving pipelines for online inference and offline training data generationDevelop high‑performance inference systems for fast AI model servingCreate scalable ML model training, evaluation, and inference infrastructure in the cloudDevelop data management systems for scalable data collection, labeling, processing, and evaluationCollaborate with ML engineers to deploy cutting‑edge models into productionApply AI tooling and high‑velocity workflows to design and ship scalable services while ensuring code correctness and security\nKey requirementsStrong programming skills in Python, JavaStrong problem‑solving with focus on system performance, scalability, and efficiencyGood understanding of distributed systems and large‑scale ML infrastructureExperience with big data processing frameworks such as Spark, Flink, or RayAbility to collaborate and work well with othersProven track record of operating highly‑available systems at scaleProactive learner with ability to adapt to evolving AI systems and toolsCollaborative mindsetProactive learningAdaptabilityPythonJavaDistributed systems","company":"Snap","rawCompany":"snap","city":"San Jose","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-15T03:50:46.680Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Software Engineer, ML Infrastructure, Content Retrieval Platform, Level 4","description":"Overview\nIn this role you will help design and optimize the ML infrastructure that supports content retrieval and recommendations at scale. You’ll collaborate across the Content Retrieval Platform and Content ML teams to build high‑performance inference, data pipelines, and training infrastructure in the cloud. You’ll deploy cutting‑edge models to production while upholding security and reliability standards. This is an opportunity to shape scalable ML systems used by hundreds of millions of users and contribute to Snap’s privacy‑first culture.\n\nCompensation / Benefitspaid parental leavecomprehensive medical coverageemotional and mental health support programsequity in RSUscompensation packages tied to long-term success\nResponsibilitiesDesign and optimize infrastructure for ML workloads at scale to improve reliability and efficiencyBuild and enhance feature generation and serving pipelines for online inference and offline training data generationDevelop high‑performance inference systems for fast AI model servingCreate scalable ML model training, evaluation, and inference infrastructure in the cloudDevelop data management systems for scalable data collection, labeling, processing, and evaluationCollaborate with ML engineers to deploy cutting‑edge models into productionApply AI tooling and high‑velocity workflows to design and ship scalable services while ensuring code correctness and security\nKey requirementsStrong programming skills in Python, JavaStrong problem‑solving with focus on system performance, scalability, and efficiencyGood understanding of distributed systems and large‑scale ML infrastructureExperience with big data processing frameworks such as Spark, Flink, or RayAbility to collaborate and work well with othersProven track record of operating highly‑available systems at scaleProactive learner with ability to adapt to evolving AI systems and toolsCollaborative mindsetProactive learningAdaptabilityPythonJavaDistributed systems","datePosted":"2026-09-15T03:50:46.680Z","dateModified":"2026-09-15T03:50:46.680Z","hiringOrganization":{"@type":"Organization","name":"Snap","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"San Jose","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"df88649ef7ecaafd95ed451c"},"url":"https://jobsearcher.com/jobs/df88649ef7ecaafd95ed451c"}}