JOBSEARCHER

Machine Learning Inference Engineer

Title : ML Inference Engineer Location : San Francisco, CA Salary : $250k base + equity An AI Unicorn startup is hiring a Senior Machine Learning Inference Engineer for a full-time role. You will be responsible for improving efficiency for AI-native infrastructure powered by generative and multimodal models. The ideal candidate has over 3 years of professional experience and a strong understanding of GPU infrastructure, Python, and PyTorch. This is a highly autonomous role with significant ownership across inference systems and model performance in production. This role is hybrid in San Francisco Bay Area and offers full benefits and equity. Experience : Building AI applications at scale from the ground upStrong understanding of GPU infrastructure including Triton, TensorRT, or vLLM frameworksHands-on experience with Python and PyTorch Building model-serving MicroservicesDiffusion and Multimodal model experience is a plusBenefits : Competitive base salaryEquity$401k matchingMedical coverageOscar Associates Limited (US) is acting as an Employment Agency in relation to this vacancy.#J-18808-Ljbffr