JOBSEARCHER

Senior ML Training Systems Engineer - Distributed GPU Infra

A leading AI technology company in San Francisco is looking for a Senior Software Engineer to build scalable infrastructure for large‐scale training and fine-tuning of foundation models. You will design distributed training systems and optimize GPU utilization while collaborating with cross-functional teams to adapt models efficiently. Ideal candidates have over 5 years of experience in ML infrastructure and a strong background in distributed training frameworks. Competitive compensation and benefits package is offered. #J-18808-Ljbffr