JOBSEARCHER

Founding ML Inference Performance Engineer

UrunMillbrae, CAL6 LeadOctober 4th, 2026
uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack. The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.#J-18808-Ljbffr