GPU Reliability Engineer - AI Supercomputing Fleet
Thinking Machines in San Francisco is hiring an engineer to ensure the reliability of their GPU supercomputing fleet. You'll be responsible for diagnosing hardware issues and collaborating with vendors to resolve them efficiently.
Ideal candidates hold a Bachelor’s degree in computer science or engineering, possess backend programming skills in Python or Rust, and have experience with large-scale systems. The position offers a competitive salary range of $350,000 to $475,000 and generous benefits including health insurance and unlimited PTO.
#J-18808-Ljbffr