AI GPU Performance Engineer - Distributed Inference
Advanced Micro Devices in San Jose, CA is seeking a software engineer to drive strategy, architecture, and tooling for AI pre-training and distributed inference on AMD GPUs. You will collaborate across hardware, AI frameworks, compilers, runtime, ROCm, and developer tools to scale performance analysis and optimization.
The role focuses on network and NIC performance, mapping model architectures to low-level software, and shaping the roadmap for out-of-box performance across multiple GPU
#J-18808-Ljbffr