LLM Inference Optimization Engineer - Frontier Performance
GMI Cloud, Inc is hiring world-class Machine Learning Engineers to advance LLM inference optimization leveraging GPUs and cutting-edge techniques. You will drive research, validation, and productionization of optimization strategies across NV platforms, with a focus on speed, efficiency, and scalability. You will collaborate with platform and infrastructure teams, contribute to open-source projects, and help define recipes and benchmarks for industry-leading inference performance.
#J-18808-Ljbffr