JOBSEARCHER

Software Engineer, Model Inference, DeepMind

Job TitleInfo: By applying to this position you will have an opportunity to share your preferred working location from the following: London, UK; Mountain View, CA, USA.Minimum Qualifications:Bachelor's degree or equivalent practical experience.8 years of experience in software development.2 years of experience in deploying and maintaining machine learning models in a live production environment.Experience in profiling, configuring, or executing ML workloads directly on hardware accelerators (e.g., GPU or TPU).Experience designing, building, or optimizing model serving infrastructure or inference backends.Preferred Qualifications:Experience with developing serving infrastructure.Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL).Experience profiling software to identify performance bottlenecks.Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism).Familiarity with writing performance-optimized kernels.Understanding of LLM architecture and inference performance dynamics (e.g., Transformer models, memory bandwidth and compute bounds, KV cache scaling).