{"schemaVersion":"jobsearcher.job.v1","id":"23d3afd0a9547113c7e44f8a","url":"https://jobsearcher.com/jobs/23d3afd0a9547113c7e44f8a","canonicalUrl":"https://jobsearcher.com/jobs/23d3afd0a9547113c7e44f8a","title":"AI Accelerator, Software Engineer- Graph Optimization/Compilers","description":"Description\n\nInvent the future with us.\n\nAmpere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.\n\nAs a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.\n\nJoin us at Ampere and work alongside a passionate and growing team - we’d love to have you apply!\n\nAbout the Role:\n\nAs a Software Engineer on Ampere’s AI Accelerator team, you will optimize deep learning computational graphs to maximize the performance, efficiency, and scalability of Ampere’s AI accelerator hardware. You will work across the software stack, from model frameworks and inference-serving systems to graph optimization, compiler infrastructure, runtimes, and compute kernels.\n\n What You’ll Achieve:\nOptimize computational graphs for performance, throughput, latency, memory efficiency, and power efficiency on Ampere AI accelerators\nEnable and optimize models, frameworks, and inference platforms, including PyTorch, Llama.cpp, vLLM, and SGLang\nDevelop graph-level optimizations such as operator fusion, pattern matching, redundancy elimination, constant folding, layout optimization, memory planning, quantization, and accelerator offload\nOptimize transformer and LLM workloads, including dynamic shapes, attention mechanisms, KV-cache management, and mixed-precision execution\nAnalyze end-to-end performance across frameworks, compilers, runtimes, kernels, and hardware\nBuild profiling, benchmarking, validation, and performance-regression infrastructure\nIdentify bottlenecks using traces, compiler diagnostics, microbenchmarks, and hardware performance data\nCollaborate with compiler, runtime, kernel, architecture, hardware, and applications teams on hardware/software co-design\nContribute to architecture, design reviews, code reviews, documentation, and engineering best practices\n\nAbout You:\nBachelor’s degree in Computer Science, Computer Engineering, Mathematics, or a related technical field & 5 years of relevant experience; or a Master’s degree with & 3 years of relevant experience\nStrong foundations in algorithms, data structures, graph algorithms, computational complexity, and systems programming\nProficiency in Python and C/C++, demonstrated through internships, research, coursework, open-source contributions, or personal projects\nStrong ability to reason about execution dependencies, memory movement, numerical correctness, and hardware execution behavior\nExperience diagnosing performance issues through profiling, benchmarking, tracing, or hardware-level analysis is a plus\nFamiliarity with deep learning concepts, neural-network architectures, tensor operations, numerical precision, quantization, and memory layout\nExperience with CUDA, ROCm, OpenCL, SYCL, Triton, GPU programming, NPU programming, or other accelerator architectures is a plus\nFamiliarity with transformer models, LLM inference, attention mechanisms, KV-cache optimization, speculative decoding, mixed-precision execution, or sparsity is a plus\nDemonstrated exceptional problem-solving ability—IOI medal, ACM ICPC medal, Codeforces Grandmaster, USACO Platinum, or equivalent achievement in research or production engineering is a strong plus\nStrong analytical and debugging skills, with the ability to investigate ambiguous technical problems and deliver robust solutions\nFast learner who can quickly understand new architectures, frameworks, compilers, and workloads\nExperience using AI-assisted development tools to accelerate implementation, testing, debugging, and code review while maintaining technical ownership and code quality\n\nWhat We’ll Offer:\n\nAt Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $159,000 and $239,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\nBenefit highlights include:\nPremium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.\nUnlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.\nA variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.\n\nAnd there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\n#LI-Hybrid#LI-DR\n#LI-Hybrid\n\nAmpere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.","company":"Ampere Computing","rawCompany":"ampere computing","city":"Santa Clara","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-16T11:34:15.250Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1251.00","title":"Computer Programmers","slug":"computer-programmers"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI Accelerator, Software Engineer- Graph Optimization/Compilers","description":"Description\n\nInvent the future with us.\n\nAmpere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.\n\nAs a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.\n\nJoin us at Ampere and work alongside a passionate and growing team - we’d love to have you apply!\n\nAbout the Role:\n\nAs a Software Engineer on Ampere’s AI Accelerator team, you will optimize deep learning computational graphs to maximize the performance, efficiency, and scalability of Ampere’s AI accelerator hardware. You will work across the software stack, from model frameworks and inference-serving systems to graph optimization, compiler infrastructure, runtimes, and compute kernels.\n\n What You’ll Achieve:\nOptimize computational graphs for performance, throughput, latency, memory efficiency, and power efficiency on Ampere AI accelerators\nEnable and optimize models, frameworks, and inference platforms, including PyTorch, Llama.cpp, vLLM, and SGLang\nDevelop graph-level optimizations such as operator fusion, pattern matching, redundancy elimination, constant folding, layout optimization, memory planning, quantization, and accelerator offload\nOptimize transformer and LLM workloads, including dynamic shapes, attention mechanisms, KV-cache management, and mixed-precision execution\nAnalyze end-to-end performance across frameworks, compilers, runtimes, kernels, and hardware\nBuild profiling, benchmarking, validation, and performance-regression infrastructure\nIdentify bottlenecks using traces, compiler diagnostics, microbenchmarks, and hardware performance data\nCollaborate with compiler, runtime, kernel, architecture, hardware, and applications teams on hardware/software co-design\nContribute to architecture, design reviews, code reviews, documentation, and engineering best practices\n\nAbout You:\nBachelor’s degree in Computer Science, Computer Engineering, Mathematics, or a related technical field & 5 years of relevant experience; or a Master’s degree with & 3 years of relevant experience\nStrong foundations in algorithms, data structures, graph algorithms, computational complexity, and systems programming\nProficiency in Python and C/C++, demonstrated through internships, research, coursework, open-source contributions, or personal projects\nStrong ability to reason about execution dependencies, memory movement, numerical correctness, and hardware execution behavior\nExperience diagnosing performance issues through profiling, benchmarking, tracing, or hardware-level analysis is a plus\nFamiliarity with deep learning concepts, neural-network architectures, tensor operations, numerical precision, quantization, and memory layout\nExperience with CUDA, ROCm, OpenCL, SYCL, Triton, GPU programming, NPU programming, or other accelerator architectures is a plus\nFamiliarity with transformer models, LLM inference, attention mechanisms, KV-cache optimization, speculative decoding, mixed-precision execution, or sparsity is a plus\nDemonstrated exceptional problem-solving ability—IOI medal, ACM ICPC medal, Codeforces Grandmaster, USACO Platinum, or equivalent achievement in research or production engineering is a strong plus\nStrong analytical and debugging skills, with the ability to investigate ambiguous technical problems and deliver robust solutions\nFast learner who can quickly understand new architectures, frameworks, compilers, and workloads\nExperience using AI-assisted development tools to accelerate implementation, testing, debugging, and code review while maintaining technical ownership and code quality\n\nWhat We’ll Offer:\n\nAt Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $159,000 and $239,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\nBenefit highlights include:\nPremium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.\nUnlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.\nA variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.\n\nAnd there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\n#LI-Hybrid#LI-DR\n#LI-Hybrid\n\nAmpere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.","datePosted":"2026-09-16T11:34:15.250Z","dateModified":"2026-09-16T11:34:15.250Z","hiringOrganization":{"@type":"Organization","name":"Ampere Computing","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Santa Clara","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"23d3afd0a9547113c7e44f8a"},"url":"https://jobsearcher.com/jobs/23d3afd0a9547113c7e44f8a"}}