{"schemaVersion":"jobsearcher.job.v1","id":"96a1cfe414918d8868fc05da","url":"https://jobsearcher.com/jobs/96a1cfe414918d8868fc05da","canonicalUrl":"https://jobsearcher.com/jobs/96a1cfe414918d8868fc05da","title":"AI Accelerator, Software Principal Engineer- Full-Stack","description":"Description\n\nInvent the future with us.\n\nAmpere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.\n\nAs a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.\n\nJoin us at Ampere and work alongside a passionate and growing team - we’d love to have you apply!\n\nAbout the Role:\n\nWe are looking for an engineer with strong experience in PyTorch-based AI deployment, accelerated inference execution, and systems integration across software components. The role involves working on temporal and multi-modal workloads, optimizing execution on target platforms, and building infrastructure to run AI models reliably in production environments.\n\nWhat You’ll Achieve:\nDeploy and validate different AI models across supported inference environments (local, on-prem, or edge/accelerated platforms), optimizing runtime performance and reliability.\n\nDevelop and tune inference graph transformations, including torch.export-based graph workflows.\n\nCollaborate with platform and infrastructure teams to improve model execution efficiency on supported AI accelerators and compute environments.\n\nIntegrate model inference pipelines into scalable services, including batching, streaming, and runtime orchestration.\n\nBuild and maintain middleware and communication layers to support modular, scalable system integration.\nSupport long-term platform development for end-to-end inference services and production readiness.\nRelevant Technical Areas:\nTemporal model architectures\n\nMulti-frame or sequence embeddings (e.g., video/text sequences)\n\nAttention-based models\n\nMulti-modal workloads (e.g., text, imaging, and other feature modalities)\n\nAutomated labeling, evaluation, and validation workflows\n\nCPU/runtime performance optimization for system components\n\nPublish-subscribe middleware and distributed communication systems\n\nAbout You:\nBachelors degree in Computer Science, Mathematics or a related technical field & 8 years of related experience; or Master's degree & 6 years\nStrong hands-on experience with PyTorch\n\nExperience deploying AI models to accelerated or constrained environments (e.g., edge, cloud GPU, or specialized accelerators)\n\nFamiliarity with graph optimization and model-performance tuning\n\nExperience working with middleware, messaging, or distributed communication layers\n\nGood understanding of hardware/software interaction in AI systems\n\nExperience collaborating with hardware or platform partners\n\nWhat We’ll Offer:\n\nAt Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $195,000 and $292,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\nBenefit highlights include:\nPremium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.\nUnlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.\nA variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.\n\nAnd there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\n#LI-CB1\n\n#LI-DR\n#LI-Hybrid\n\nAmpere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.","company":"Ampere Computing","rawCompany":"ampere computing","city":"Santa Clara","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-04T21:56:54.872Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI Accelerator, Software Principal Engineer- Full-Stack","description":"Description\n\nInvent the future with us.\n\nAmpere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.\n\nAs a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.\n\nJoin us at Ampere and work alongside a passionate and growing team - we’d love to have you apply!\n\nAbout the Role:\n\nWe are looking for an engineer with strong experience in PyTorch-based AI deployment, accelerated inference execution, and systems integration across software components. The role involves working on temporal and multi-modal workloads, optimizing execution on target platforms, and building infrastructure to run AI models reliably in production environments.\n\nWhat You’ll Achieve:\nDeploy and validate different AI models across supported inference environments (local, on-prem, or edge/accelerated platforms), optimizing runtime performance and reliability.\n\nDevelop and tune inference graph transformations, including torch.export-based graph workflows.\n\nCollaborate with platform and infrastructure teams to improve model execution efficiency on supported AI accelerators and compute environments.\n\nIntegrate model inference pipelines into scalable services, including batching, streaming, and runtime orchestration.\n\nBuild and maintain middleware and communication layers to support modular, scalable system integration.\nSupport long-term platform development for end-to-end inference services and production readiness.\nRelevant Technical Areas:\nTemporal model architectures\n\nMulti-frame or sequence embeddings (e.g., video/text sequences)\n\nAttention-based models\n\nMulti-modal workloads (e.g., text, imaging, and other feature modalities)\n\nAutomated labeling, evaluation, and validation workflows\n\nCPU/runtime performance optimization for system components\n\nPublish-subscribe middleware and distributed communication systems\n\nAbout You:\nBachelors degree in Computer Science, Mathematics or a related technical field & 8 years of related experience; or Master's degree & 6 years\nStrong hands-on experience with PyTorch\n\nExperience deploying AI models to accelerated or constrained environments (e.g., edge, cloud GPU, or specialized accelerators)\n\nFamiliarity with graph optimization and model-performance tuning\n\nExperience working with middleware, messaging, or distributed communication layers\n\nGood understanding of hardware/software interaction in AI systems\n\nExperience collaborating with hardware or platform partners\n\nWhat We’ll Offer:\n\nAt Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $195,000 and $292,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\nBenefit highlights include:\nPremium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.\nUnlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.\nA variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.\n\nAnd there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.\n\n#LI-CB1\n\n#LI-DR\n#LI-Hybrid\n\nAmpere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.","datePosted":"2026-08-04T21:56:54.872Z","dateModified":"2026-08-04T21:56:54.872Z","hiringOrganization":{"@type":"Organization","name":"Ampere Computing","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Santa Clara","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"96a1cfe414918d8868fc05da"},"url":"https://jobsearcher.com/jobs/96a1cfe414918d8868fc05da"}}