{"schemaVersion":"jobsearcher.job.v1","id":"d0934e4e320a0f6ab005e1ea","url":"https://jobsearcher.com/jobs/d0934e4e320a0f6ab005e1ea","canonicalUrl":"https://jobsearcher.com/jobs/d0934e4e320a0f6ab005e1ea","title":"Senior Solutions Architect, Generative AI","description":"Overview\nAs an AI Solutions Architect, you will design and optimize large-scale GPU AI infrastructure for leading tech partners, accelerating workloads and maximizing utilization. You’ll work cross-functionally to drive performance and reliability, shaping solutions with NVIDIA technologies. This role offers exposure to frontier labs and consumer internet customers, with impact across benchmarking, onboarding, and technical collateral. You join a mission-driven team advancing AI infrastructure at scale.\n\nCompensation / Benefitsequity and benefitsremote work optionstravel up to 20% for on-site visitscompetitive salariesopportunity to influence industry infrastructure trendsdynamic, innovative culture\nResponsibilitiesCollaborate with customers to maximize GPU utilization and end-to-end workload throughput while reducing costsDesign and optimize large-scale AI clusters across compute, networking, storage, scheduling, orchestration, and observabilityProfile distributed training and inference workloads to identify bottlenecks across hardware and software stacksDiagnose intricate infrastructure and distributed system issues spanning InfiniBand, RoCE, cloud interconnects, RDMA, NCCL, NVLink, NVSwitchLead POCs and performance studies for large-scale AI infrastructure, develop benchmarking tools, automation, runbooks, and collateralPartner with NVIDIA engineering, product, and sales teams to secure design wins and deliver customer-centric solutions\nKey requirementsBS, MS, or PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or equivalent experience6+ years in AI infrastructure, systems engineering, HPC, networking, SRE, or related fieldDeep Linux, distributed computing, GPU architectures knowledgeHands-on experience with InfiniBand, RoCE, or GPUDirect RDMA in on-prem or cloud environmentsExperience debugging NCCL communication and distributed performance, including topology and routingExperience profiling AI workloads across compute, networking, storage, and orchestrationExperience with Kubernetes and Slurm, containers, and production monitoring systemsProficiency in Python, shell scripting, or similar for automation and troubleshootingCollaboration across cross-functional teamsStrong problem-solving and debugging mindsetClear communication and ability to present complex conceptsLinux systemsDistributed computing and GPU clustersInfiniBand, RoCE, GPUDirect RDMA","company":"NVIDIA","rawCompany":"nvidia","city":"Austin","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-09-15T04:14:30.511Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1241.00","title":"Computer Network Architects","slug":"computer-network-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Senior Solutions Architect, Generative AI","description":"Overview\nAs an AI Solutions Architect, you will design and optimize large-scale GPU AI infrastructure for leading tech partners, accelerating workloads and maximizing utilization. You’ll work cross-functionally to drive performance and reliability, shaping solutions with NVIDIA technologies. This role offers exposure to frontier labs and consumer internet customers, with impact across benchmarking, onboarding, and technical collateral. You join a mission-driven team advancing AI infrastructure at scale.\n\nCompensation / Benefitsequity and benefitsremote work optionstravel up to 20% for on-site visitscompetitive salariesopportunity to influence industry infrastructure trendsdynamic, innovative culture\nResponsibilitiesCollaborate with customers to maximize GPU utilization and end-to-end workload throughput while reducing costsDesign and optimize large-scale AI clusters across compute, networking, storage, scheduling, orchestration, and observabilityProfile distributed training and inference workloads to identify bottlenecks across hardware and software stacksDiagnose intricate infrastructure and distributed system issues spanning InfiniBand, RoCE, cloud interconnects, RDMA, NCCL, NVLink, NVSwitchLead POCs and performance studies for large-scale AI infrastructure, develop benchmarking tools, automation, runbooks, and collateralPartner with NVIDIA engineering, product, and sales teams to secure design wins and deliver customer-centric solutions\nKey requirementsBS, MS, or PhD in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or equivalent experience6+ years in AI infrastructure, systems engineering, HPC, networking, SRE, or related fieldDeep Linux, distributed computing, GPU architectures knowledgeHands-on experience with InfiniBand, RoCE, or GPUDirect RDMA in on-prem or cloud environmentsExperience debugging NCCL communication and distributed performance, including topology and routingExperience profiling AI workloads across compute, networking, storage, and orchestrationExperience with Kubernetes and Slurm, containers, and production monitoring systemsProficiency in Python, shell scripting, or similar for automation and troubleshootingCollaboration across cross-functional teamsStrong problem-solving and debugging mindsetClear communication and ability to present complex conceptsLinux systemsDistributed computing and GPU clustersInfiniBand, RoCE, GPUDirect RDMA","datePosted":"2026-09-15T04:14:30.511Z","dateModified":"2026-09-15T04:14:30.511Z","hiringOrganization":{"@type":"Organization","name":"NVIDIA","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Austin","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"d0934e4e320a0f6ab005e1ea"},"url":"https://jobsearcher.com/jobs/d0934e4e320a0f6ab005e1ea"}}