{"schemaVersion":"jobsearcher.job.v1","id":"0e038c2d4ce74b1ad2fd235d","url":"https://jobsearcher.com/jobs/0e038c2d4ce74b1ad2fd235d","canonicalUrl":"https://jobsearcher.com/jobs/0e038c2d4ce74b1ad2fd235d","title":"Platform Engineer","description":"Key Responsibilities:Design and implement scalable infrastructure for LLM and GenAI workloads across multi-GPU environmentsPerform GPU profiling, benchmarking, and performance optimization for distributed training workloadsManage and schedule compute-intensive jobs using Slurm-based clusters and OpenShift/Kubernetes environmentsEnable and optimize the NVIDIA GPU stack (CUDA, cuDNN, NCCL, Triton, RAPIDS, etc.)Collaborate with cross-functional teams to deploy models in research and production environmentsBuild and support GenAI pipelines (fine-tuning, RAG, multi-modal inferencing, LLMOps)Develop reusable infrastructure templates using tools like Terraform and HelmContribute to internal innovation (PoCs, workshops) and support client-facing delivery engagementsBasic Qualifications:Strong experience with Slurm and distributed training environmentsHands-on expertise with Red Hat OpenShift and/or KubernetesDeep knowledge of the NVIDIA GPU ecosystem (CUDA, cuDNN, NCCL, Nsight, Triton/TensorRT)Strong foundation in Linux systems, performance tuning, and multi-GPU optimizationExperience deploying GenAI workloads (LLM fine-tuning, RAG pipelines, multi-modal systems)Familiarity with Infrastructure-as-Code tools (Terraform, Ansible)Experience with cloud GPU environments (GCP, Azure, AWS, OCI) and/or on-prem GPU clustersOther Qualifications (OQs):Experience with NVIDIA NIMs, DGX systems, or GPU-accelerated containersKnowledge of LLMOps frameworks and MLOps integrationFamiliarity with vector databases and retrieval systems for RAG architecturesComfortable working in client-facing environments and collaborating with AI solution teamsHealthcare Domain Experience (Nice to Have):Experience working with FHIR R4, HL7 v2, or SMART on FHIRIntegration with EHR systems (e.g., Epic)Understanding of HIPAA compliance and healthcare data privacyExposure to clinical workflows, CDS Hooks, or patient-facing applicationsExperience building clinical decision support systems or healthcare interoperability solutions","company":"Quantiphi","rawCompany":"quantiphi","city":"Denver","state":"CO","isRemote":false,"isActive":false,"createdAt":"2026-08-14T14:59:37.218Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Platform Engineer","description":"Key Responsibilities:Design and implement scalable infrastructure for LLM and GenAI workloads across multi-GPU environmentsPerform GPU profiling, benchmarking, and performance optimization for distributed training workloadsManage and schedule compute-intensive jobs using Slurm-based clusters and OpenShift/Kubernetes environmentsEnable and optimize the NVIDIA GPU stack (CUDA, cuDNN, NCCL, Triton, RAPIDS, etc.)Collaborate with cross-functional teams to deploy models in research and production environmentsBuild and support GenAI pipelines (fine-tuning, RAG, multi-modal inferencing, LLMOps)Develop reusable infrastructure templates using tools like Terraform and HelmContribute to internal innovation (PoCs, workshops) and support client-facing delivery engagementsBasic Qualifications:Strong experience with Slurm and distributed training environmentsHands-on expertise with Red Hat OpenShift and/or KubernetesDeep knowledge of the NVIDIA GPU ecosystem (CUDA, cuDNN, NCCL, Nsight, Triton/TensorRT)Strong foundation in Linux systems, performance tuning, and multi-GPU optimizationExperience deploying GenAI workloads (LLM fine-tuning, RAG pipelines, multi-modal systems)Familiarity with Infrastructure-as-Code tools (Terraform, Ansible)Experience with cloud GPU environments (GCP, Azure, AWS, OCI) and/or on-prem GPU clustersOther Qualifications (OQs):Experience with NVIDIA NIMs, DGX systems, or GPU-accelerated containersKnowledge of LLMOps frameworks and MLOps integrationFamiliarity with vector databases and retrieval systems for RAG architecturesComfortable working in client-facing environments and collaborating with AI solution teamsHealthcare Domain Experience (Nice to Have):Experience working with FHIR R4, HL7 v2, or SMART on FHIRIntegration with EHR systems (e.g., Epic)Understanding of HIPAA compliance and healthcare data privacyExposure to clinical workflows, CDS Hooks, or patient-facing applicationsExperience building clinical decision support systems or healthcare interoperability solutions","datePosted":"2026-08-14T14:59:37.218Z","dateModified":"2026-08-14T14:59:37.218Z","hiringOrganization":{"@type":"Organization","name":"Quantiphi","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Denver","addressRegion":"CO","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"0e038c2d4ce74b1ad2fd235d"},"url":"https://jobsearcher.com/jobs/0e038c2d4ce74b1ad2fd235d"}}