{"schemaVersion":"jobsearcher.job.v1","id":"dfeceb3da7c94bae8419fa32","url":"https://jobsearcher.com/jobs/dfeceb3da7c94bae8419fa32","canonicalUrl":"https://jobsearcher.com/jobs/dfeceb3da7c94bae8419fa32","title":"Parallel Computing Engineer","description":"Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.Job TitleParallel Computing EngineerLocation: 100% Remote (U.S.)Position Type: Full-time, Direct W2Salary Range: $130,000–$180,000 AnnuallyExperience Required: 10+ YearsSponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.Job SummaryBright Vision Technologies is seeking a highly experienced Parallel Computing Engineer with 10+ years of experience in High-Performance Computing (HPC), GPU programming, and parallel computing to optimize AI, machine learning, and scientific computing workloads. The ideal candidate will possess deep expertise in CUDA, GPU architecture, distributed computing, performance optimization, and large-scale AI infrastructure, with a proven track record of designing high-performance computing solutions for enterprise and research environments.Key ResponsibilitiesDesign, develop, and optimize high-performance CUDA kernels for AI, deep learning, and scientific computing applicationsAnalyze, profile, and optimize GPU workloads using NVIDIA Nsight Systems, Nsight Compute, CUDA Profiler, and related performance analysis toolsOptimize GPU memory management, kernel execution, multi-GPU scaling, and distributed computing performanceDesign scalable distributed training and inference architectures using NCCL, MPI, CUDA-aware communication libraries, and high-performance networking technologiesDevelop custom GPU operators and optimized kernels for PyTorch, JAX, Triton, TensorFlow, or similar AI frameworksImprove training and inference performance for large language models (LLMs), deep learning, and high-performance AI workloadsCollaborate with AI researchers, ML engineers, and software architects to accelerate production AI applicationsBuild automated benchmarking frameworks, performance regression testing, and optimization pipelinesEvaluate emerging GPU technologies, programming models, and accelerator architectures to improve computational efficiencyMentor engineers and provide technical leadership in GPU optimization, HPC architecture, and parallel programming best practicesRequired QualificationsBachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical discipline10+ years of professional experience in GPU programming, High-Performance Computing (HPC), or parallel computingExpert-level proficiency in CUDA C/C++, GPU architecture, and massively parallel programming techniquesExtensive experience with NCCL, MPI, CUDA-aware MPI, and distributed GPU communication frameworksStrong understanding of GPU memory hierarchy, kernel optimization, occupancy tuning, and performance analysisHands-on experience integrating custom GPU kernels into PyTorch, TensorFlow, JAX, Triton, or other machine learning frameworksStrong C/C++ programming skills with expertise in debugging, profiling, and performance optimizationExperience developing scalable AI or HPC solutions on cloud platforms or large GPU clustersExcellent analytical, communication, collaboration, and technical leadership skillsPreferred QualificationsExperience with Triton, CUTLASS, TensorRT, FasterTransformer, vLLM, DeepSpeed, or similar GPU optimization frameworksKnowledge of LLVM, MLIR, compiler optimization techniques, or code generation technologiesExperience with large-scale distributed AI training, model parallelism, pipeline parallelism, and inference optimizationFamiliarity with cloud-based GPU infrastructure on AWS, Microsoft Azure, or Google Cloud Platform (GCP)Contributions to open-source GPU libraries, research publications, patents, or technical presentationsExperience with emerging accelerator technologies such as AMD ROCm, Intel How To ApplyWould you like to know more about this opportunity? For immediate consideration, please send your resume to jaya@bvteck.com or contact us at (908) 505-3545. Learn more about Bright Vision Technologies at www.bvteck.com.Bright Vision Technologies is an Equal Opportunity Employer.Equal Employment Opportunity (EEO) StatementBright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.Powered by JazzHR678TDStXzV","company":"Bright Vision Technologies","rawCompany":"bright vision technologies","city":"Apex","state":"NC","isRemote":false,"isActive":false,"createdAt":"2026-08-05T09:44:39.031Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Parallel Computing Engineer","description":"Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.Job TitleParallel Computing EngineerLocation: 100% Remote (U.S.)Position Type: Full-time, Direct W2Salary Range: $130,000–$180,000 AnnuallyExperience Required: 10+ YearsSponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.Job SummaryBright Vision Technologies is seeking a highly experienced Parallel Computing Engineer with 10+ years of experience in High-Performance Computing (HPC), GPU programming, and parallel computing to optimize AI, machine learning, and scientific computing workloads. The ideal candidate will possess deep expertise in CUDA, GPU architecture, distributed computing, performance optimization, and large-scale AI infrastructure, with a proven track record of designing high-performance computing solutions for enterprise and research environments.Key ResponsibilitiesDesign, develop, and optimize high-performance CUDA kernels for AI, deep learning, and scientific computing applicationsAnalyze, profile, and optimize GPU workloads using NVIDIA Nsight Systems, Nsight Compute, CUDA Profiler, and related performance analysis toolsOptimize GPU memory management, kernel execution, multi-GPU scaling, and distributed computing performanceDesign scalable distributed training and inference architectures using NCCL, MPI, CUDA-aware communication libraries, and high-performance networking technologiesDevelop custom GPU operators and optimized kernels for PyTorch, JAX, Triton, TensorFlow, or similar AI frameworksImprove training and inference performance for large language models (LLMs), deep learning, and high-performance AI workloadsCollaborate with AI researchers, ML engineers, and software architects to accelerate production AI applicationsBuild automated benchmarking frameworks, performance regression testing, and optimization pipelinesEvaluate emerging GPU technologies, programming models, and accelerator architectures to improve computational efficiencyMentor engineers and provide technical leadership in GPU optimization, HPC architecture, and parallel programming best practicesRequired QualificationsBachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical discipline10+ years of professional experience in GPU programming, High-Performance Computing (HPC), or parallel computingExpert-level proficiency in CUDA C/C++, GPU architecture, and massively parallel programming techniquesExtensive experience with NCCL, MPI, CUDA-aware MPI, and distributed GPU communication frameworksStrong understanding of GPU memory hierarchy, kernel optimization, occupancy tuning, and performance analysisHands-on experience integrating custom GPU kernels into PyTorch, TensorFlow, JAX, Triton, or other machine learning frameworksStrong C/C++ programming skills with expertise in debugging, profiling, and performance optimizationExperience developing scalable AI or HPC solutions on cloud platforms or large GPU clustersExcellent analytical, communication, collaboration, and technical leadership skillsPreferred QualificationsExperience with Triton, CUTLASS, TensorRT, FasterTransformer, vLLM, DeepSpeed, or similar GPU optimization frameworksKnowledge of LLVM, MLIR, compiler optimization techniques, or code generation technologiesExperience with large-scale distributed AI training, model parallelism, pipeline parallelism, and inference optimizationFamiliarity with cloud-based GPU infrastructure on AWS, Microsoft Azure, or Google Cloud Platform (GCP)Contributions to open-source GPU libraries, research publications, patents, or technical presentationsExperience with emerging accelerator technologies such as AMD ROCm, Intel How To ApplyWould you like to know more about this opportunity? For immediate consideration, please send your resume to jaya@bvteck.com or contact us at (908) 505-3545. Learn more about Bright Vision Technologies at www.bvteck.com.Bright Vision Technologies is an Equal Opportunity Employer.Equal Employment Opportunity (EEO) StatementBright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.Powered by JazzHR678TDStXzV","datePosted":"2026-08-05T09:44:39.031Z","dateModified":"2026-08-05T09:44:39.031Z","hiringOrganization":{"@type":"Organization","name":"Bright Vision Technologies","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Apex","addressRegion":"NC","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"dfeceb3da7c94bae8419fa32"},"url":"https://jobsearcher.com/jobs/dfeceb3da7c94bae8419fa32"}}