{"schemaVersion":"jobsearcher.job.v1","id":"dae0e2a26bce1e57fac986f8","url":"https://jobsearcher.com/jobs/dae0e2a26bce1e57fac986f8","canonicalUrl":"https://jobsearcher.com/jobs/dae0e2a26bce1e57fac986f8","title":"GPU Kernel Developer - Fully Remote | Upto $90/hr","description":"Job DescriptionAbout the jobMercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.Position: GPU Kernel ExpertType: ContractCompensation: $70–$90/hourLocation: RemoteRole ResponsibilitiesEvaluate the quality and correctness of GPU/accelerator kernel development tasks for training and evaluating AI models.Assess numerical correctness, performance-benchmarking fairness, and task scoping across diverse kernel task types.Review compilation and runtime validity, providing clear, rubric-based written feedback.Collaborate with AI research teams to improve model training and evaluation processes.Work independently and asynchronously to meet deadlines and enhance AI model performance.QualificationsMust-Have3+ years of experience with CUDA, Triton, NKI, or Pallas (JAX).Strong understanding of numerical-correctness criteria for kernels.Experience with performance profiling and benchmarking (e.g., nsight, ncu).Familiarity with common compilation and runtime failure modes.Experience with at least three kernel task types.PreferredExperience with both NVIDIA GPU and custom-accelerator ecosystems.Background in compiler engineering, MLIR, or intermediate-representation lowering.Understanding of memory-hierarchy optimization.Contributions to kernel libraries like cuBLAS, cuDNN, or Triton.Application Process (Takes 20–30 mins to complete)Upload resumeAI interview based on your resumeSubmit formResources & SupportFor details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcomeFor any help or support, reach out to: support@mercor.comPS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.","company":"Mercor","rawCompany":"mercor","city":"New York","state":"NY","isRemote":true,"isActive":false,"createdAt":"2026-08-30T03:48:12.682Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-1251.00","title":"Computer Programmers","slug":"computer-programmers"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"GPU Kernel Developer - Fully Remote | Upto $90/hr","description":"Job DescriptionAbout the jobMercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.Position: GPU Kernel ExpertType: ContractCompensation: $70–$90/hourLocation: RemoteRole ResponsibilitiesEvaluate the quality and correctness of GPU/accelerator kernel development tasks for training and evaluating AI models.Assess numerical correctness, performance-benchmarking fairness, and task scoping across diverse kernel task types.Review compilation and runtime validity, providing clear, rubric-based written feedback.Collaborate with AI research teams to improve model training and evaluation processes.Work independently and asynchronously to meet deadlines and enhance AI model performance.QualificationsMust-Have3+ years of experience with CUDA, Triton, NKI, or Pallas (JAX).Strong understanding of numerical-correctness criteria for kernels.Experience with performance profiling and benchmarking (e.g., nsight, ncu).Familiarity with common compilation and runtime failure modes.Experience with at least three kernel task types.PreferredExperience with both NVIDIA GPU and custom-accelerator ecosystems.Background in compiler engineering, MLIR, or intermediate-representation lowering.Understanding of memory-hierarchy optimization.Contributions to kernel libraries like cuBLAS, cuDNN, or Triton.Application Process (Takes 20–30 mins to complete)Upload resumeAI interview based on your resumeSubmit formResources & SupportFor details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcomeFor any help or support, reach out to: support@mercor.comPS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.","datePosted":"2026-08-30T03:48:12.682Z","dateModified":"2026-08-30T03:48:12.682Z","hiringOrganization":{"@type":"Organization","name":"Mercor","sameAs":"https://jobsearcher.com"},"jobLocationType":"TELECOMMUTE","applicantLocationRequirements":{"@type":"Country","name":"US"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"New York","addressRegion":"NY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"dae0e2a26bce1e57fac986f8"},"url":"https://jobsearcher.com/jobs/dae0e2a26bce1e57fac986f8"}}