{"schemaVersion":"jobsearcher.job.v1","id":"87917469ffa35e9cd6049910","url":"https://jobsearcher.com/jobs/87917469ffa35e9cd6049910","canonicalUrl":"https://jobsearcher.com/jobs/87917469ffa35e9cd6049910","title":"Machine Learning Engineer","description":"Compensation: $200K–$290K + EquityLocation: San Jose, CaliforniaAn innovative AI company is expanding its systems engineering team and is searching for an engineer who enjoys making large language models faster, more scalable, and more efficient in production.If your interests include GPU optimization, distributed computing, and extracting every ounce of performance from modern hardware, this role offers the opportunity to work on some of today's most demanding inference challenges.ResponsibilitiesYour primary focus will be improving the speed and efficiency of production AI systems. You'll evaluate performance bottlenecks, build benchmarking tools, optimize inference pipelines, and help scale distributed GPU environments.Working closely with research and infrastructure teams, you'll transform new modeling techniques into reliable production systems while continuously improving latency, throughput, and hardware utilization.We're Looking For Someone Who HasA degree in Computer Science, Electrical Engineering, or a related disciplineStrong Python and C++ (CUDA preferred) programming skillsKnowledge of modern LLM serving technologies such as vLLM, SGLang, PyTorch, or comparable frameworksUnderstanding of GPU architecture and parallel computingExperience with model serving, distributed inference, quantization, or batching strategiesStrong profiling, debugging, and systems optimization skillsAdditional Experience That Stands OutRay or similar distributed computing frameworksPerformance tuning at the systems or kernel levelHigh-performance computing environmentsWhat's OfferedCompetitive salary with meaningful equity participation401(k)Unlimited PTOModern engineering workspace with premium employee amenitiesOpportunity to solve technically complex AI infrastructure challenges alongside a highly experienced engineering team #J-18808-Ljbffr","company":"Ic Resources","rawCompany":"ic resources","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-05T23:32:08.257Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Machine Learning Engineer","description":"Compensation: $200K–$290K + EquityLocation: San Jose, CaliforniaAn innovative AI company is expanding its systems engineering team and is searching for an engineer who enjoys making large language models faster, more scalable, and more efficient in production.If your interests include GPU optimization, distributed computing, and extracting every ounce of performance from modern hardware, this role offers the opportunity to work on some of today's most demanding inference challenges.ResponsibilitiesYour primary focus will be improving the speed and efficiency of production AI systems. You'll evaluate performance bottlenecks, build benchmarking tools, optimize inference pipelines, and help scale distributed GPU environments.Working closely with research and infrastructure teams, you'll transform new modeling techniques into reliable production systems while continuously improving latency, throughput, and hardware utilization.We're Looking For Someone Who HasA degree in Computer Science, Electrical Engineering, or a related disciplineStrong Python and C++ (CUDA preferred) programming skillsKnowledge of modern LLM serving technologies such as vLLM, SGLang, PyTorch, or comparable frameworksUnderstanding of GPU architecture and parallel computingExperience with model serving, distributed inference, quantization, or batching strategiesStrong profiling, debugging, and systems optimization skillsAdditional Experience That Stands OutRay or similar distributed computing frameworksPerformance tuning at the systems or kernel levelHigh-performance computing environmentsWhat's OfferedCompetitive salary with meaningful equity participation401(k)Unlimited PTOModern engineering workspace with premium employee amenitiesOpportunity to solve technically complex AI infrastructure challenges alongside a highly experienced engineering team #J-18808-Ljbffr","datePosted":"2026-08-05T23:32:08.257Z","dateModified":"2026-08-05T23:32:08.257Z","hiringOrganization":{"@type":"Organization","name":"Ic Resources","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"87917469ffa35e9cd6049910"},"url":"https://jobsearcher.com/jobs/87917469ffa35e9cd6049910"}}