{"schemaVersion":"jobsearcher.job.v1","id":"7efe86c00a7ab3de320b6d46","url":"https://jobsearcher.com/jobs/7efe86c00a7ab3de320b6d46","canonicalUrl":"https://jobsearcher.com/jobs/7efe86c00a7ab3de320b6d46","title":"AI Research Engineer, Inference","description":"Job Title: AI Research Engineer – Deep Learning InferenceLocation: New York, NY | London, UKAbout The OpportunityJoin an industry-leading global quantitative technology organization at the forefront of machine learning innovation. We are seeking a High-Performance AI Research Engineer to drive speed and scalability across our real-time predictive modeling infrastructure. In this role, you will bridge the gap between machine learning research and hardware acceleration, building low-latency inference systems that process continuous global data streams to drive core business decisions.ResponsibilitiesDrive performance optimizations across all aspects of large-scale model execution, including custom kernel creation, data streaming pipelines, and novel hardware integration. Partner directly with machine learning researchers to co-design neural architectures optimized for extreme real-time execution. Author low-level primitives and custom operations to extract maximum compute throughput from modern hardware architectures. Evaluate, benchmark, and deploy novel hardware acceleration tech, ranging from off-the-shelf accelerators to custom specialized silicon. Formulate and execute engineering initiatives that address complex, non-obvious bottlenecks in ultra-low-latency deep learning inference. Requirements (Must-Have)At least two years of hands-on experience engineering production-grade deep learning systems within any complex domain (such as robotics, computer vision, audio, NLP, physics, or recommender platforms). Strong lower-level engineering foundation, including experience writing custom compute kernels (e.g., CUDA, Triton, Pallas, or CuTe DSLs). Proficiency with framework compilation internals and low-level runtime environments (PyTorch, JAX, XLA, or CUDA Graphs). Practical exposure to hardware acceleration tech, such as FPGAs, ASICs, or specialized AI processors. Proven ability to adapt algorithms and technical concepts across different domain applications. Preferred QualificationsExperience optimizing or serving Large Language Models (LLMs) and foundation architectures. Note: Prior background in quantitative finance or trading is explicitly NOT required. Compensation & BenefitsHighly competitive base salary, performance-based bonus incentive, and premium health/wellness benefits package. Equal Opportunity Employer.","company":"Objective Partners","rawCompany":"objective partners","city":"Chicago","state":"IL","isRemote":false,"isActive":false,"createdAt":"2026-08-19T12:53:06.437Z","occupations":[{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"17-2061.00","title":"Computer Hardware Engineers","slug":"computer-hardware-engineers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"334111","title":"Electronic Computer Manufacturing","slug":"electronic-computer-manufacturing"},{"code":"541715","title":"Research and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)","slug":"research-and-development-in-the-physical-engineering-and-life-sciences-except-nanotechnology-and-biotechnology"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI Research Engineer, Inference","description":"Job Title: AI Research Engineer – Deep Learning InferenceLocation: New York, NY | London, UKAbout The OpportunityJoin an industry-leading global quantitative technology organization at the forefront of machine learning innovation. We are seeking a High-Performance AI Research Engineer to drive speed and scalability across our real-time predictive modeling infrastructure. In this role, you will bridge the gap between machine learning research and hardware acceleration, building low-latency inference systems that process continuous global data streams to drive core business decisions.ResponsibilitiesDrive performance optimizations across all aspects of large-scale model execution, including custom kernel creation, data streaming pipelines, and novel hardware integration. Partner directly with machine learning researchers to co-design neural architectures optimized for extreme real-time execution. Author low-level primitives and custom operations to extract maximum compute throughput from modern hardware architectures. Evaluate, benchmark, and deploy novel hardware acceleration tech, ranging from off-the-shelf accelerators to custom specialized silicon. Formulate and execute engineering initiatives that address complex, non-obvious bottlenecks in ultra-low-latency deep learning inference. Requirements (Must-Have)At least two years of hands-on experience engineering production-grade deep learning systems within any complex domain (such as robotics, computer vision, audio, NLP, physics, or recommender platforms). Strong lower-level engineering foundation, including experience writing custom compute kernels (e.g., CUDA, Triton, Pallas, or CuTe DSLs). Proficiency with framework compilation internals and low-level runtime environments (PyTorch, JAX, XLA, or CUDA Graphs). Practical exposure to hardware acceleration tech, such as FPGAs, ASICs, or specialized AI processors. Proven ability to adapt algorithms and technical concepts across different domain applications. Preferred QualificationsExperience optimizing or serving Large Language Models (LLMs) and foundation architectures. Note: Prior background in quantitative finance or trading is explicitly NOT required. Compensation & BenefitsHighly competitive base salary, performance-based bonus incentive, and premium health/wellness benefits package. Equal Opportunity Employer.","datePosted":"2026-08-19T12:53:06.437Z","dateModified":"2026-08-19T12:53:06.437Z","hiringOrganization":{"@type":"Organization","name":"Objective Partners","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Chicago","addressRegion":"IL","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"7efe86c00a7ab3de320b6d46"},"url":"https://jobsearcher.com/jobs/7efe86c00a7ab3de320b6d46"}}