{"schemaVersion":"jobsearcher.job.v1","id":"dce8ba11c794cb16df52dbc6","url":"https://jobsearcher.com/jobs/dce8ba11c794cb16df52dbc6","canonicalUrl":"https://jobsearcher.com/jobs/dce8ba11c794cb16df52dbc6","title":"Research Engineer - Reinforcement Learning","description":"About the Role\nAs a Research Engineer in our Reasoning team, you'll play a crucial role in shaping our technological direction, focusing on our test-time compute scaling research ideas. This role is ideal for individuals who enjoy working with synthetic data and teaching LLMs reasoning abilities. You will be contributing to Prime Intellect's mission of building the open superintelligence stack, from frontier agentic models to the infrastructure that enables anyone to create, train, and deploy them. This involves aggregating and orchestrating global compute into a single control plane and pairing it with a full RL post-training stack, including environments, secure sandboxes, verifiable evals, and an async RL trainer.\n\nResponsibilities\n\nLead and participate in novel research to build a massive scale synthetic data generation pipeline and orchestration solution.\n\nOptimize the performance, cost, and resource utilization of AI inference workloads by leveraging the most recent advances for compute & memory optimization techniques.\n\nContribute to the development of our open-source libraries and frameworks for synthetic data generation and distributed RL frameworks.\n\nPublish research in top-tier AI conferences such as ICML & NeurIPS.\n\nDistill highly technical project outcomes in layman approachable technical blogs to our customers and developers.\n\nStay up-to-date with the latest advancements in AI/ML infrastructure and tools, synthetic data generation research and proactively identify opportunities to enhance our platform's capabilities and user experience.\n\nRequirements\n\nStrong background in AI/ML engineering , with extensive experience in designing and implementing end-to-end pipelines for the inference or training of large-scale AI models.\n\nDeep expertise in distributed inference techniques and frameworks (e.g., vllm , sglang ) for optimizing the performance and scalability of AI workloads.\n\nSolid understanding of MLOps best practices, including model versioning , experiment tracking , and continuous integration/deployment (CI/CD) pipelines .\n\nPassion for advancing the state-of-the-art in reasoning and democratizing access to AI capabilities for researchers, developers, and businesses worldwide.\n\n#J-18808-Ljbffr","company":"Gravity Engineering Services","rawCompany":"gravity engineering services","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-07-16T03:15:14.706Z","occupations":[{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541715","title":"Research and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)","slug":"research-and-development-in-the-physical-engineering-and-life-sciences-except-nanotechnology-and-biotechnology"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Research Engineer - Reinforcement Learning","description":"About the Role\nAs a Research Engineer in our Reasoning team, you'll play a crucial role in shaping our technological direction, focusing on our test-time compute scaling research ideas. This role is ideal for individuals who enjoy working with synthetic data and teaching LLMs reasoning abilities. You will be contributing to Prime Intellect's mission of building the open superintelligence stack, from frontier agentic models to the infrastructure that enables anyone to create, train, and deploy them. This involves aggregating and orchestrating global compute into a single control plane and pairing it with a full RL post-training stack, including environments, secure sandboxes, verifiable evals, and an async RL trainer.\n\nResponsibilities\n\nLead and participate in novel research to build a massive scale synthetic data generation pipeline and orchestration solution.\n\nOptimize the performance, cost, and resource utilization of AI inference workloads by leveraging the most recent advances for compute & memory optimization techniques.\n\nContribute to the development of our open-source libraries and frameworks for synthetic data generation and distributed RL frameworks.\n\nPublish research in top-tier AI conferences such as ICML & NeurIPS.\n\nDistill highly technical project outcomes in layman approachable technical blogs to our customers and developers.\n\nStay up-to-date with the latest advancements in AI/ML infrastructure and tools, synthetic data generation research and proactively identify opportunities to enhance our platform's capabilities and user experience.\n\nRequirements\n\nStrong background in AI/ML engineering , with extensive experience in designing and implementing end-to-end pipelines for the inference or training of large-scale AI models.\n\nDeep expertise in distributed inference techniques and frameworks (e.g., vllm , sglang ) for optimizing the performance and scalability of AI workloads.\n\nSolid understanding of MLOps best practices, including model versioning , experiment tracking , and continuous integration/deployment (CI/CD) pipelines .\n\nPassion for advancing the state-of-the-art in reasoning and democratizing access to AI capabilities for researchers, developers, and businesses worldwide.\n\n#J-18808-Ljbffr","datePosted":"2026-07-16T03:15:14.706Z","dateModified":"2026-07-16T03:15:14.706Z","hiringOrganization":{"@type":"Organization","name":"Gravity Engineering Services","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"dce8ba11c794cb16df52dbc6"},"url":"https://jobsearcher.com/jobs/dce8ba11c794cb16df52dbc6"}}