{"schemaVersion":"jobsearcher.job.v1","id":"91cf3f34418c7af962ef939c","url":"https://jobsearcher.com/jobs/91cf3f34418c7af962ef939c","canonicalUrl":"https://jobsearcher.com/jobs/91cf3f34418c7af962ef939c","title":"Applied Research - Forward-Deployed","description":"Own Your Intelligence\n\nPrime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team.\n\nOur platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own.\n\nPrime Intellect has raised $150M in total funding from Founders Fund, Radical Ventures, NVIDIA, and exceptional AI, infrastructure, and enterprise operators — including Andrej Karpathy, Dwarkesh Patel, and leaders and founders from Ramp, Perplexity, Harvey, Mercor, Zapier, Datadog, Cognition, OpenAI, Thinking Machines, Together AI, SemiAnalysis, LangChain, Browserbase, Cloudflare, Sierra, Databricks, Airbnb, OpenRouter, Standard Intelligence, Fleet, Core Auto, and more. We are looking for people who want to build at the intersection of frontier research, real infrastructure, and go-to-market for a category that does not fully exist yet.\n\nAbout the Role\n\nWe're looking for a Forward-Deployed Research Engineer (FDRE) to serve as the primary technical interface between Prime Intellect and our most important customers: AI companies, research labs, and enterprises running post-training and agentic RL on our platform.\n\nThis is not a traditional research role. You'll spend most of your time embedded with customers, understanding their models, workflows, and goals. Then, you'll translate those objectives into concrete training runs, environment designs, evaluation harnesses, and deployment recipes using the Lab stack. You are the person who makes the platform work in practice for real workloads.\n\nYou'll work closely with our research, product, and infrastructure teams to feed field insights back into the platform, shaping what we build next based on what customers actually need.\n\nWhat You'll Do\nCustomer Engagement & Technical Delivery\n\nEmbed directly with strategic customers to understand their agent architectures, failure modes, and product goals\n\nDesign and build custom RL environments, evaluation harnesses, and verifiers that capture what \"good\" looks like for each customer's domain\n\nArchitect agent scaffolding — tool use, multi-step reasoning, memory, sandbox execution — tailored to customer workflows\n\nConfigure and launch training runs on Lab, iterating on reward functions, rollout strategies, and evaluation criteria\n\nServe as the technical lead for engagements end-to-end: from discovery through deployed, improved models\n\nPlatform Feedback & Ecosystem\n\nIdentify repeatable patterns from customer engagements and codify them into reference implementations, templates, and documentation\n\nServe as the voice of the customer internally, shaping the roadmap for Lab, verifiers, the Environments Hub, and training infrastructure\n\nBuild high-quality examples and \"recipes\" that make it easy for new customers and open-source contributors to extend the stack\n\nContribute to technical content (blog posts, tutorials, case studies) that demonstrates real-world platform usage\n\nApplied Research & Experimentation\n\nDevelop novel evaluation methodologies for agentic behavior — multi-step reasoning, tool use correctness, recovery from failure, long-horizon task completion\n\nPrototype and iterate on agent harnesses for real-world tasks: code generation, workflow automation, document processing, and more\n\nExperiment with reward design, rubric construction, and environment shaping to improve training signal quality\n\nStay current on the frontier of agentic AI, evals, and post-training methods, and bring that knowledge directly into customer work\n\nWhat We're Looking For\n\nDeep hands-on experience building, evaluating, or deploying LLM-based agents in the past 1–2 years — you've seen what breaks in production and know what good evals look like\n\nStrong intuition for evaluation design: you can look at a customer's agent and quickly identify what to measure, how to construct a rubric, and where the reward signal is weak\n\nWorking understanding of RL and post-training concepts (GRPO, RLHF, reward modeling, SFT) — you don't need to have written a trainer from scratch, but you should understand what the knobs do and why they matter\n\nStrong Python skills and comfort with the modern AI stack (Hugging Face, inference engines, agent frameworks)\n\nExperience in a customer-facing or consulting-adjacent technical role, or as a technical founder — you're comfortable in a room with a customer's engineering team figuring out what to build\n\nExcellent written and verbal communication — you can write a clear environment spec, a compelling case study, and a useful Slack message to a frustrated customer\n\nHigh agency and comfort with ambiguity. You don't wait for specs; you scope the problem, ship a solution, and iterate\n\nNice-to-Haves\n\nExperience with agent frameworks and tooling (DSPy, LangGraph, MCP, Stagehand, browser automation)\n\nExperience building or running LLM evaluation pipelines at scale (benchmarks, synthetic data generation, model grading)\n\nResearch experience — publications, open-source contributions, or benchmarks in ML/RL/agents\n\nFamiliarity with sandbox/code execution environments for agent evaluation\n\nWeb programming experience (React, TypeScript, Next.js) for building demos and customer-facing tooling\n\nWhat We Offer\n\nCash Compensation Range of $150-300k + equity incentives\n\nFlexible Work (San Francisco or hybrid-remote)\n\nVisa Sponsorship & relocation support\n\nProfessional Development budget\n\nTeam Off-sites & conference attendance\n\nGrowth Opportunity\n\nYou’ll join a mission-driven team working at the frontier of open, superintelligence infra. In this role, you’ll have the opportunity to:\n\nShape the evolution of agent-driven solutions—from research breakthroughs to production systems used by real customers.\n\nCollaborate with leading researchers, engineers, and partners pushing the boundaries of RL and post-training.\n\nGrow with a fast-moving organization where your contributions directly influence both the technical direction and the broader AI ecosystem.\n\nIf you’re excited to move fast, build boldly, and help define how agentic AI is developed and deployed, we’d love to hear from you.\n\nReady to build the open superintelligence infrastructure of tomorrow?\nApply now to help us make powerful, open AGI accessible to everyone.","company":"Primeintellect","rawCompany":"primeintellect","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-04T09:15:00.120Z","occupations":[{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Applied Research - Forward-Deployed","description":"Own Your Intelligence\n\nPrime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team.\n\nOur platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own.\n\nPrime Intellect has raised $150M in total funding from Founders Fund, Radical Ventures, NVIDIA, and exceptional AI, infrastructure, and enterprise operators — including Andrej Karpathy, Dwarkesh Patel, and leaders and founders from Ramp, Perplexity, Harvey, Mercor, Zapier, Datadog, Cognition, OpenAI, Thinking Machines, Together AI, SemiAnalysis, LangChain, Browserbase, Cloudflare, Sierra, Databricks, Airbnb, OpenRouter, Standard Intelligence, Fleet, Core Auto, and more. We are looking for people who want to build at the intersection of frontier research, real infrastructure, and go-to-market for a category that does not fully exist yet.\n\nAbout the Role\n\nWe're looking for a Forward-Deployed Research Engineer (FDRE) to serve as the primary technical interface between Prime Intellect and our most important customers: AI companies, research labs, and enterprises running post-training and agentic RL on our platform.\n\nThis is not a traditional research role. You'll spend most of your time embedded with customers, understanding their models, workflows, and goals. Then, you'll translate those objectives into concrete training runs, environment designs, evaluation harnesses, and deployment recipes using the Lab stack. You are the person who makes the platform work in practice for real workloads.\n\nYou'll work closely with our research, product, and infrastructure teams to feed field insights back into the platform, shaping what we build next based on what customers actually need.\n\nWhat You'll Do\nCustomer Engagement & Technical Delivery\n\nEmbed directly with strategic customers to understand their agent architectures, failure modes, and product goals\n\nDesign and build custom RL environments, evaluation harnesses, and verifiers that capture what \"good\" looks like for each customer's domain\n\nArchitect agent scaffolding — tool use, multi-step reasoning, memory, sandbox execution — tailored to customer workflows\n\nConfigure and launch training runs on Lab, iterating on reward functions, rollout strategies, and evaluation criteria\n\nServe as the technical lead for engagements end-to-end: from discovery through deployed, improved models\n\nPlatform Feedback & Ecosystem\n\nIdentify repeatable patterns from customer engagements and codify them into reference implementations, templates, and documentation\n\nServe as the voice of the customer internally, shaping the roadmap for Lab, verifiers, the Environments Hub, and training infrastructure\n\nBuild high-quality examples and \"recipes\" that make it easy for new customers and open-source contributors to extend the stack\n\nContribute to technical content (blog posts, tutorials, case studies) that demonstrates real-world platform usage\n\nApplied Research & Experimentation\n\nDevelop novel evaluation methodologies for agentic behavior — multi-step reasoning, tool use correctness, recovery from failure, long-horizon task completion\n\nPrototype and iterate on agent harnesses for real-world tasks: code generation, workflow automation, document processing, and more\n\nExperiment with reward design, rubric construction, and environment shaping to improve training signal quality\n\nStay current on the frontier of agentic AI, evals, and post-training methods, and bring that knowledge directly into customer work\n\nWhat We're Looking For\n\nDeep hands-on experience building, evaluating, or deploying LLM-based agents in the past 1–2 years — you've seen what breaks in production and know what good evals look like\n\nStrong intuition for evaluation design: you can look at a customer's agent and quickly identify what to measure, how to construct a rubric, and where the reward signal is weak\n\nWorking understanding of RL and post-training concepts (GRPO, RLHF, reward modeling, SFT) — you don't need to have written a trainer from scratch, but you should understand what the knobs do and why they matter\n\nStrong Python skills and comfort with the modern AI stack (Hugging Face, inference engines, agent frameworks)\n\nExperience in a customer-facing or consulting-adjacent technical role, or as a technical founder — you're comfortable in a room with a customer's engineering team figuring out what to build\n\nExcellent written and verbal communication — you can write a clear environment spec, a compelling case study, and a useful Slack message to a frustrated customer\n\nHigh agency and comfort with ambiguity. You don't wait for specs; you scope the problem, ship a solution, and iterate\n\nNice-to-Haves\n\nExperience with agent frameworks and tooling (DSPy, LangGraph, MCP, Stagehand, browser automation)\n\nExperience building or running LLM evaluation pipelines at scale (benchmarks, synthetic data generation, model grading)\n\nResearch experience — publications, open-source contributions, or benchmarks in ML/RL/agents\n\nFamiliarity with sandbox/code execution environments for agent evaluation\n\nWeb programming experience (React, TypeScript, Next.js) for building demos and customer-facing tooling\n\nWhat We Offer\n\nCash Compensation Range of $150-300k + equity incentives\n\nFlexible Work (San Francisco or hybrid-remote)\n\nVisa Sponsorship & relocation support\n\nProfessional Development budget\n\nTeam Off-sites & conference attendance\n\nGrowth Opportunity\n\nYou’ll join a mission-driven team working at the frontier of open, superintelligence infra. In this role, you’ll have the opportunity to:\n\nShape the evolution of agent-driven solutions—from research breakthroughs to production systems used by real customers.\n\nCollaborate with leading researchers, engineers, and partners pushing the boundaries of RL and post-training.\n\nGrow with a fast-moving organization where your contributions directly influence both the technical direction and the broader AI ecosystem.\n\nIf you’re excited to move fast, build boldly, and help define how agentic AI is developed and deployed, we’d love to hear from you.\n\nReady to build the open superintelligence infrastructure of tomorrow?\nApply now to help us make powerful, open AGI accessible to everyone.","datePosted":"2026-09-04T09:15:00.120Z","dateModified":"2026-09-04T09:15:00.120Z","hiringOrganization":{"@type":"Organization","name":"Primeintellect","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"91cf3f34418c7af962ef939c"},"url":"https://jobsearcher.com/jobs/91cf3f34418c7af962ef939c"}}