{"schemaVersion":"jobsearcher.job.v1","id":"39b8e5b6331b0d8e9bd1511c","url":"https://jobsearcher.com/jobs/39b8e5b6331b0d8e9bd1511c","canonicalUrl":"https://jobsearcher.com/jobs/39b8e5b6331b0d8e9bd1511c","title":"Agent Engineer","description":"About us\nRifa AI is building the AI agents platform for contact centers in regulated industries.\nEnterprises in these industries want AI agents handling their customer operations and mostly can't deploy them. It's not a model problem. Horizontal platforms lack governance, release processes, and change management, and in a domain where every call can be reviewed by a regulator, that's disqualifying. Building an AI agent has never been easier. Deploying one an enterprise can trust has never been harder. That harder problem is the one we work on.\nOur platform turns a company's written procedures into AI agents: voicebots that hold real-time conversations, take actions in the client's CRM, and stay within the limits the client has set. Every release is gated by an automated testing suite, and every conversation feeds a post-analysis platform. We're live in production today, handling debt collection calls for US financial services clients.\nThe engineering convictions behind it: evaluation methodology, not model capability, is the bottleneck. Every rule and decision trace becomes part of a company's context graph. Observability goes beyond logging. And our Agent Studio lets engineers, non-engineers, and auditors collaborate to build and improve AI agents with the right guardrails to achieve the expected business outcomes.\nRifa was founded by Sameer Fulzele (IIT Bombay). We're a small team of exceptionally capable, passionate engineers with paying enterprise clients and growing revenue, backed by Seaborne Capital, a founder-first firm of exceptional industry operators that works closely with its founders, and by angel investors who are veterans of enterprise software and operators in the accounts receivable industry.\nWhat you'll do\nAn Agent Engineer owns a client's voice or chat agent in production. Not a component of it, the whole thing: the procedure it follows, the instructions that govern how it speaks, the code that connects it to the client's systems, and the tests that prove it does what the documentation says. When a client says \"the bot offered a payment plan below our minimum,\" you're the person who works out why, fixes it, and explains the fix to the client. You'll ship changes that speak to real callers in your first two weeks.\nOwn a client delivery end to end. Take requirements from first conversation through pilot, production, and continuous iteration as procedures change, volumes grow, and models improve.\nEngineer the agent's behavior. Write and maintain the instruction sets that determine what the agent says and when. Instructions are versioned, tested, and reviewed like code, because a wrong word in a disclosure is a compliance incident, not a UX bug.\nBuild the evaluation gate. Automated tests that replay past conversations and check the agent follows the procedure. Nothing ships without passing them.\nDebug live conversations. Look up what the system knew at each moment, read the exact instructions the model was given, and work out why it responded the way it did.\nWork directly with clients. Sit in on calls with US enterprise clients, own the technical conversation, and watch your changes move their business metrics. Not many engineering roles put you this close to the people using what you build.\nExample projects\nRecent work by engineers on this team, all on live debt collection agents:\nBuild an intent identification layer for a voice agent, so every response is grounded in what the caller actually asked rather than what the model assumes, sharply reducing hallucinations on live calls\nExtend the negotiation flow so the agent offers payment plans only within the limits the client has set, with guardrails that make out-of-bounds offers impossible rather than just unlikely\nGrow the eval suite that replays real collection calls and verifies every legally required disclosure was delivered, word for word, before any release ships\nWhat we work with\nPython with FastAPI, PostgreSQL, Temporal for background workflows, WebSockets holding real-time conversations open, and STT and TTS providers on the speech side, with LLMs via OpenAI and similar underneath.\nEverything runs on Kubernetes with ArgoCD-driven GitOps deployments, SigNoz for observability, DeepEval driving the eval suite our CI runs before anything reaches production, Ory and OpenFGA for identity and access management, and a lot of beautifully built internal agent architecture underneath.\nWhat you'll bring\n2 to 5 years building and running production systems. You've been on call for something you built and debugged it under pressure.\nYou write Python another person can read, review others' work thoughtfully, and debug by forming a theory and testing it, not by changing things until they work. Reading code you didn't write and working out what it does is the single most important skill in the role.\nYou've worked with async code and a relational database.\nComfort working directly with customers to understand their needs and solve real-world problems: take vague feedback, ask the right questions, leave with a scoped change.\nStrong written communication. Agent instructions, client explanations, and incident writeups are all writing, and an agent is only as precise as the instructions behind it.\nEven better\nLLM systems in production: eval frameworks, agent tooling, RAG pipelines, structured prompting\nConversational AI experience, voice or chat: dialogue design, IVR systems, chatbots, speech interfaces\nA regulated industry: finance, healthcare, insurance, collections\nReal-time or telephony systems: WebSockets, streaming audio, Twilio or similar\nFounder or founding engineer experience\nExplicitly not required: prior \"AI engineering\" as a job title, or a computer science degree. Careful engineers who own outcomes pick this up fast.\nHow we work\nSmall teams per client, with real ownership. You'll have peers to pair with and review your work, and you'll review theirs. We review all code, we write things down, and we'd rather hear an honest \"I don't know yet\" than a confident wrong answer.\nOur values\nTrust: We do the right thing, especially when nobody is checking. Clients hand us regulated conversations with their own customers, and we earn that every day.\nTransparency: We write things down, share the real numbers, and say \"I don't know yet\" out loud, with each other and with clients.\nTechnically best solution: We choose what's right, not what's easiest or trendiest. When we notice we got it wrong, we fix it.\nDecisiveness: We decide quickly with the information we have, commit, and correct course fast when reality disagrees.\nSimplicity: We keep systems, processes, and words simple. Complexity is a cost we pay only when it clearly buys something.\nCompensation Range: $80K - $150K","company":"Rifa Ai","rawCompany":"rifa ai","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-13T14:43:49.881Z","occupations":[{"code":"17-2199.00","title":"Engineers, All Other","slug":"engineers-all-other"},{"code":"17-2199.08","title":"Robotics Engineers","slug":"robotics-engineers"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"561440","title":"Collection Agencies","slug":"collection-agencies"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Agent Engineer","description":"About us\nRifa AI is building the AI agents platform for contact centers in regulated industries.\nEnterprises in these industries want AI agents handling their customer operations and mostly can't deploy them. It's not a model problem. Horizontal platforms lack governance, release processes, and change management, and in a domain where every call can be reviewed by a regulator, that's disqualifying. Building an AI agent has never been easier. Deploying one an enterprise can trust has never been harder. That harder problem is the one we work on.\nOur platform turns a company's written procedures into AI agents: voicebots that hold real-time conversations, take actions in the client's CRM, and stay within the limits the client has set. Every release is gated by an automated testing suite, and every conversation feeds a post-analysis platform. We're live in production today, handling debt collection calls for US financial services clients.\nThe engineering convictions behind it: evaluation methodology, not model capability, is the bottleneck. Every rule and decision trace becomes part of a company's context graph. Observability goes beyond logging. And our Agent Studio lets engineers, non-engineers, and auditors collaborate to build and improve AI agents with the right guardrails to achieve the expected business outcomes.\nRifa was founded by Sameer Fulzele (IIT Bombay). We're a small team of exceptionally capable, passionate engineers with paying enterprise clients and growing revenue, backed by Seaborne Capital, a founder-first firm of exceptional industry operators that works closely with its founders, and by angel investors who are veterans of enterprise software and operators in the accounts receivable industry.\nWhat you'll do\nAn Agent Engineer owns a client's voice or chat agent in production. Not a component of it, the whole thing: the procedure it follows, the instructions that govern how it speaks, the code that connects it to the client's systems, and the tests that prove it does what the documentation says. When a client says \"the bot offered a payment plan below our minimum,\" you're the person who works out why, fixes it, and explains the fix to the client. You'll ship changes that speak to real callers in your first two weeks.\nOwn a client delivery end to end. Take requirements from first conversation through pilot, production, and continuous iteration as procedures change, volumes grow, and models improve.\nEngineer the agent's behavior. Write and maintain the instruction sets that determine what the agent says and when. Instructions are versioned, tested, and reviewed like code, because a wrong word in a disclosure is a compliance incident, not a UX bug.\nBuild the evaluation gate. Automated tests that replay past conversations and check the agent follows the procedure. Nothing ships without passing them.\nDebug live conversations. Look up what the system knew at each moment, read the exact instructions the model was given, and work out why it responded the way it did.\nWork directly with clients. Sit in on calls with US enterprise clients, own the technical conversation, and watch your changes move their business metrics. Not many engineering roles put you this close to the people using what you build.\nExample projects\nRecent work by engineers on this team, all on live debt collection agents:\nBuild an intent identification layer for a voice agent, so every response is grounded in what the caller actually asked rather than what the model assumes, sharply reducing hallucinations on live calls\nExtend the negotiation flow so the agent offers payment plans only within the limits the client has set, with guardrails that make out-of-bounds offers impossible rather than just unlikely\nGrow the eval suite that replays real collection calls and verifies every legally required disclosure was delivered, word for word, before any release ships\nWhat we work with\nPython with FastAPI, PostgreSQL, Temporal for background workflows, WebSockets holding real-time conversations open, and STT and TTS providers on the speech side, with LLMs via OpenAI and similar underneath.\nEverything runs on Kubernetes with ArgoCD-driven GitOps deployments, SigNoz for observability, DeepEval driving the eval suite our CI runs before anything reaches production, Ory and OpenFGA for identity and access management, and a lot of beautifully built internal agent architecture underneath.\nWhat you'll bring\n2 to 5 years building and running production systems. You've been on call for something you built and debugged it under pressure.\nYou write Python another person can read, review others' work thoughtfully, and debug by forming a theory and testing it, not by changing things until they work. Reading code you didn't write and working out what it does is the single most important skill in the role.\nYou've worked with async code and a relational database.\nComfort working directly with customers to understand their needs and solve real-world problems: take vague feedback, ask the right questions, leave with a scoped change.\nStrong written communication. Agent instructions, client explanations, and incident writeups are all writing, and an agent is only as precise as the instructions behind it.\nEven better\nLLM systems in production: eval frameworks, agent tooling, RAG pipelines, structured prompting\nConversational AI experience, voice or chat: dialogue design, IVR systems, chatbots, speech interfaces\nA regulated industry: finance, healthcare, insurance, collections\nReal-time or telephony systems: WebSockets, streaming audio, Twilio or similar\nFounder or founding engineer experience\nExplicitly not required: prior \"AI engineering\" as a job title, or a computer science degree. Careful engineers who own outcomes pick this up fast.\nHow we work\nSmall teams per client, with real ownership. You'll have peers to pair with and review your work, and you'll review theirs. We review all code, we write things down, and we'd rather hear an honest \"I don't know yet\" than a confident wrong answer.\nOur values\nTrust: We do the right thing, especially when nobody is checking. Clients hand us regulated conversations with their own customers, and we earn that every day.\nTransparency: We write things down, share the real numbers, and say \"I don't know yet\" out loud, with each other and with clients.\nTechnically best solution: We choose what's right, not what's easiest or trendiest. When we notice we got it wrong, we fix it.\nDecisiveness: We decide quickly with the information we have, commit, and correct course fast when reality disagrees.\nSimplicity: We keep systems, processes, and words simple. Complexity is a cost we pay only when it clearly buys something.\nCompensation Range: $80K - $150K","datePosted":"2026-08-13T14:43:49.881Z","dateModified":"2026-08-13T14:43:49.881Z","hiringOrganization":{"@type":"Organization","name":"Rifa Ai","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"39b8e5b6331b0d8e9bd1511c"},"url":"https://jobsearcher.com/jobs/39b8e5b6331b0d8e9bd1511c"}}