{"schemaVersion":"jobsearcher.job.v1","id":"53aa1f853e8bfecd0f077920","url":"https://jobsearcher.com/jobs/53aa1f853e8bfecd0f077920","canonicalUrl":"https://jobsearcher.com/jobs/53aa1f853e8bfecd0f077920","title":"Staff Engineer - Core Engineering","description":"About TrueFoundry Every production AI system, whether it's powering customer support, writing code, analyzing financial data, or diagnosing medical conditions, needs the same foundational infrastructure. A way to route between models. A way to manage tools and integrate them securely. A way to orchestrate agents and enforce governance. A unified compute layer to run it all.\nThat infrastructure layer is being built right now. We're TrueFoundry, and we're building it. We're looking for a Staff Engineer – Core Engineering to join the team.\nThe Problem We're Solving Companies are moving beyond simple chatbots to production agentic systems. These systems route between OpenAI, Anthropic, Google, and self-hosted models. They integrate dozens of tools via protocols like MCP. They orchestrate multi-agent workflows where agents coordinate with other agents.\nThe infrastructure to support this doesn't exist yet. You can't just duct-tape together a few API calls and call it production‑ready.\nYou need a control plane that handles:\nIntelligent routing with observability, cost policies, and fallback logic\nCentralized tool and MCP server management with security and lifecycle controls\nAgent orchestration with governance and guardrails\nA unified compute layer to run self-hosted models, custom tools, and agents\nWe've built two products to solve this:\nAI Gateway is the control plane, five composable components (Prompts, LLM Gateway, MCP Gateway, Guardrails, Agent Gateway) that handle routing, orchestration, and governance.\nAI Deploy is the compute layer, a Kubernetes‑based platform that abstracts ML workloads as standard software primitives, so everything runs on unified infrastructure.\nWe're Series A, backed by Intel Capital and Sequoia. Companies like CVS, Mastercard, Siemens, Paytm, Synopsys, and Zscaler run production AI workloads on our platform.\nThe Role We are seeking a Staff / Principal Engineer to join our Core Engineering team as a senior technical leader based in the United States. You will:\nSolve some of the most complex Engineering problems and drive it alongside a team of engineers & ML researchers.\nBuild a deep, holistic understanding of the TrueFoundry platform across all components and shape the product vision and implementation.\nAct as the technical face of engineering for customer-related discussions and escalations.\nGuide and unblock engineers across projects in the US region.\nPartner closely with our CTO and India‑based engineering team to drive system design, architecture, and implementation of complex products.\nLead technical design, critical customer problem‑solving, and platform scalability initiatives end‑to‑end.\nThis is a high‑ownership, high‑impact role designed for an engineer who loves combining world‑class systems thinking with real‑world execution.\nWhat You’ll Do Develop deep expertise across TrueFoundry’s platform stack — infrastructure, deployment systems, LLM/ML orchestration, observability, cost optimization, and more.\nDrive the system architecture and design for complex, distributed, cloud‑native systems.\nAct as the technical point‑of‑contact for enterprise customer engineering needs and escalations.\nLead and participate in design reviews, code reviews, and critical incident responses.\nCollaborate closely with the CTO on architectural decisions, scaling strategies, and technical roadmap prioritization.\nGuide and mentor US‑based engineers across multiple initiatives, helping them deliver high‑quality, scalable systems.\nIdentify and drive technical debt cleanup, performance improvements, and resilience upgrades across the platform.\nBring a product engineering mindset, ensuring that customer needs and feedback translate into scalable engineering solutions.\nWho You Are 8+ years of strong backend / systems engineering experience at top technology companies or startups.\nDeep expertise in distributed systems, cloud‑native architectures, and scalable system design.\nStrong working knowledge of Kubernetes, containerized workloads, and infrastructure engineering.\nPractical experience building or deploying ML/GenAI applications (or closely working with ML/DS teams).\nSkilled in programming languages such as Python, Go, or typescript.\nSolid understanding of system observability, resiliency design, and SRE practices.\nStrong technical leadership and communication skills — able to work with both customers and engineering teams.\nAbility to think strategically while also executing hands‑on when required.\nBonus Experience supporting enterprise deployments of AI/ML infrastructure , model training , or inference systems .\nWhy Join TrueFoundry? Work directly with ex‑Facebook engineers and founders from IIT Kharagpur, UC Berkeley, and YCombinator alumni.\nFirst‑hand exposure to building and scaling a deep‑tech startup—insights you’ll carry if you want to start your own one day.\nBe part of a fearlessly experimental culture focused on customer success and long‑term impact.\nFlexible hours, learning credits, and the opportunity to work shoulder‑to‑shoulder with the co‑founders (Abhishek & Nikunj).\n\n#J-18808-Ljbffr","company":"Truefoundry","rawCompany":"truefoundry","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-07-16T03:29:23.029Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"17-2199.00","title":"Engineers, All Other","slug":"engineers-all-other"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff Engineer - Core Engineering","description":"About TrueFoundry Every production AI system, whether it's powering customer support, writing code, analyzing financial data, or diagnosing medical conditions, needs the same foundational infrastructure. A way to route between models. A way to manage tools and integrate them securely. A way to orchestrate agents and enforce governance. A unified compute layer to run it all.\nThat infrastructure layer is being built right now. We're TrueFoundry, and we're building it. We're looking for a Staff Engineer – Core Engineering to join the team.\nThe Problem We're Solving Companies are moving beyond simple chatbots to production agentic systems. These systems route between OpenAI, Anthropic, Google, and self-hosted models. They integrate dozens of tools via protocols like MCP. They orchestrate multi-agent workflows where agents coordinate with other agents.\nThe infrastructure to support this doesn't exist yet. You can't just duct-tape together a few API calls and call it production‑ready.\nYou need a control plane that handles:\nIntelligent routing with observability, cost policies, and fallback logic\nCentralized tool and MCP server management with security and lifecycle controls\nAgent orchestration with governance and guardrails\nA unified compute layer to run self-hosted models, custom tools, and agents\nWe've built two products to solve this:\nAI Gateway is the control plane, five composable components (Prompts, LLM Gateway, MCP Gateway, Guardrails, Agent Gateway) that handle routing, orchestration, and governance.\nAI Deploy is the compute layer, a Kubernetes‑based platform that abstracts ML workloads as standard software primitives, so everything runs on unified infrastructure.\nWe're Series A, backed by Intel Capital and Sequoia. Companies like CVS, Mastercard, Siemens, Paytm, Synopsys, and Zscaler run production AI workloads on our platform.\nThe Role We are seeking a Staff / Principal Engineer to join our Core Engineering team as a senior technical leader based in the United States. You will:\nSolve some of the most complex Engineering problems and drive it alongside a team of engineers & ML researchers.\nBuild a deep, holistic understanding of the TrueFoundry platform across all components and shape the product vision and implementation.\nAct as the technical face of engineering for customer-related discussions and escalations.\nGuide and unblock engineers across projects in the US region.\nPartner closely with our CTO and India‑based engineering team to drive system design, architecture, and implementation of complex products.\nLead technical design, critical customer problem‑solving, and platform scalability initiatives end‑to‑end.\nThis is a high‑ownership, high‑impact role designed for an engineer who loves combining world‑class systems thinking with real‑world execution.\nWhat You’ll Do Develop deep expertise across TrueFoundry’s platform stack — infrastructure, deployment systems, LLM/ML orchestration, observability, cost optimization, and more.\nDrive the system architecture and design for complex, distributed, cloud‑native systems.\nAct as the technical point‑of‑contact for enterprise customer engineering needs and escalations.\nLead and participate in design reviews, code reviews, and critical incident responses.\nCollaborate closely with the CTO on architectural decisions, scaling strategies, and technical roadmap prioritization.\nGuide and mentor US‑based engineers across multiple initiatives, helping them deliver high‑quality, scalable systems.\nIdentify and drive technical debt cleanup, performance improvements, and resilience upgrades across the platform.\nBring a product engineering mindset, ensuring that customer needs and feedback translate into scalable engineering solutions.\nWho You Are 8+ years of strong backend / systems engineering experience at top technology companies or startups.\nDeep expertise in distributed systems, cloud‑native architectures, and scalable system design.\nStrong working knowledge of Kubernetes, containerized workloads, and infrastructure engineering.\nPractical experience building or deploying ML/GenAI applications (or closely working with ML/DS teams).\nSkilled in programming languages such as Python, Go, or typescript.\nSolid understanding of system observability, resiliency design, and SRE practices.\nStrong technical leadership and communication skills — able to work with both customers and engineering teams.\nAbility to think strategically while also executing hands‑on when required.\nBonus Experience supporting enterprise deployments of AI/ML infrastructure , model training , or inference systems .\nWhy Join TrueFoundry? Work directly with ex‑Facebook engineers and founders from IIT Kharagpur, UC Berkeley, and YCombinator alumni.\nFirst‑hand exposure to building and scaling a deep‑tech startup—insights you’ll carry if you want to start your own one day.\nBe part of a fearlessly experimental culture focused on customer success and long‑term impact.\nFlexible hours, learning credits, and the opportunity to work shoulder‑to‑shoulder with the co‑founders (Abhishek & Nikunj).\n\n#J-18808-Ljbffr","datePosted":"2026-07-16T03:29:23.029Z","dateModified":"2026-07-16T03:29:23.029Z","hiringOrganization":{"@type":"Organization","name":"Truefoundry","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"53aa1f853e8bfecd0f077920"},"url":"https://jobsearcher.com/jobs/53aa1f853e8bfecd0f077920"}}