Technical Program Manager
About AI FabrikAI Fabrik builds an edge inference delivery network for high-performance tokens, with faster time-to-market from grid to tokens. Our mission is to build the inference infrastructure we wished every enterprise already had — close to users, close to the cloud, and extremely resilient for real-time workloads. We are builders, architects, engineers, and researchers with hands-on experience in real-world AI deployment in production, and decades of data center experience that taught us exactly what needs to change.AI Fabrik was incubated inside Gruve and backed by MayfieldXora (Temasek), Acclimate VenturesCisco Investments — existing investors from Gruve who followed us into this new chapter. We are deploying five initial production sites, with the first one coming online in July 2026.About the RoleThis role sits at the intersection of physical infrastructure deployment and software engineering. Our programs span the buildout of hardware sites across dozens of locations — facilities, power, networking, and compute — alongside the delivery of a deep software stack that must be production-ready when each site comes online.The multiple related workstreams run in parallel and are tightly coupled. Hardware procurement and site construction operate on fixed schedules with limited room for adjustment. Software delivery must be planned and tracked with the same discipline. The TPM in this role is accountable for keeping both on track and ensuring the dependencies between them are managed proactively, not discovered late.Key ResponsibilitiesOwn the end-to-end planning and execution of large-scale technical programs: define scope, set milestones, sequence dependencies, and keep delivery on track across multiple engineering teamsEngage directly with system designs, architecture decisions, and infrastructure trade-offs — well enough to identify risks, ask the right questions, and flag issues before they compoundManage cross-functional dependencies across engineering, product, design, data, and infrastructure; keep teams aligned without creating process overhead that slows them downCommunicate program status, risks, and decisions to different audiences — architectural context for engineering leads, delivery forecasts for product, business impact for executives — adapting the level of detail to what each stakeholder actually needsIdentify, track, and escalate technical risks early; own the mitigation plan and follow through to resolution rather than handing it offRun the operational rhythm of the program: standups, steering meetings, decision logs, and retrospectives — with a bias toward async where possibleWrite and maintain clear program documentation: specs, decision records, dependency maps, and status reports that teams actually read and referenceDrive post-launch reviews and feed learnings back into how the next program runsBasic Qualifications4+ years of experience in a technical role — software engineering, infrastructure, or a closely related discipline — before or alongside program management workDemonstrable experience running large, cross-functional technical programs from planning through deliveryComfortable reading system design documents, architecture diagrams, and API specs; able to engage meaningfully with engineers on technical trade-offs without needing to write the codeStrong written and verbal communication skills, with the ability to adjust detail and framing depending on the audienceExperience managing dependencies and coordinating delivery across multiple teams simultaneouslyFamiliarity with project and program management tools (Jira, Confluence, Linear, or equivalent)Solid understanding of modern software delivery practices: CI/CD pipelines, agile methodologies, and release managementPreferred QualificationsBackground in cloud infrastructure, distributed systems, or platform engineering — giving you the vocabulary and intuition to engage on architecture without needing a technical translatorExperience running programs in high-ambiguity environments where the scope evolves and the requirements are still being worked outFamiliarity with risk frameworks and structured approaches to dependency mapping across large engineering organizationsExperience coordinating security, compliance, or data privacy requirements as part of a technical programTrack record of improving how programs run — not just delivering on time, but leaving better processes behindSalary Range$170,000 - $240,000 USD + BenefitsWhy AI FabrikAt AI Fabrik, we hire for impact. We want those who challenge how inference infrastructure is built and who excel at delivering it in production. We are builders, architects, engineers, and researchers. We move fast, work with rigor, and care deeply about what runs in the real world.We are committed to building a diverse and inclusive team. AI Fabrik is an equal opportunity employer. We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.Please note that this is an onsite position based out of AI Fabrik's Redwood City, California office.