{"schemaVersion":"jobsearcher.job.v1","id":"c5e70eb241f29b1b9272ba6a","url":"https://jobsearcher.com/jobs/c5e70eb241f29b1b9272ba6a","canonicalUrl":"https://jobsearcher.com/jobs/c5e70eb241f29b1b9272ba6a","title":"Software Engineer, AI Distributed Systems","description":"Software Engineer — Distributed AI SystemsSan Francisco, CA | Five Days On-Site | $180K–$250K + EquityAI does not change the world when it works in a demo.It changes the world when it can operate continuously, reliably, and accurately inside the complex environments where consequential work actually happens.We’re partnering with a well-funded AI startup building the infrastructure required to make that possible. Its technology is already operating in demanding enterprise environments, processing enormous volumes of information and turning AI-generated outputs into reliable systems people can use to make better decisions and drive meaningful change.The company has fewer than 20 employees.That means the systems you build, the standards you establish, and the decisions you make will materially influence the product and the company.The OpportunityAs a Software Engineer focused on Distributed AI Systems, you’ll build and own the data, orchestration, and infrastructure layer supporting a continuously running AI platform.You’ll work on genuinely difficult engineering problems involving large-scale data processing, distributed compute, AI-driven transformations, and production reliability.This is not an analytics, dashboarding, or warehouse-modeling role. It is not a narrow infrastructure seat where you maintain systems designed by someone else.You’ll own company-critical technology from architecture through production—and help determine what the business is capable of building next.What You’ll OwnBuild and evolve large-scale, AI-driven data and transformation pipelinesTurn probabilistic AI outputs into reliable, structured, usable dataDesign resilient systems for retries, backfills, partial failures, and recoveryImprove platform accuracy, reliability, throughput, latency, and costBuild observability, evaluation, and debugging tools for distributed workloadsDevelop the internal platform capabilities used by other engineersMake architectural decisions with direct product and business impactParticipate in a separately compensated on-call rotationWhat We’re Looking For3+ years of experience in data platform, backend, infrastructure, or distributed-systems engineeringStrong production Python experienceEnd-to-end ownership of production data platforms or ETL pipelinesExperience operating systems at terabyte-to-petabyte scaleStrong understanding of distributed-system reliability and failure modesHands-on experience with AWS and infrastructure as codeAbility to explain not only what you built, but why you built it that wayComfort solving problems without an existing blueprintWillingness to work in the San Francisco office five days per weekEspecially RelevantAI-enabled ETL or production LLM systemsDAG and workflow-orchestration platformsDistributed compute, queues, sharding, partitioning, or streamingSpark, Ray, Flink, Kafka, Airflow, Dagster, Prefect, or TemporalPostgreSQL, vector databases, or high-scale data systemsInfrastructure ownership through a major company-growth inflectionQuantitative, research, or other unusually strong problem-solving experienceWhy This Role MattersYou’ll build the foundation. The company’s AI capabilities depend on the systems you own.You’ll solve real production problems. This technology is already operating in demanding customer environments.You’ll work at meaningful scale. These are continuously running distributed workloads—not occasional experiments.You’ll have a seat at the table. With fewer than 20 people, one exceptional engineer can materially influence the architecture, product, and engineering culture.You’ll build without a playbook. The hardest problems here do not have obvious answers or established solutions.You’ll see your impact. Your work will directly determine how reliably the product performs and how much the company can accomplish.Compensation and Logistics$180K–$250K base salaryCompetitive equityComprehensive benefitsAdditional compensation for on-call participationVisa transfers considered for qualified candidatesFive days per week in the San Francisco officeThe world has plenty of AI products that look impressive when everything goes right.This team is building the systems that keep AI accurate, reliable, and useful when the data is enormous, the environment is complicated, and the outcome actually matters.If that is the kind of problem you want to own, apply today because we need to talk!","company":"Epic Placements","rawCompany":"epic placements","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-03T10:05:06.454Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Software Engineer, AI Distributed Systems","description":"Software Engineer — Distributed AI SystemsSan Francisco, CA | Five Days On-Site | $180K–$250K + EquityAI does not change the world when it works in a demo.It changes the world when it can operate continuously, reliably, and accurately inside the complex environments where consequential work actually happens.We’re partnering with a well-funded AI startup building the infrastructure required to make that possible. Its technology is already operating in demanding enterprise environments, processing enormous volumes of information and turning AI-generated outputs into reliable systems people can use to make better decisions and drive meaningful change.The company has fewer than 20 employees.That means the systems you build, the standards you establish, and the decisions you make will materially influence the product and the company.The OpportunityAs a Software Engineer focused on Distributed AI Systems, you’ll build and own the data, orchestration, and infrastructure layer supporting a continuously running AI platform.You’ll work on genuinely difficult engineering problems involving large-scale data processing, distributed compute, AI-driven transformations, and production reliability.This is not an analytics, dashboarding, or warehouse-modeling role. It is not a narrow infrastructure seat where you maintain systems designed by someone else.You’ll own company-critical technology from architecture through production—and help determine what the business is capable of building next.What You’ll OwnBuild and evolve large-scale, AI-driven data and transformation pipelinesTurn probabilistic AI outputs into reliable, structured, usable dataDesign resilient systems for retries, backfills, partial failures, and recoveryImprove platform accuracy, reliability, throughput, latency, and costBuild observability, evaluation, and debugging tools for distributed workloadsDevelop the internal platform capabilities used by other engineersMake architectural decisions with direct product and business impactParticipate in a separately compensated on-call rotationWhat We’re Looking For3+ years of experience in data platform, backend, infrastructure, or distributed-systems engineeringStrong production Python experienceEnd-to-end ownership of production data platforms or ETL pipelinesExperience operating systems at terabyte-to-petabyte scaleStrong understanding of distributed-system reliability and failure modesHands-on experience with AWS and infrastructure as codeAbility to explain not only what you built, but why you built it that wayComfort solving problems without an existing blueprintWillingness to work in the San Francisco office five days per weekEspecially RelevantAI-enabled ETL or production LLM systemsDAG and workflow-orchestration platformsDistributed compute, queues, sharding, partitioning, or streamingSpark, Ray, Flink, Kafka, Airflow, Dagster, Prefect, or TemporalPostgreSQL, vector databases, or high-scale data systemsInfrastructure ownership through a major company-growth inflectionQuantitative, research, or other unusually strong problem-solving experienceWhy This Role MattersYou’ll build the foundation. The company’s AI capabilities depend on the systems you own.You’ll solve real production problems. This technology is already operating in demanding customer environments.You’ll work at meaningful scale. These are continuously running distributed workloads—not occasional experiments.You’ll have a seat at the table. With fewer than 20 people, one exceptional engineer can materially influence the architecture, product, and engineering culture.You’ll build without a playbook. The hardest problems here do not have obvious answers or established solutions.You’ll see your impact. Your work will directly determine how reliably the product performs and how much the company can accomplish.Compensation and Logistics$180K–$250K base salaryCompetitive equityComprehensive benefitsAdditional compensation for on-call participationVisa transfers considered for qualified candidatesFive days per week in the San Francisco officeThe world has plenty of AI products that look impressive when everything goes right.This team is building the systems that keep AI accurate, reliable, and useful when the data is enormous, the environment is complicated, and the outcome actually matters.If that is the kind of problem you want to own, apply today because we need to talk!","datePosted":"2026-09-03T10:05:06.454Z","dateModified":"2026-09-03T10:05:06.454Z","hiringOrganization":{"@type":"Organization","name":"Epic Placements","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"c5e70eb241f29b1b9272ba6a"},"url":"https://jobsearcher.com/jobs/c5e70eb241f29b1b9272ba6a"}}