{"schemaVersion":"jobsearcher.job.v1","id":"45d950a08e4cf339cd372ff7","url":"https://jobsearcher.com/jobs/45d950a08e4cf339cd372ff7","canonicalUrl":"https://jobsearcher.com/jobs/45d950a08e4cf339cd372ff7","title":"Senior Data Engineer - Data Lead","description":"About Us\nFoundry Robotics is building an AI-native robotics manufacturing company focused on deploying advanced assembly and production capability for leading robotics companies and national-security-critical hardware. Basically, we’re building robots that build robots.\nWe are reimagining manufacturing through advanced robotics. Our mission is to rebuild the American manufacturing industry as an AI-first, assembly-focused, dual-use contract manufacturer. We aim to empower manufacturers with intelligent, efficient, and adaptable robotic systems that redefine productivity and quality.\nThe Role\nWe are hiring a Sr Data Engineer / Data Lead to own the data layer that powers Factory OS. You will design and operate petabyte-scale data pipelines that ingest visual, telemetry, and manufacturing data from the factory floor, move it reliably across edge, on-prem, and cloud environments, and make it available for ML training, analytics, and operational decision-making. You will define the data roadmap—architecture, governance, lifecycle, and tooling—and be the person the rest of the engineering org depends on for clean, reliable, well-modeled data. You will also contribute to backend services where data meets application logic. This is a hands-on engineering role. You will ship.\nKey Responsibilities\nPetabyte-Scale Data Pipelines\nDesign, build, and operate PB-scale data pipelines for ingesting, curating, indexing, and preparing manufacturing, visual, and telemetry data\nImplement reliable data movement across embedded, edge, on-prem, and cloud compute environments\nBuild streaming and batch processing systems that handle high-throughput factory-floor data in real time\nImplement data lifecycle management: retention, archival, compaction, and cost optimization at scale\nData Architecture + Roadmap\nDefine and own the Factory OS data roadmap—architecture, governance, quality, and tooling strategy\nDesign data models, schemas, and contracts that serve application, ML, and analytics consumers\nEstablish data cataloging, lineage tracking, and discoverability across the platform\nDrive data quality standards: validation, monitoring, alerting, and anomaly detection\nEvaluate and adopt data technologies (warehouses, lakehouses, streaming platforms) as the platform scales\nBackend Services + Integration\nContribute to backend services where data pipelines meet application logic (APIs, event-driven systems, database layers)\nPartner with ML engineers to ensure training and inference pipelines have access to clean, well-prepared datasets\nCollaborate with full-stack developers, robotics teams, and operations to translate data needs into reliable infrastructure\nImplement observability: metrics, logging, and tracing across data systems\nWhat We're Looking For\nBetween 5-7 years of experience in data engineering, with demonstrated work at terabyte-to-petabyte scale\nDeep hands-on experience with data pipeline frameworks (Spark, Flink, Kafka, Airflow, dbt, or similar)\nStrong proficiency in Python and SQL; backend experience in Go, Java, or TypeScript is a plus\nExperience designing data architectures spanning streaming, batch, lakehouse, and warehouse patterns\nSolid understanding of distributed systems, storage engines, and data modeling (relational and NoSQL)\nExperience with cloud data services (AWS S3/Glue/Redshift, GCP BigQuery/Dataflow, or Azure equivalents)\nComfortable defining roadmaps, making architectural decisions, and driving alignment across teams\nComfortable operating independently and owning complex systems end-to-end\nNice to Have (Not Required)\nExperience with visual or image data pipelines at scale (video ingestion, frame extraction, annotation workflows)\nBackground in manufacturing, robotics, or industrial IoT data systems\nFamiliarity with ML data preparation workflows (feature stores, dataset versioning, labeling pipelines)\nExperience with edge-to-cloud data architectures and intermittent connectivity patterns\nExposure to infrastructure-as-code (Terraform, Pulumi) and container orchestration (Kubernetes)\nGrowth Opportunities\nDefine the entire data architecture and roadmap for an AI-native factory\nWork with petabyte-scale visual, telemetry, and manufacturing data in a real production environment\nGrow into a data leadership role as the team and platform scale\nDirect influence on how data drives quality, throughput, and cost improvements in physical manufacturing\nWhy Join Us?\nThis is one of the only places where world-class manufacturing operators, mechanical engineers, robotics researchers, and software engineers sit in the same room — building production systems together.\nWe are committed to being deeply embedded in the U.S. industrial base. Our focus is simple: build adaptive robotic assembly systems that make American manufacturing scalable, resilient, and competitive again.\nIf you want to run a mature, well-defined commercial org, this may not be the role.\nIf you want to build the commercial engine that brings AI-driven manufacturing to every industrial and energy customer in America — this is it.\nThe base salary range for this full-time position in the location of San Francisco is:\n$150,000—$250,000 USD\nCompensation packages at Foundry Robotics for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position, determined by work location and additional factors, including job-related skills, experience, interview performance, and relevant education or training. Foundry Robotics employees in eligible roles are also granted equity based compensation, subject to Board of Director approval. You'll also receive benefits including, but not limited to: Comprehensive health, dental and vision coverage, and generous PTO.\nCompensation Range: $150K - $250K","company":"Foundry Robotics","rawCompany":"foundry robotics","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-03T16:25:57.097Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Senior Data Engineer - Data Lead","description":"About Us\nFoundry Robotics is building an AI-native robotics manufacturing company focused on deploying advanced assembly and production capability for leading robotics companies and national-security-critical hardware. Basically, we’re building robots that build robots.\nWe are reimagining manufacturing through advanced robotics. Our mission is to rebuild the American manufacturing industry as an AI-first, assembly-focused, dual-use contract manufacturer. We aim to empower manufacturers with intelligent, efficient, and adaptable robotic systems that redefine productivity and quality.\nThe Role\nWe are hiring a Sr Data Engineer / Data Lead to own the data layer that powers Factory OS. You will design and operate petabyte-scale data pipelines that ingest visual, telemetry, and manufacturing data from the factory floor, move it reliably across edge, on-prem, and cloud environments, and make it available for ML training, analytics, and operational decision-making. You will define the data roadmap—architecture, governance, lifecycle, and tooling—and be the person the rest of the engineering org depends on for clean, reliable, well-modeled data. You will also contribute to backend services where data meets application logic. This is a hands-on engineering role. You will ship.\nKey Responsibilities\nPetabyte-Scale Data Pipelines\nDesign, build, and operate PB-scale data pipelines for ingesting, curating, indexing, and preparing manufacturing, visual, and telemetry data\nImplement reliable data movement across embedded, edge, on-prem, and cloud compute environments\nBuild streaming and batch processing systems that handle high-throughput factory-floor data in real time\nImplement data lifecycle management: retention, archival, compaction, and cost optimization at scale\nData Architecture + Roadmap\nDefine and own the Factory OS data roadmap—architecture, governance, quality, and tooling strategy\nDesign data models, schemas, and contracts that serve application, ML, and analytics consumers\nEstablish data cataloging, lineage tracking, and discoverability across the platform\nDrive data quality standards: validation, monitoring, alerting, and anomaly detection\nEvaluate and adopt data technologies (warehouses, lakehouses, streaming platforms) as the platform scales\nBackend Services + Integration\nContribute to backend services where data pipelines meet application logic (APIs, event-driven systems, database layers)\nPartner with ML engineers to ensure training and inference pipelines have access to clean, well-prepared datasets\nCollaborate with full-stack developers, robotics teams, and operations to translate data needs into reliable infrastructure\nImplement observability: metrics, logging, and tracing across data systems\nWhat We're Looking For\nBetween 5-7 years of experience in data engineering, with demonstrated work at terabyte-to-petabyte scale\nDeep hands-on experience with data pipeline frameworks (Spark, Flink, Kafka, Airflow, dbt, or similar)\nStrong proficiency in Python and SQL; backend experience in Go, Java, or TypeScript is a plus\nExperience designing data architectures spanning streaming, batch, lakehouse, and warehouse patterns\nSolid understanding of distributed systems, storage engines, and data modeling (relational and NoSQL)\nExperience with cloud data services (AWS S3/Glue/Redshift, GCP BigQuery/Dataflow, or Azure equivalents)\nComfortable defining roadmaps, making architectural decisions, and driving alignment across teams\nComfortable operating independently and owning complex systems end-to-end\nNice to Have (Not Required)\nExperience with visual or image data pipelines at scale (video ingestion, frame extraction, annotation workflows)\nBackground in manufacturing, robotics, or industrial IoT data systems\nFamiliarity with ML data preparation workflows (feature stores, dataset versioning, labeling pipelines)\nExperience with edge-to-cloud data architectures and intermittent connectivity patterns\nExposure to infrastructure-as-code (Terraform, Pulumi) and container orchestration (Kubernetes)\nGrowth Opportunities\nDefine the entire data architecture and roadmap for an AI-native factory\nWork with petabyte-scale visual, telemetry, and manufacturing data in a real production environment\nGrow into a data leadership role as the team and platform scale\nDirect influence on how data drives quality, throughput, and cost improvements in physical manufacturing\nWhy Join Us?\nThis is one of the only places where world-class manufacturing operators, mechanical engineers, robotics researchers, and software engineers sit in the same room — building production systems together.\nWe are committed to being deeply embedded in the U.S. industrial base. Our focus is simple: build adaptive robotic assembly systems that make American manufacturing scalable, resilient, and competitive again.\nIf you want to run a mature, well-defined commercial org, this may not be the role.\nIf you want to build the commercial engine that brings AI-driven manufacturing to every industrial and energy customer in America — this is it.\nThe base salary range for this full-time position in the location of San Francisco is:\n$150,000—$250,000 USD\nCompensation packages at Foundry Robotics for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position, determined by work location and additional factors, including job-related skills, experience, interview performance, and relevant education or training. Foundry Robotics employees in eligible roles are also granted equity based compensation, subject to Board of Director approval. You'll also receive benefits including, but not limited to: Comprehensive health, dental and vision coverage, and generous PTO.\nCompensation Range: $150K - $250K","datePosted":"2026-08-03T16:25:57.097Z","dateModified":"2026-08-03T16:25:57.097Z","hiringOrganization":{"@type":"Organization","name":"Foundry Robotics","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"45d950a08e4cf339cd372ff7"},"url":"https://jobsearcher.com/jobs/45d950a08e4cf339cd372ff7"}}