{"schemaVersion":"jobsearcher.job.v1","id":"e600bb40eed228f1cd35c101","url":"https://jobsearcher.com/jobs/e600bb40eed228f1cd35c101","canonicalUrl":"https://jobsearcher.com/jobs/e600bb40eed228f1cd35c101","title":"Lead Data Engineer, Data Platform","description":"About CrewAI\nCrewAI is the leading framework and enterprise platform for building and orchestrating multi-agent AI systems, powering 300M+ agent executions per month across thousands of companies. As the product, platform, and customer base scale, data is becoming one of the most important systems in the company: how we understand usage, reliability, activation, customer health, cost, governance, and where to invest next.\nToday, we have meaningful data already, but it is spread across product telemetry, trace data, application databases, analytics tables, Cube models, Metabase dashboards, and team-specific queries. We need someone to turn that into a coherent, trusted, useful data foundation.\nThe Role\nYou’ll be CrewAI’s first dedicated data engineering hire. Your job is to own the data foundation end to end: rationalize what exists, improve the infrastructure, define trusted metrics, close instrumentation gaps, and make data accessible enough that product, growth, engineering, customer success, and leadership can actually use it.\nThis is a foundational role with real range. The center of gravity is data infrastructure and analytics engineering: pipelines, warehouse/lake design, semantic modeling, metric definitions, data quality, and self-serve access. You’ll also be the person who turns messy questions into clear analysis, reliable dashboards, and better product decisions.\nThis is not a maintenance role. It is a “make data legible and useful for the company” role.\nWhat You’ll Do\nOwn and evolve CrewAI’s data platform across ingestion, transformation, storage, semantic modeling, BI, and operational data quality.\nRationalize the existing data estate: product events, execution telemetry, OpenTelemetry-derived traces, application tables, Cube models, Redshift/data-lake tables, Metabase dashboards, and team-specific reporting.\nEstablish trusted source-of-truth metrics for the business and product, including executions, active builders/users, activation, deployment health, token and cost usage, customer health, governance adoption, retention, and feature usage.\nBuild and maintain the models, pipelines, and metric layers that make those numbers consistent across teams.\nPartner with product and engineering to improve instrumentation, event taxonomy, data contracts, and telemetry coverage for new features.\nMake data self-serve through clear dashboards, documented datasets, reusable metric definitions, and sensible access patterns.\nImprove reliability and trust in the stack through data quality checks, freshness monitoring, lineage, alerting, backfills, and incident/debug workflows.\nPartner with Discovery, product, and go-to-market teams on analysis behind recommendations, customer signals, usage patterns, and roadmap decisions.\nKeep the stack secure and cost-aware, including access control, PII handling, retention, and warehouse/query efficiency.\nHelp define how CrewAI uses data internally as the company scales.\n\nRequirements\n\nWhat We’re Looking For\nStrong data engineering or analytics engineering experience, especially building data foundations in fast-moving product companies.\nExcellent SQL and data modeling skills, with experience designing reliable datasets, fact/dimension models, and metric definitions.\nExperience operating a warehouse or analytics store such as Redshift, Snowflake, BigQuery, Postgres, or similar.\nFamiliarity with transformation and modeling tools such as dbt, Cube, semantic layers, or equivalent systems.\nExperience with event pipelines, product telemetry, application data, and BI tools such as Metabase, Looker, Mode, or similar.\nStrong Python for data work, automation, validation, and operational workflows.\nProduct sense: you can turn ambiguous questions into useful metrics, and you care whether the numbers are understood correctly.\nPragmatism: you are comfortable inheriting messy systems, improving them incrementally, and choosing boring reliable solutions when they are right.\nStrong communication and documentation habits. You make data easier for other people to use.\nComfort being the first dedicated owner in an early-stage, high-growth environment.\nBonus\nExperience with LLM, agent, observability, trace, usage, or cost analytics.\nExperience with OpenTelemetry, high-volume event data, or operational telemetry.\nExperience with experimentation, causal analysis, activation/retention modeling, or customer health scoring.\nExperience defining event taxonomies and instrumentation standards for SaaS products.\nFamiliarity with Rails/Postgres application data, background jobs, and product analytics in B2B SaaS.\nLightweight ML or recommendation experience, especially where it supports product or customer workflows.","company":"Crewai","rawCompany":"crewai","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-07-11T12:08:47.504Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Lead Data Engineer, Data Platform","description":"About CrewAI\nCrewAI is the leading framework and enterprise platform for building and orchestrating multi-agent AI systems, powering 300M+ agent executions per month across thousands of companies. As the product, platform, and customer base scale, data is becoming one of the most important systems in the company: how we understand usage, reliability, activation, customer health, cost, governance, and where to invest next.\nToday, we have meaningful data already, but it is spread across product telemetry, trace data, application databases, analytics tables, Cube models, Metabase dashboards, and team-specific queries. We need someone to turn that into a coherent, trusted, useful data foundation.\nThe Role\nYou’ll be CrewAI’s first dedicated data engineering hire. Your job is to own the data foundation end to end: rationalize what exists, improve the infrastructure, define trusted metrics, close instrumentation gaps, and make data accessible enough that product, growth, engineering, customer success, and leadership can actually use it.\nThis is a foundational role with real range. The center of gravity is data infrastructure and analytics engineering: pipelines, warehouse/lake design, semantic modeling, metric definitions, data quality, and self-serve access. You’ll also be the person who turns messy questions into clear analysis, reliable dashboards, and better product decisions.\nThis is not a maintenance role. It is a “make data legible and useful for the company” role.\nWhat You’ll Do\nOwn and evolve CrewAI’s data platform across ingestion, transformation, storage, semantic modeling, BI, and operational data quality.\nRationalize the existing data estate: product events, execution telemetry, OpenTelemetry-derived traces, application tables, Cube models, Redshift/data-lake tables, Metabase dashboards, and team-specific reporting.\nEstablish trusted source-of-truth metrics for the business and product, including executions, active builders/users, activation, deployment health, token and cost usage, customer health, governance adoption, retention, and feature usage.\nBuild and maintain the models, pipelines, and metric layers that make those numbers consistent across teams.\nPartner with product and engineering to improve instrumentation, event taxonomy, data contracts, and telemetry coverage for new features.\nMake data self-serve through clear dashboards, documented datasets, reusable metric definitions, and sensible access patterns.\nImprove reliability and trust in the stack through data quality checks, freshness monitoring, lineage, alerting, backfills, and incident/debug workflows.\nPartner with Discovery, product, and go-to-market teams on analysis behind recommendations, customer signals, usage patterns, and roadmap decisions.\nKeep the stack secure and cost-aware, including access control, PII handling, retention, and warehouse/query efficiency.\nHelp define how CrewAI uses data internally as the company scales.\n\nRequirements\n\nWhat We’re Looking For\nStrong data engineering or analytics engineering experience, especially building data foundations in fast-moving product companies.\nExcellent SQL and data modeling skills, with experience designing reliable datasets, fact/dimension models, and metric definitions.\nExperience operating a warehouse or analytics store such as Redshift, Snowflake, BigQuery, Postgres, or similar.\nFamiliarity with transformation and modeling tools such as dbt, Cube, semantic layers, or equivalent systems.\nExperience with event pipelines, product telemetry, application data, and BI tools such as Metabase, Looker, Mode, or similar.\nStrong Python for data work, automation, validation, and operational workflows.\nProduct sense: you can turn ambiguous questions into useful metrics, and you care whether the numbers are understood correctly.\nPragmatism: you are comfortable inheriting messy systems, improving them incrementally, and choosing boring reliable solutions when they are right.\nStrong communication and documentation habits. You make data easier for other people to use.\nComfort being the first dedicated owner in an early-stage, high-growth environment.\nBonus\nExperience with LLM, agent, observability, trace, usage, or cost analytics.\nExperience with OpenTelemetry, high-volume event data, or operational telemetry.\nExperience with experimentation, causal analysis, activation/retention modeling, or customer health scoring.\nExperience defining event taxonomies and instrumentation standards for SaaS products.\nFamiliarity with Rails/Postgres application data, background jobs, and product analytics in B2B SaaS.\nLightweight ML or recommendation experience, especially where it supports product or customer workflows.","datePosted":"2026-07-11T12:08:47.504Z","dateModified":"2026-07-11T12:08:47.504Z","hiringOrganization":{"@type":"Organization","name":"Crewai","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"e600bb40eed228f1cd35c101"},"url":"https://jobsearcher.com/jobs/e600bb40eed228f1cd35c101"}}