JOBSEARCHER

Software Engineer (lead)

ARCHIVED

We can't find an active application page for this role right now. It may reopen or be listed elsewhere. Use Next Steps to search for an active apply link and similar live jobs.

Location: San Francisco Bay Area (Remote OK with up to 30% travel, primarily to the Bay Area)Compensation: $180,000-$250,000 base + generous equity + benefitsExperience: 7+ yearsAbout DatastratoDatastrato, the original creator of Apache Gravitino, is building the open metadata platform for the AI era, increasingly critical as enterprises endeavor to deploy AI and AI agents at scale. Led by prominent open source leaders with deep expertise in data infra/large-scale distributed systems and Silicon Valley veterans, in just over 2 years, we have built a compelling product, cultivated a vibrant open source community (Apple, Pinterest, Roku, Tencent, Uber, etc), and signed enterprises including 2 of the top 20 US Internet companies as paying customers.We're defining the next generation of data infrastructure for AI and are well-positioned to emerge as a category leader.Core ResponsibilitiesLead architectural design and implementation of major components in Apache Gravitino and Datastrato's commercial offeringsBuild and optimize high-performance, scalable systems across metadata management, storage, and distributed compute engines, and clouds; Contribute to and help shape relevant open source projectsDrive technical direction, mentor engineers, and influence cross-team architecture decisionsEngage with customers (typically top tech companies), engineering leaders, and the broader data and AI community, including speaking at industry eventsIdeal Candidate Profile7+ years in building large-scale distributed systems, including experience as a tech leadStrong foundation in CS, systems design and programming, data and AI infrastructure technologies, especially OSS such as Iceberg, Spark, Gravitino, Trino, vLLM, DaftProficiency in at least one systems language (Java, Go, C++, Rust) and solid Linux/OS fundamentalsTrack record of meaningful open-source contributionsStrong PlusExperience at early-stage data or AI infrastructure, or developer-first startups and top tech companiesExperience with database/data engine internals (query execution, transactions, storage), lakehouse technologies (Iceberg, Hudi) and metadata/governance systemsFamiliarity with modern AI infrastructure, LLMs, or AI agent systemsTechnical writing, developer evangelism, or community engagement experience#J-18808-Ljbffr