{"schemaVersion":"jobsearcher.job.v1","id":"de496360779084e8d3205d6b","url":"https://jobsearcher.com/jobs/de496360779084e8d3205d6b","canonicalUrl":"https://jobsearcher.com/jobs/de496360779084e8d3205d6b","title":"Staff Data Engineer","description":"About the Company\r\nGemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of simple, reliable, and secure crypto products and services to individuals and institutions in over 70 countries. Our mission is to unlock the next era of financial, creative, and personal freedom by providing trusted access to the decentralized future. We envision a world where crypto reshapes the global financial system, internet, and money to create greater choice, independence, and opportunity for all — bridging traditional finance with the emerging cryptoeconomy in a way that is more open, fair, and secure. As a publicly traded company, Gemini is poised to accelerate this vision with greater scale, reach, and impact.\r\nThe Department: Data\r\nAt Gemini, our Data Team is the engine that powers insight, innovation, and trust across the company. We bring together world-class data engineers, platform engineers, machine-learning engineers, analytics engineers, and data scientists — all working in harmony to transform raw information into secure, reliable, and actionable intelligence. From building scalable pipelines and platforms, to enabling cutting-edge machine learning, to ensuring governance and cost efficiency, we deliver the foundation for smarter decisions and breakthrough products. We thrive at the intersection of crypto, technology, and finance, and we're united by a shared mission: to unlock the full potential of Gemini's data to drive growth, efficiency, and customer impact.\r\nThe Role: Staff Data Engineer\r\nThe Data team is responsible for designing and operating the data infrastructure that powers insight, reporting, analytics, and machine learning across the business. As a Staff Data Engineer, you will lead architectural initiatives, mentor others, and build high-scale systems that impact the entire organization. You will partner closely with product, analytics, ML, finance, operations, and engineering teams to move, transform, and model data reliably, with observability, resilience, and agility.\r\nThis role is required to be in person twice a week at either our San Francisco, CA or New York City, NY office.\r\nResponsibilities\r\nLead the architecture, design, and implementation of data infrastructure and pipelines, spanning both batch and real-time / streaming workloads\r\nBuild and maintain scalable, efficient, and reliable ETL/ELT pipelines using languages and frameworks such as Python, SQL, Spark, Flink, Beam, or equivalents\r\nWork on real-time or near-real-time data solutions (e.g. CDC, streaming, micro-batch) for use cases that require timely data\r\nPartner with data scientists, ML engineers, analysts, and product teams to understand data requirements, define SLAs, and deliver coherent data products that others can self-serve\r\nEstablish data quality, validation, observability, and monitoring frameworks (data auditing, alerting, anomaly detection, data lineage)\r\nInvestigate and resolve complex production issues: root cause analysis, performance bottlenecks, data integrity, fault tolerance\r\nMentor and guide more junior and mid-level data engineers: lead code reviews, design reviews, and best-practice evangelism\r\nStay up to date on new tools, technologies, and patterns in the data and cloud space, bringing proposals and proof-of-concepts when appropriate\r\nDocument data flows, data dictionaries, architecture patterns, and operational runbooks\r\nMinimum Qualifications\r\n8+ years of experience in data engineering (or similar) roles\r\nStrong experience in ETL/ELT pipeline design, implementation, and optimization\r\nDeep expertise in Python and SQL writing production-quality, maintainable, testable code\r\nExperience with large-scale data warehouses (e.g. Databricks, BigQuery, Snowflake)\r\nSolid grounding in software engineering fundamentals, data structures, and systems thinking\r\nHands-on experience in data modeling (dimensional modeling, normalization, schema design)\r\nExperience building systems with real-time or streaming data (e.g. Kafka, Kinesis, Flink, Spark Streaming), and familiarity with CDC frameworks\r\nExperience with orchestration / workflow frameworks (e.g. Airflow)\r\nFamiliarity with data governance, lineage, metadata, cataloging, and data quality practices\r\nStrong cross-functional communication skills; ability to translate between technical and non-technical stakeholders\r\nProven experience in mentoring, leading design discussions, and influencing data-engineering best practices across teams\r\nPreferred Qualifications\r\nExperience with crypto, financial services, trading, markets, or exchange systems\r\nExperience with blockchain, crypto, Web3 data — e.g. blocks, transactions, contract calls, token transfers, UTXO/account models, on-chain indexing, chain APIs, etc.\r\nExperience with infrastructure as code, containerization, and CI/CD pipelines\r\nHands-on experience managing and optimizing Databricks on AWS\r\nIt Pays to Work Here\r\nThe compensation & benefits package for this role includes:\r\nCompetitive starting salary\r\nA discretionary annual bonus\r\nLong-term incentive in the form of a new hire equity grant\r\nComprehensive health plans\r\n401K with company matching\r\nPaid Parental Leave\r\nFlexible time off\r\nSalary Range\r\nThe base salary range for this role is between $168,000 - $240,000 in the State of New York, the State of California and the State of Washington. This range is not inclusive of our discretionary bonus or equity package. When determining a candidate's compensation, we consider a number of factors including skillset, experience, job scope, and current market data.\r\nIn the United States, we offer a hybrid work approach at our hub offices, balancing the benefits of in-person collaboration with the flexibility of remote work. Expectations may vary by location and role, so candidates are encouraged to connect with their recruiter to learn more about the specific policy for the role. Employees who do not live near one of our hubs are part of our remote workforce.\r\nAt Gemini, we strive to build diverse teams that reflect the people we want to empower through our products, and we are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, or Veteran status. Equal Opportunity is the Law, and Gemini is proud to be an equal opportunity workplace. If you have a specific need that requires accommodation, please let a member of the People Team know.\r\nJ-18808-Ljbffr","company":"Gemini","rawCompany":"gemini","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-04-11T05:04:53.363Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff Data Engineer","description":"About the Company\r\nGemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of simple, reliable, and secure crypto products and services to individuals and institutions in over 70 countries. Our mission is to unlock the next era of financial, creative, and personal freedom by providing trusted access to the decentralized future. We envision a world where crypto reshapes the global financial system, internet, and money to create greater choice, independence, and opportunity for all — bridging traditional finance with the emerging cryptoeconomy in a way that is more open, fair, and secure. As a publicly traded company, Gemini is poised to accelerate this vision with greater scale, reach, and impact.\r\nThe Department: Data\r\nAt Gemini, our Data Team is the engine that powers insight, innovation, and trust across the company. We bring together world-class data engineers, platform engineers, machine-learning engineers, analytics engineers, and data scientists — all working in harmony to transform raw information into secure, reliable, and actionable intelligence. From building scalable pipelines and platforms, to enabling cutting-edge machine learning, to ensuring governance and cost efficiency, we deliver the foundation for smarter decisions and breakthrough products. We thrive at the intersection of crypto, technology, and finance, and we're united by a shared mission: to unlock the full potential of Gemini's data to drive growth, efficiency, and customer impact.\r\nThe Role: Staff Data Engineer\r\nThe Data team is responsible for designing and operating the data infrastructure that powers insight, reporting, analytics, and machine learning across the business. As a Staff Data Engineer, you will lead architectural initiatives, mentor others, and build high-scale systems that impact the entire organization. You will partner closely with product, analytics, ML, finance, operations, and engineering teams to move, transform, and model data reliably, with observability, resilience, and agility.\r\nThis role is required to be in person twice a week at either our San Francisco, CA or New York City, NY office.\r\nResponsibilities\r\nLead the architecture, design, and implementation of data infrastructure and pipelines, spanning both batch and real-time / streaming workloads\r\nBuild and maintain scalable, efficient, and reliable ETL/ELT pipelines using languages and frameworks such as Python, SQL, Spark, Flink, Beam, or equivalents\r\nWork on real-time or near-real-time data solutions (e.g. CDC, streaming, micro-batch) for use cases that require timely data\r\nPartner with data scientists, ML engineers, analysts, and product teams to understand data requirements, define SLAs, and deliver coherent data products that others can self-serve\r\nEstablish data quality, validation, observability, and monitoring frameworks (data auditing, alerting, anomaly detection, data lineage)\r\nInvestigate and resolve complex production issues: root cause analysis, performance bottlenecks, data integrity, fault tolerance\r\nMentor and guide more junior and mid-level data engineers: lead code reviews, design reviews, and best-practice evangelism\r\nStay up to date on new tools, technologies, and patterns in the data and cloud space, bringing proposals and proof-of-concepts when appropriate\r\nDocument data flows, data dictionaries, architecture patterns, and operational runbooks\r\nMinimum Qualifications\r\n8+ years of experience in data engineering (or similar) roles\r\nStrong experience in ETL/ELT pipeline design, implementation, and optimization\r\nDeep expertise in Python and SQL writing production-quality, maintainable, testable code\r\nExperience with large-scale data warehouses (e.g. Databricks, BigQuery, Snowflake)\r\nSolid grounding in software engineering fundamentals, data structures, and systems thinking\r\nHands-on experience in data modeling (dimensional modeling, normalization, schema design)\r\nExperience building systems with real-time or streaming data (e.g. Kafka, Kinesis, Flink, Spark Streaming), and familiarity with CDC frameworks\r\nExperience with orchestration / workflow frameworks (e.g. Airflow)\r\nFamiliarity with data governance, lineage, metadata, cataloging, and data quality practices\r\nStrong cross-functional communication skills; ability to translate between technical and non-technical stakeholders\r\nProven experience in mentoring, leading design discussions, and influencing data-engineering best practices across teams\r\nPreferred Qualifications\r\nExperience with crypto, financial services, trading, markets, or exchange systems\r\nExperience with blockchain, crypto, Web3 data — e.g. blocks, transactions, contract calls, token transfers, UTXO/account models, on-chain indexing, chain APIs, etc.\r\nExperience with infrastructure as code, containerization, and CI/CD pipelines\r\nHands-on experience managing and optimizing Databricks on AWS\r\nIt Pays to Work Here\r\nThe compensation & benefits package for this role includes:\r\nCompetitive starting salary\r\nA discretionary annual bonus\r\nLong-term incentive in the form of a new hire equity grant\r\nComprehensive health plans\r\n401K with company matching\r\nPaid Parental Leave\r\nFlexible time off\r\nSalary Range\r\nThe base salary range for this role is between $168,000 - $240,000 in the State of New York, the State of California and the State of Washington. This range is not inclusive of our discretionary bonus or equity package. When determining a candidate's compensation, we consider a number of factors including skillset, experience, job scope, and current market data.\r\nIn the United States, we offer a hybrid work approach at our hub offices, balancing the benefits of in-person collaboration with the flexibility of remote work. Expectations may vary by location and role, so candidates are encouraged to connect with their recruiter to learn more about the specific policy for the role. Employees who do not live near one of our hubs are part of our remote workforce.\r\nAt Gemini, we strive to build diverse teams that reflect the people we want to empower through our products, and we are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, or Veteran status. Equal Opportunity is the Law, and Gemini is proud to be an equal opportunity workplace. If you have a specific need that requires accommodation, please let a member of the People Team know.\r\nJ-18808-Ljbffr","datePosted":"2026-04-11T05:04:53.363Z","dateModified":"2026-04-11T05:04:53.363Z","hiringOrganization":{"@type":"Organization","name":"Gemini","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"de496360779084e8d3205d6b"},"url":"https://jobsearcher.com/jobs/de496360779084e8d3205d6b"}}