Staff Backline Engineer – ML/AI
Overview
In this role you will act as a bridge between frontline support and engineering, tackling high-priority issues in the Data & AI ecosystem. You’ll perform deep-dive analysis and root-cause investigations to stabilize large-scale production workloads, while collaborating with product and engineering teams to improve the platform. You’ll develop tooling and automation to streamline troubleshooting and scale the organization. This is a hands-on role with impact on reliability, performance, and customer satisfaction.
ResponsibilitiesDeep dive forensics into Spark core internals and the Databricks Data & AI ecosystem to resolve architectural failures and anomaliesPerform advanced code-level root-cause analysis and resource profiling to ensure stability of high-scale workloadsOptimize architectural performance by refining execution parameters and best-practice strategiesInfluence product roadmap through analysis of global issue trends and collaboration with Product EngineeringDevelop reproduction frameworks, automated workflows, and AI-driven diagnostic tools to standardize resolutions
Key requirements10+ years of relevant experienceSpecialization in one of three tracks (Data Engineering Track OR Product Supportability Track OR AI Track)Experience managing both customers and technical stakeholderscustomer-obsessedstrong problem-solving abilitiesmentoring and guiding othersSpark, Delta Lake, Hive (Data Engineering Track)Python, SQL, Scala (data engineering or code-level tasks)Java, Scala, Python (distributed systems and profiling)