{"schemaVersion":"jobsearcher.job.v1","id":"42c9982838aed3e759240314","url":"https://jobsearcher.com/jobs/42c9982838aed3e759240314","canonicalUrl":"https://jobsearcher.com/jobs/42c9982838aed3e759240314","title":"Staff Software Engineer, ML Platform","description":"About Stack:\nStack is developing revolutionary AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous technology incorporates cutting-edge advancements in artificial intelligence, robotics, machine learning, and cloud technologies, empowering us to create innovative solutions that address the needs and challenges of the dynamic trucking transportation industry. With decades of experience creating and deploying real world systems for demanding environments, the Stack team is dedicated to developing an autonomous solution ecosystem tailored to the trucking industry's unique demands.\nAbout the Role:\nIn the ML Data Understanding team, our mission is to provide trusted and useful data to efficiently power all of Stack's ML applications end-to-end from mining to training to safety evaluation. We work hand in hand with AV autonomy teams to provide cutting edge solutions to all their data needs, working across data engineering, mining, modeling and infrastructure. In particular, we provide services to find (data mining), curate (datasets), annotate (data labeling), search and serve (high throughput data access) data for all ML needs.\nData Mining: We are building a framework and infrastructure to find interesting events quickly and flexibly. As part of this mission, you would be setting the direction for and helping us build an inference service using LLMs, open-world models and vector databases.\nSemantic Search for Data Mining: We are building the infrastructure of a highly scalable semantic search service for multimodal data to find interesting events quickly and flexibly. As part of this mission, you would be setting the direction for and helping us build an inference service using the latest AI models & approaches.\nDataset management for training: We are building state of the art infrastructure to support machine learning training and inference workloads using OSS components such as Ray, Spark, Lance and Iceberg.\nResponsibilities:\nBuild state-of-art multimodal data mining and semantic search solutions to power AV product development.\nDevelop data understanding platform infrastructure for real-time querying/vector databases and batch/stream processing using technologies like Ray, Spark, Lance, or similar.\nDeliver end-to-end data mining solutions that span onboard (C++) and offboard (ML & Data Infra) infrastructure to accelerate AV product development.\nDevelop e2e solution for real-time semantic search services (text/images/videos) and vector DBs.\nDiscover and identify key issues in existing ML infra and proactively improve system performance.\nBuild low latency/high throughput batch or stream processing pipelines.\nDrive technical discussions across multiple orgs and deliver solutions on a timely basis.\nArchitect and tune ETL pipelines to maximize GPU/CPU/Ram utilization.\nWrite readable and high-performance Python/C++ code.\nQualifications:\nExperience with both ML platforms and building ML-based applications (modeling experience is a bonus).\nProven track record of building scalable, reliable infrastructure in a fast-paced environment.\nAbility to collaborate effectively across teams.\nExperience building or using ML infrastructure for a large number of customer teams.\nDeep understanding of design trade-offs with the ability to articulate those trade-offs and achieve alignment with others.\nExperience in building ML models or infrastructure in domains such as autonomous vehicles, perception, and decision-making (desirable but not required).\nExperience with model training, model optimization, or large data processing pipelines.\nPrior experience in autonomous vehicles (AV) is a plus.\n6+ years of experience with:\nMultimodal data indexing and inference pipelines.\nBuilding semantic search service, embedding generation for video/images and vector DB.\nLarge scale ML pipelines (Airflow/Flyte) and model optimization.\nWe are proud to be an equal opportunity workplace. We believe that diverse teams produce the best ideas and outcomes. We are committed to building a culture of inclusion, entrepreneurship, and innovation across gender, race, age, sexual orientation, religion, disability, and identity.\nCheck out our Privacy Policy.\nPlease Note: Pursuant to its business activities and use of technology, Stack AV complies with all applicable U.S. national security laws, regulations, and administrative requirements, which can restrict Stack AV's ability to employ certain persons in certain positions pursuant to a range of national security-related requirements. As such, this position may be contingent upon Stack AV verifying a candidate's residence, U.S. person status, and/or citizenship status. This position may also involve working with software and technologies subject to U.S. export control regulations. Under these regulations, it may be necessary for Stack AV to obtain a U.S. government export license prior to releasing its technologies to certain persons. If Stack AV determines that a candidate's residence, U.S. person status, and/or citizenship status will require a license, prohibit the candidate from working in this position, or otherwise be subject to national security-related restrictions, Stack AV expressly reserves the right to either consider the candidate for a different position that is not subject to such restrictions, on whatever terms and conditions Stack AV shall establish in its sole discretion, or, in the alternative, decline to move forward with the candidate's application.","company":"Stackav","rawCompany":"stackav","city":"Pittsburgh","state":"PA","isRemote":false,"isActive":false,"createdAt":"2026-08-05T11:55:32.882Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff Software Engineer, ML Platform","description":"About Stack:\nStack is developing revolutionary AI and advanced autonomous systems designed to enhance safety, reliability, and efficiency of modern operations. Stack's autonomous technology incorporates cutting-edge advancements in artificial intelligence, robotics, machine learning, and cloud technologies, empowering us to create innovative solutions that address the needs and challenges of the dynamic trucking transportation industry. With decades of experience creating and deploying real world systems for demanding environments, the Stack team is dedicated to developing an autonomous solution ecosystem tailored to the trucking industry's unique demands.\nAbout the Role:\nIn the ML Data Understanding team, our mission is to provide trusted and useful data to efficiently power all of Stack's ML applications end-to-end from mining to training to safety evaluation. We work hand in hand with AV autonomy teams to provide cutting edge solutions to all their data needs, working across data engineering, mining, modeling and infrastructure. In particular, we provide services to find (data mining), curate (datasets), annotate (data labeling), search and serve (high throughput data access) data for all ML needs.\nData Mining: We are building a framework and infrastructure to find interesting events quickly and flexibly. As part of this mission, you would be setting the direction for and helping us build an inference service using LLMs, open-world models and vector databases.\nSemantic Search for Data Mining: We are building the infrastructure of a highly scalable semantic search service for multimodal data to find interesting events quickly and flexibly. As part of this mission, you would be setting the direction for and helping us build an inference service using the latest AI models & approaches.\nDataset management for training: We are building state of the art infrastructure to support machine learning training and inference workloads using OSS components such as Ray, Spark, Lance and Iceberg.\nResponsibilities:\nBuild state-of-art multimodal data mining and semantic search solutions to power AV product development.\nDevelop data understanding platform infrastructure for real-time querying/vector databases and batch/stream processing using technologies like Ray, Spark, Lance, or similar.\nDeliver end-to-end data mining solutions that span onboard (C++) and offboard (ML & Data Infra) infrastructure to accelerate AV product development.\nDevelop e2e solution for real-time semantic search services (text/images/videos) and vector DBs.\nDiscover and identify key issues in existing ML infra and proactively improve system performance.\nBuild low latency/high throughput batch or stream processing pipelines.\nDrive technical discussions across multiple orgs and deliver solutions on a timely basis.\nArchitect and tune ETL pipelines to maximize GPU/CPU/Ram utilization.\nWrite readable and high-performance Python/C++ code.\nQualifications:\nExperience with both ML platforms and building ML-based applications (modeling experience is a bonus).\nProven track record of building scalable, reliable infrastructure in a fast-paced environment.\nAbility to collaborate effectively across teams.\nExperience building or using ML infrastructure for a large number of customer teams.\nDeep understanding of design trade-offs with the ability to articulate those trade-offs and achieve alignment with others.\nExperience in building ML models or infrastructure in domains such as autonomous vehicles, perception, and decision-making (desirable but not required).\nExperience with model training, model optimization, or large data processing pipelines.\nPrior experience in autonomous vehicles (AV) is a plus.\n6+ years of experience with:\nMultimodal data indexing and inference pipelines.\nBuilding semantic search service, embedding generation for video/images and vector DB.\nLarge scale ML pipelines (Airflow/Flyte) and model optimization.\nWe are proud to be an equal opportunity workplace. We believe that diverse teams produce the best ideas and outcomes. We are committed to building a culture of inclusion, entrepreneurship, and innovation across gender, race, age, sexual orientation, religion, disability, and identity.\nCheck out our Privacy Policy.\nPlease Note: Pursuant to its business activities and use of technology, Stack AV complies with all applicable U.S. national security laws, regulations, and administrative requirements, which can restrict Stack AV's ability to employ certain persons in certain positions pursuant to a range of national security-related requirements. As such, this position may be contingent upon Stack AV verifying a candidate's residence, U.S. person status, and/or citizenship status. This position may also involve working with software and technologies subject to U.S. export control regulations. Under these regulations, it may be necessary for Stack AV to obtain a U.S. government export license prior to releasing its technologies to certain persons. If Stack AV determines that a candidate's residence, U.S. person status, and/or citizenship status will require a license, prohibit the candidate from working in this position, or otherwise be subject to national security-related restrictions, Stack AV expressly reserves the right to either consider the candidate for a different position that is not subject to such restrictions, on whatever terms and conditions Stack AV shall establish in its sole discretion, or, in the alternative, decline to move forward with the candidate's application.","datePosted":"2026-08-05T11:55:32.882Z","dateModified":"2026-08-05T11:55:32.882Z","hiringOrganization":{"@type":"Organization","name":"Stackav","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Pittsburgh","addressRegion":"PA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"42c9982838aed3e759240314"},"url":"https://jobsearcher.com/jobs/42c9982838aed3e759240314"}}