{"schemaVersion":"jobsearcher.job.v1","id":"fe76e528e18de462bb5a72ff","url":"https://jobsearcher.com/jobs/fe76e528e18de462bb5a72ff","canonicalUrl":"https://jobsearcher.com/jobs/fe76e528e18de462bb5a72ff","title":"Senior Software Engineer - ML Infrastructure","description":"We're looking for a Senior Software Engineer - ML Infrastructure to build and scale the infrastructure that powers our AI-driven warehouse intelligence platform. You'll own the end-to-end lifecycle of computer vision models — from training pipelines through optimized cloud deployment — ensuring our cutting-edge computer vision and multi-modal AI systems run reliably and efficiently in production. Your work will directly enable the real-time perception and autonomous decision-making capabilities at the core of our platform.\nThis is a deeply technical role at the intersection of machine learning, distributed systems, and cloud infrastructure. You'll design scalable GPU compute clusters, build robust orchestration pipelines, and optimize model serving for low-latency inference at scale. You'll work closely with our research scientists, computer vision engineers, and product teams to bridge the gap between experimental models and production-ready systems that operate across diverse warehouse environments. We've found tremendous value in collaborative problem-solving, thus our team works from our SF office three days a week.\nResponsibilities\nDevelop and maintain distributed cloud GPU infrastructure for large-scale world model training and low-latency inference.\nBuild end-to-end computer vision pipelines — from data ingestion and preprocessing through model training, evaluation, and deployment — and integrate them into core product workflows.\nDeploy and optimize state-of-the-art machine learning models in the cloud using model serving platforms and inference optimization techniques, including VLMs and VLAs.\nDesign and operate orchestration systems that enable both engineers and non-engineers to build and manage data and ML pipelines.\nEstablish monitoring, benchmarking, and evaluation frameworks to ensure model performance and reliability in production environments.\nRequired Experience\nB.S. / M.S. in Computer Science, Robotics, or similar technical field, or equivalent practical experience.\n7+ years of professional software engineering experience, with at least 3 years in machine learning infrastructure — developing, scaling, training, deploying, and optimizing large-scale ML systems from data to model.\nTrack record of deploying machine learning models in production environments with real-world constraints.\nExperience with distributed messaging and compute systems (Kafka, gRPC, ROS2, or similar).\nStrong programming skills in Python with solid software engineering practices.\nPreferred Experience\nExperience with training and/or deployment of machine learning models in the computer vision domain.\nExperience developing, running, and managing orchestration systems (Flyte, Temporal, Airflow, or similar) for ML and data pipelines.\nProficiency with ML frameworks (PyTorch, TensorFlow, DeepSpeed) and model serving platforms (TorchServe, TensorFlow Serving, NVIDIA Triton Inference Server, or similar).\nDeep understanding of state-of-the-art machine learning models such as auto-regressive transformers and familiarity with inference optimization techniques (TensorRT, quantization, custom kernels).\nExperience with C++ or CUDA programming for GPU acceleration.\nPrior experience working at autonomous vehicles or robotics companies.\nEqual Opportunity Statement\nWe’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.\nBenefits\nAt Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, parental leave, and unlimited vacation.\nCompensation Range: $170K - $190K","company":"Claryo","rawCompany":"claryo","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-09T14:16:52.508Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Senior Software Engineer - ML Infrastructure","description":"We're looking for a Senior Software Engineer - ML Infrastructure to build and scale the infrastructure that powers our AI-driven warehouse intelligence platform. You'll own the end-to-end lifecycle of computer vision models — from training pipelines through optimized cloud deployment — ensuring our cutting-edge computer vision and multi-modal AI systems run reliably and efficiently in production. Your work will directly enable the real-time perception and autonomous decision-making capabilities at the core of our platform.\nThis is a deeply technical role at the intersection of machine learning, distributed systems, and cloud infrastructure. You'll design scalable GPU compute clusters, build robust orchestration pipelines, and optimize model serving for low-latency inference at scale. You'll work closely with our research scientists, computer vision engineers, and product teams to bridge the gap between experimental models and production-ready systems that operate across diverse warehouse environments. We've found tremendous value in collaborative problem-solving, thus our team works from our SF office three days a week.\nResponsibilities\nDevelop and maintain distributed cloud GPU infrastructure for large-scale world model training and low-latency inference.\nBuild end-to-end computer vision pipelines — from data ingestion and preprocessing through model training, evaluation, and deployment — and integrate them into core product workflows.\nDeploy and optimize state-of-the-art machine learning models in the cloud using model serving platforms and inference optimization techniques, including VLMs and VLAs.\nDesign and operate orchestration systems that enable both engineers and non-engineers to build and manage data and ML pipelines.\nEstablish monitoring, benchmarking, and evaluation frameworks to ensure model performance and reliability in production environments.\nRequired Experience\nB.S. / M.S. in Computer Science, Robotics, or similar technical field, or equivalent practical experience.\n7+ years of professional software engineering experience, with at least 3 years in machine learning infrastructure — developing, scaling, training, deploying, and optimizing large-scale ML systems from data to model.\nTrack record of deploying machine learning models in production environments with real-world constraints.\nExperience with distributed messaging and compute systems (Kafka, gRPC, ROS2, or similar).\nStrong programming skills in Python with solid software engineering practices.\nPreferred Experience\nExperience with training and/or deployment of machine learning models in the computer vision domain.\nExperience developing, running, and managing orchestration systems (Flyte, Temporal, Airflow, or similar) for ML and data pipelines.\nProficiency with ML frameworks (PyTorch, TensorFlow, DeepSpeed) and model serving platforms (TorchServe, TensorFlow Serving, NVIDIA Triton Inference Server, or similar).\nDeep understanding of state-of-the-art machine learning models such as auto-regressive transformers and familiarity with inference optimization techniques (TensorRT, quantization, custom kernels).\nExperience with C++ or CUDA programming for GPU acceleration.\nPrior experience working at autonomous vehicles or robotics companies.\nEqual Opportunity Statement\nWe’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.\nBenefits\nAt Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, parental leave, and unlimited vacation.\nCompensation Range: $170K - $190K","datePosted":"2026-08-09T14:16:52.508Z","dateModified":"2026-08-09T14:16:52.508Z","hiringOrganization":{"@type":"Organization","name":"Claryo","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"fe76e528e18de462bb5a72ff"},"url":"https://jobsearcher.com/jobs/fe76e528e18de462bb5a72ff"}}