{"schemaVersion":"jobsearcher.job.v1","id":"dd0f47493ac1e8a67ddb31b6","url":"https://jobsearcher.com/jobs/dd0f47493ac1e8a67ddb31b6","canonicalUrl":"https://jobsearcher.com/jobs/dd0f47493ac1e8a67ddb31b6","title":"MLOps / ML Platform Engineer","description":"1. About Our Client:The organization operates within the machine learning infrastructure and platform engineering space. It addresses challenges related to managing scalable, reliable, and cost-efficient machine learning workflows and production systems. By focusing on optimizing compute resources, ensuring operational reliability, and automating deployment pipelines, the organization supports high-throughput model workflows and the integration of machine learning models into production environments.2. About the Opportunity:The MLOps / ML Platform Engineer role is responsible for designing, operating, and optimizing machine learning infrastructure that supports the end-to-end lifecycle of model development and deployment. This position plays a critical role in ensuring that machine learning models are efficiently trained, evaluated, and served in production while maintaining reliability, security, and cost-effectiveness. The role contributes by enabling smooth collaboration between research and engineering teams and by developing tools and automation that enhance the repeatability and safety of ML workflows.3. Responsibilities:Design and manage ML infrastructure including data, training, serving, and inference systemsBuild scalable, reproducible training and evaluation pipelines with versioning and schedulingOptimize compute workloads and costs through tuning and cluster managementOperate production model serving with autoscaling and deployment safety measuresDefine and monitor service level objectives (SLOs) and track system observability metricsManage security aspects such as IAM, secrets, and container securityAutomate deployment pipelines using CI/CD and infrastructure as codeCollaborate with research scientists and AI engineers to facilitate model productionCreate documentation, templates, and internal tools to improve ML workflow efficiency4. Requirements:Minimum 4 years of experience in ML platform, DevOps, or infrastructure engineeringStrong knowledge of Kubernetes, CI/CD, containers, and cloud platforms (AWS, GCP, or Azure)Experience managing GPU clusters and ML training/inference pipelinesFamiliarity with data orchestration and storage formats such as Delta, Parquet, Polars, and SparkProven ability to deploy and operate production ML systems with defined SLOsProficient in Python and automation using infrastructure as codeExperience with observability tools and cost optimization at scale5. Pay Range and Compensation Package:The pay range and compensation package for this role will be determined based on the candidate’s experience, skills, and other relevant factors.6. Benefits & Perks:Comprehensive health insurance planRetirement savings plan (401k) with company matchRemote working environmentFlexible, unlimited time off policyThirteen paid holidays annually, including the Monday after the Super BowlEqual Opportunity Statement:Equal Opportunity Statement: Our client is an equal opportunity employer. They celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, or national origin.Note:RemoteHunter is a recruitment partner of this role. Please note that all employment decisions, including candidate assessment, interviews, hiring, compensation, and employment terms, are made exclusively by the hiring employer.","company":"Remotehunter","rawCompany":"remotehunter","city":"Denver","state":"CO","isRemote":false,"isActive":false,"createdAt":"2026-09-28T09:47:22.639Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"MLOps / ML Platform Engineer","description":"1. About Our Client:The organization operates within the machine learning infrastructure and platform engineering space. It addresses challenges related to managing scalable, reliable, and cost-efficient machine learning workflows and production systems. By focusing on optimizing compute resources, ensuring operational reliability, and automating deployment pipelines, the organization supports high-throughput model workflows and the integration of machine learning models into production environments.2. About the Opportunity:The MLOps / ML Platform Engineer role is responsible for designing, operating, and optimizing machine learning infrastructure that supports the end-to-end lifecycle of model development and deployment. This position plays a critical role in ensuring that machine learning models are efficiently trained, evaluated, and served in production while maintaining reliability, security, and cost-effectiveness. The role contributes by enabling smooth collaboration between research and engineering teams and by developing tools and automation that enhance the repeatability and safety of ML workflows.3. Responsibilities:Design and manage ML infrastructure including data, training, serving, and inference systemsBuild scalable, reproducible training and evaluation pipelines with versioning and schedulingOptimize compute workloads and costs through tuning and cluster managementOperate production model serving with autoscaling and deployment safety measuresDefine and monitor service level objectives (SLOs) and track system observability metricsManage security aspects such as IAM, secrets, and container securityAutomate deployment pipelines using CI/CD and infrastructure as codeCollaborate with research scientists and AI engineers to facilitate model productionCreate documentation, templates, and internal tools to improve ML workflow efficiency4. Requirements:Minimum 4 years of experience in ML platform, DevOps, or infrastructure engineeringStrong knowledge of Kubernetes, CI/CD, containers, and cloud platforms (AWS, GCP, or Azure)Experience managing GPU clusters and ML training/inference pipelinesFamiliarity with data orchestration and storage formats such as Delta, Parquet, Polars, and SparkProven ability to deploy and operate production ML systems with defined SLOsProficient in Python and automation using infrastructure as codeExperience with observability tools and cost optimization at scale5. Pay Range and Compensation Package:The pay range and compensation package for this role will be determined based on the candidate’s experience, skills, and other relevant factors.6. Benefits & Perks:Comprehensive health insurance planRetirement savings plan (401k) with company matchRemote working environmentFlexible, unlimited time off policyThirteen paid holidays annually, including the Monday after the Super BowlEqual Opportunity Statement:Equal Opportunity Statement: Our client is an equal opportunity employer. They celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, or national origin.Note:RemoteHunter is a recruitment partner of this role. Please note that all employment decisions, including candidate assessment, interviews, hiring, compensation, and employment terms, are made exclusively by the hiring employer.","datePosted":"2026-09-28T09:47:22.639Z","dateModified":"2026-09-28T09:47:22.639Z","hiringOrganization":{"@type":"Organization","name":"Remotehunter","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Denver","addressRegion":"CO","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"dd0f47493ac1e8a67ddb31b6"},"url":"https://jobsearcher.com/jobs/dd0f47493ac1e8a67ddb31b6"}}