{"schemaVersion":"jobsearcher.job.v1","id":"44910c580c499be39417692c","url":"https://jobsearcher.com/jobs/44910c580c499be39417692c","canonicalUrl":"https://jobsearcher.com/jobs/44910c580c499be39417692c","title":"Data / ML Platform Engineer (Python)","description":"Overview:Data / ML Platform Engineer (Python, LLM Pipelines, Batch Processing)Job SummaryWe are looking for aData / ML Platform Engineerto design and managePython-based batch pipelinesfor evaluating machine learning and LLM models.The role involves building scalable pipelines that process large datasets, interact with LLM APIs, and store structured outputs for analytics and benchmarking.Key ResponsibilitiesML / Data Pipeline DevelopmentBuild and maintainPython batch pipelinesfor:Model evaluation\r\nDataset processingDesignresumable and fault-tolerant workflows\r\nLLM IntegrationIntegrate with:GPT / Claude / other LLM APIsExecute large-scale evaluation workflows\r\nData EngineeringLoad and process data from:Cloud storage (S3, GCS, Azure Blob)Transform and structure data for analytics\r\nDatabase & StorageStore results in:MongoDB (structured outputs)Ensure efficient data retrieval and storage\r\nPerformance & ScalabilityOptimize:Pipeline execution\r\nCost and latencyHandle high-volume batch jobs\r\nCloud & DevOpsDeploy pipelines on:AWS / GCP / AzureUse:CI/CD pipelines\r\nVersion control (Git)Monitoring & ReliabilityImplement:Logging and monitoring\r\nError handling & retriesEnsure pipeline stability\r\n•Required SkillsCore SkillsStrongPython development\r\nExperience with:Batch pipelines / ETL\r\nData processingAI / MLHands-on with:LLM APIs (GPT, Claude)\r\nPrompt engineeringExperience with:RAG pipelines (preferred)Data & StorageMongoDB / NoSQL databases\r\nData structuring and transformation\r\nCloudAWS / GCP / Azure\r\nCloud storage services\r\nDevOpsGit / GitHub\r\nCI/CD pipelines\r\nExperience Required5-8 years total experience\r\n2+ years inML / data pipelines / LLM-based systems","company":"Purple Drive","rawCompany":"purple drive","city":"Tampa","state":"FL","isRemote":false,"isActive":false,"createdAt":"2026-08-09T01:19:09.152Z","occupations":[{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data / ML Platform Engineer (Python)","description":"Overview:Data / ML Platform Engineer (Python, LLM Pipelines, Batch Processing)Job SummaryWe are looking for aData / ML Platform Engineerto design and managePython-based batch pipelinesfor evaluating machine learning and LLM models.The role involves building scalable pipelines that process large datasets, interact with LLM APIs, and store structured outputs for analytics and benchmarking.Key ResponsibilitiesML / Data Pipeline DevelopmentBuild and maintainPython batch pipelinesfor:Model evaluation\r\nDataset processingDesignresumable and fault-tolerant workflows\r\nLLM IntegrationIntegrate with:GPT / Claude / other LLM APIsExecute large-scale evaluation workflows\r\nData EngineeringLoad and process data from:Cloud storage (S3, GCS, Azure Blob)Transform and structure data for analytics\r\nDatabase & StorageStore results in:MongoDB (structured outputs)Ensure efficient data retrieval and storage\r\nPerformance & ScalabilityOptimize:Pipeline execution\r\nCost and latencyHandle high-volume batch jobs\r\nCloud & DevOpsDeploy pipelines on:AWS / GCP / AzureUse:CI/CD pipelines\r\nVersion control (Git)Monitoring & ReliabilityImplement:Logging and monitoring\r\nError handling & retriesEnsure pipeline stability\r\n•Required SkillsCore SkillsStrongPython development\r\nExperience with:Batch pipelines / ETL\r\nData processingAI / MLHands-on with:LLM APIs (GPT, Claude)\r\nPrompt engineeringExperience with:RAG pipelines (preferred)Data & StorageMongoDB / NoSQL databases\r\nData structuring and transformation\r\nCloudAWS / GCP / Azure\r\nCloud storage services\r\nDevOpsGit / GitHub\r\nCI/CD pipelines\r\nExperience Required5-8 years total experience\r\n2+ years inML / data pipelines / LLM-based systems","datePosted":"2026-08-09T01:19:09.152Z","dateModified":"2026-08-09T01:19:09.152Z","hiringOrganization":{"@type":"Organization","name":"Purple Drive","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Tampa","addressRegion":"FL","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"44910c580c499be39417692c"},"url":"https://jobsearcher.com/jobs/44910c580c499be39417692c"}}