{"schemaVersion":"jobsearcher.job.v1","id":"3fdb9e591f3d0bc7ef4b6369","url":"https://jobsearcher.com/jobs/3fdb9e591f3d0bc7ef4b6369","canonicalUrl":"https://jobsearcher.com/jobs/3fdb9e591f3d0bc7ef4b6369","title":"AI DevOps Engineer","description":"About the roleSeeking a highly skilled AI DevOps Engineer to design and manage scalable, secure infrastructure for AI and LLM-powered applications within a regulated financial services environment. This role combines DevOps, Platform Engineering, and Site Reliability Engineering to support high-performance AI systems and workflows.Key responsibilities- Design, deploy, and maintain scalable infrastructure for production AI and LLM applications, ensuring high availability and security.- Develop and manage Infrastructure-as-Code using Terraform to enable secure and repeatable deployments.- Implement and operate Kubernetes environments to support containerized AI workloads at scale.- Establish monitoring, alerting, and incident response procedures to maintain system reliability and performance.- Collaborate with security and compliance teams to uphold regulatory standards and improve automation processes for AI infrastructure.Required skills and experience- Bachelor’s degree in Computer Science, Engineering, or equivalent experience.- Proven experience in DevOps, Platform Engineering, or Site Reliability Engineering roles managing large-scale production infrastructure.- Strong expertise with Terraform and Infrastructure-as-Code methodologies.- Hands-on experience deploying and operating Kubernetes clusters in production environments.- Experience supporting AI platforms or LLM-based workloads with focus on automation, scalability, and cloud-native architectures.Nice to have- Experience supporting production-grade LLM applications and AI agent workflows.- Familiarity with vector databases such as Pinecone, Weaviate, or PostgreSQL with pgvector.- Exposure to AI developer tooling and internal platform support.- Understanding of observability, monitoring, and capacity planning for AI/ML systems.- Experience within financial services or other regulated industries and strong cross-functional communication skills.LocationNew York, NY, United States","company":"Alignity","rawCompany":"alignity","city":"Brooklyn","state":"NY","isRemote":false,"isActive":false,"createdAt":"2026-05-04T03:14:58.759Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI DevOps Engineer","description":"About the roleSeeking a highly skilled AI DevOps Engineer to design and manage scalable, secure infrastructure for AI and LLM-powered applications within a regulated financial services environment. This role combines DevOps, Platform Engineering, and Site Reliability Engineering to support high-performance AI systems and workflows.Key responsibilities- Design, deploy, and maintain scalable infrastructure for production AI and LLM applications, ensuring high availability and security.- Develop and manage Infrastructure-as-Code using Terraform to enable secure and repeatable deployments.- Implement and operate Kubernetes environments to support containerized AI workloads at scale.- Establish monitoring, alerting, and incident response procedures to maintain system reliability and performance.- Collaborate with security and compliance teams to uphold regulatory standards and improve automation processes for AI infrastructure.Required skills and experience- Bachelor’s degree in Computer Science, Engineering, or equivalent experience.- Proven experience in DevOps, Platform Engineering, or Site Reliability Engineering roles managing large-scale production infrastructure.- Strong expertise with Terraform and Infrastructure-as-Code methodologies.- Hands-on experience deploying and operating Kubernetes clusters in production environments.- Experience supporting AI platforms or LLM-based workloads with focus on automation, scalability, and cloud-native architectures.Nice to have- Experience supporting production-grade LLM applications and AI agent workflows.- Familiarity with vector databases such as Pinecone, Weaviate, or PostgreSQL with pgvector.- Exposure to AI developer tooling and internal platform support.- Understanding of observability, monitoring, and capacity planning for AI/ML systems.- Experience within financial services or other regulated industries and strong cross-functional communication skills.LocationNew York, NY, United States","datePosted":"2026-05-04T03:14:58.759Z","dateModified":"2026-05-04T03:14:58.759Z","hiringOrganization":{"@type":"Organization","name":"Alignity","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Brooklyn","addressRegion":"NY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"3fdb9e591f3d0bc7ef4b6369"},"url":"https://jobsearcher.com/jobs/3fdb9e591f3d0bc7ef4b6369"}}