DevOps Engineer
About The RoleThe role owns the reliability, scalability, and security of production infrastructure supporting high-traffic applications serving millions of users globally.You will work alongside software engineering teams to design resilient systems, automate deployment pipelines, and establish rigorous observability standards.Key ResponsibilitiesDesign, provision, and manage cloud infrastructure on AWS using Terraform and infrastructure-as-code best practicesBuild and maintain robust CI/CD pipelines using GitHub Actions or GitLab CI for seamless, automated deploymentsManage Kubernetes clusters in production, ensuring high availability, optimal resource utilization, and secure cluster configurationImplement comprehensive monitoring, logging, and alerting systems using Prometheus, Grafana, and DatadogParticipate in an on-call rotation to troubleshoot and resolve production incidents, conducting thorough post-mortems to prevent recurrenceEnforce security baselines, manage IAM roles, and participate in compliance and vulnerability remediation effortsWhat We Are Looking For3–6 years of experience in DevOps, Site Reliability Engineering, or systems engineering roles in high-growth environmentsStrong proficiency with AWS core services (EKS, EC2, RDS, IAM, VPC) and Infrastructure-as-Code tools like TerraformDeep operational experience with Kubernetes, Docker containerization, and service mesh architecturesSolid programming and scripting skills in Python, Go, or Bash for automation and toolingDeep understanding of networking concepts (TCP/IP, DNS, TLS, load balancing) and Linux system administrationBonus: Experience with service mesh technologies like Istio, FinOps cost optimization practices, or holding active AWS/CKA certifications