Cloud DevOps~AWS DevOps and Automation
Job Description -NEED MINIMUM 3 YEARS EXPERIENCE WORKING IN APPLESkills: Digital : Cloud DevOps~AWS DevOps and AutomationLocation: Sunnyvale, CAOnsite positionRole DescriptionsSupport and administer large-scale Kubernetes platforms hosting mission-critical applications with high availability requirementsDesign, implement, and maintain CI/CD pipelines using GitOps methodologies (Flux), Jenkins, and GitLab for automated build, test, and deployment workflowsDevelop and maintain Infrastructure-as-Code solutions using Terraform and Ansible to provision AWS resources (EKS, EC2, RDS, S3, VPC, IAM, Route 53, Lambda)Build, deploy, and manage containerized applications using Docker, Helm, and KustomizeConfigure and maintain monitoring, alerting, and observability using Prometheus and Grafana; define and track SLIs/SLOs and manage error budgetsPerform incident triage, root cause analysis (RCA), and resolution of complex network, platform, and application issuesParticipate in on-call rotations, responding to high-severity production incidents and restoring services within defined SLOsHarden Kubernetes clusters through network policies, secrets management, and secure container image practicesAutomate operational processes through scripting in Python, Bash, and ShellSupport messaging and data platforms including Kafka, Redis, and RabbitMQCollaborate with Development, QA, and Product teams to resolve environment issues and reduce release bottlenecksOptimize cloud infrastructure costs through right-sizing, autoscaling, and workload schedulingRequired Qualifications5+ years of experience in DevOps, SRE, or platform/cloud engineering rolesStrong hands-on experience administering Kubernetes in production environmentsProficiency with Terraform and Ansible for IaC and configuration managementSolid experience with AWS core services (EKS, EC2, S3, RDS, Lambda, VPC, IAM, Route 53)Experience implementing CI/CD pipelines (Jenkins, GitLab, Azure Pipelines) and GitOps workflows (FluxCD)Scripting proficiency in Python, Bash, or GoExperience with observability tooling (Prometheus, Grafana, Splunk)Experience with containerization (Docker) and Kubernetes packaging tools (Helm, Kustomize)Strong troubleshooting skills across network, platform, and application layersExperience participating in on-call rotations and production incident responseEducationBachelor's degree in Computer Science, Electronics/Communications Engineering, or a related technical discipline (or equivalent experience)