Terraform IaC Engineer
Responsibilities
Design, develop, and maintain reusable Terraform modules and infrastructure deployments.
Develop and maintain Ansible playbooks for configuration management and infrastructure automation.
Create and maintain Linux and Windows golden images using HashiCorp Packer.
Build and maintain CI/CD pipelines using Azure DevOps.
Implement HashiCorp Vault for secure secrets and certificate management.
Automate provisioning and configuration across Nutanix AHV and VMware environments.
Develop automated infrastructure validation, recovery, and disaster recovery workflows.
Support infrastructure monitoring and observability using Grafana, Prometheus, Loki, and Alertmanager.
Apply SRE principles to improve system reliability, availability, and operational efficiency.
Implement secure, version-controlled infrastructure practices and reduce configuration drift.
Develop and maintain technical documentation, operational runbooks, and knowledge-transfer materials.
Collaborate with infrastructure, operations, security, database, and application teams.
Participate in architecture discussions, troubleshooting, root cause analysis, and reliability improvement initiatives.
Support cybersecurity resilience, business continuity, and disaster recovery efforts.
Required Qualifications
Strong hands-on experience with Terraform and Infrastructure as Code (IaC).
Experience developing automation using Ansible.
Experience with Azure DevOps, Git, CI/CD, and DevOps practices.
Experience with HashiCorp Packer and automated image management.
Hands-on experience with HashiCorp Vault and secrets management.
Strong knowledge of Linux, including Red Hat and Oracle Linux.
Experience with Windows Server environments.
Experience with Nutanix AHV and VMware virtualization platforms.
Proficiency in Python and Shell scripting.
Experience with monitoring and observability platforms such as Grafana and Prometheus.
Familiarity with Loki and Alertmanager.
Understanding of SRE, disaster recovery, infrastructure security, and system reliability.
Strong troubleshooting, documentation, and root cause analysis skills.