Job Requirements
New York, NY
Top Secret Polygraph Unspecified
Career Level not specified
Salary not specified
Join Premium to unlock estimated salaries
Job Description
Job Title: Infrastructure as Code Engineer - Nutanix, Terraform & Ansible
Job Location: 100% Remote
Duration: 24 months
Interview Mode: 1st round remote via Teams, 2nd round onsite in New York, NY
CANDIDATES MUST BE ABLE TO GO TO 2ND ROUND INTERVIEW ONSITE IN NEW YORK, NY AND ONBOARD FOR 2-3 DAYS ONSITE IN NYC
Position Overview
Our Fortune 500 client located in New York, NY is seeking an experienced Senior Infrastructure as Code Engineer - Nutanix, Terraform & Ansible to design, develop, and maintain automation solutions supporting critical Operational Technology (OT) infrastructure and enterprise applications.
This is a hands-on infrastructure automation and Site Reliability Engineering role focused on transforming traditional infrastructure operations into a repeatable, secure, version-controlled, and highly automated delivery model. The engineer will develop Infrastructure as Code (IaC) and Configuration as Code (CaC) solutions using Terraform, Ansible, HashiCorp Packer, Vault, Git, and Azure DevOps across Nutanix AHV, VMware, Linux, and Windows environments.
The ideal candidate brings strong infrastructure engineering fundamentals combined with deep automation expertise. This individual will build reusable infrastructure modules, standardized server images, CI/CD pipelines, automated recovery processes, and observability solutions that improve platform reliability, security, recoverability, and operational efficiency.
Key Responsibilities
Infrastructure as Code & Automation
The ideal candidate is a hands-on infrastructure automation engineer who thinks like an SRE. You should be equally comfortable writing Terraform modules and Ansible automation, building Packer images, developing Azure DevOps pipelines, scripting with Python or Linux shell, and troubleshooting Nutanix, VMware, and Linux infrastructure.
This position goes beyond traditional infrastructure administration. We are looking for someone who can turn manual operational procedures into reusable, version-controlled automation, establish repeatable infrastructure baselines, improve observability, and build automated recovery capabilities.
Success in this role means creating infrastructure that is not only easier to deploy, but also more reliable, secure, consistent, auditable, and recoverable.
Job Location: 100% Remote
Duration: 24 months
Interview Mode: 1st round remote via Teams, 2nd round onsite in New York, NY
CANDIDATES MUST BE ABLE TO GO TO 2ND ROUND INTERVIEW ONSITE IN NEW YORK, NY AND ONBOARD FOR 2-3 DAYS ONSITE IN NYC
Position Overview
Our Fortune 500 client located in New York, NY is seeking an experienced Senior Infrastructure as Code Engineer - Nutanix, Terraform & Ansible to design, develop, and maintain automation solutions supporting critical Operational Technology (OT) infrastructure and enterprise applications.
This is a hands-on infrastructure automation and Site Reliability Engineering role focused on transforming traditional infrastructure operations into a repeatable, secure, version-controlled, and highly automated delivery model. The engineer will develop Infrastructure as Code (IaC) and Configuration as Code (CaC) solutions using Terraform, Ansible, HashiCorp Packer, Vault, Git, and Azure DevOps across Nutanix AHV, VMware, Linux, and Windows environments.
The ideal candidate brings strong infrastructure engineering fundamentals combined with deep automation expertise. This individual will build reusable infrastructure modules, standardized server images, CI/CD pipelines, automated recovery processes, and observability solutions that improve platform reliability, security, recoverability, and operational efficiency.
Key Responsibilities
Infrastructure as Code & Automation
- Design, develop, maintain, and enhance Terraform modules for automated infrastructure provisioning and lifecycle management.
- Build reusable Ansible playbooks, roles, and configuration automation frameworks.
- Develop automation solutions that reduce manual infrastructure administration and eliminate configuration drift.
- Implement Infrastructure as Code and Configuration as Code practices across critical infrastructure environments.
- Develop automation using Python, Linux shell scripting, and related technologies.
- Maintain infrastructure and configuration code through Git-based version control and established DevOps practices.
- Develop reusable automation standards that enable consistent deployment across environments.
- Automate the provisioning, configuration, and ongoing management of Nutanix AHV and VMware environments.
- Support virtualized data center infrastructure across enterprise and Operational Technology environments.
- Engineer and automate solutions supporting Oracle Linux, Red Hat Linux, Windows, and virtualized infrastructure.
- Create and maintain standardized Oracle Linux and Windows golden images using HashiCorp Packer.
- Establish repeatable server and infrastructure baselines that improve consistency, security, and recoverability.
- Troubleshoot complex infrastructure, virtualization, operating system, and automation issues.
- Design, develop, and maintain CI/CD pipelines within Microsoft Azure DevOps.
- Integrate infrastructure provisioning, configuration management, validation, and deployment processes into automated pipelines.
- Apply Git, DevOps, CI/CD, and version-control practices to infrastructure engineering.
- Implement automated testing and validation to improve deployment quality and reliability.
- Maintain traceability and auditability for infrastructure changes through source control and automated delivery processes.
- Implement HashiCorp Vault for secure secrets management, credential handling, and certificate management.
- Partner with cybersecurity and infrastructure teams to implement secure-by-design automation and infrastructure solutions.
- Integrate security requirements into infrastructure code, server baselines, deployment pipelines, and operational processes.
- Support cybersecurity resilience and infrastructure-hardening initiatives.
- Help ensure infrastructure configurations remain standardized, controlled, auditable, and recoverable.
- Apply Site Reliability Engineering (SRE) principles to improve system availability, reliability, performance, and operational efficiency.
- Develop and support observability solutions utilizing Grafana, Prometheus, Loki, Alertmanager, and related platforms.
- Build dashboards, monitoring, logging, metrics, and alerting capabilities that improve visibility into infrastructure health.
- Participate in troubleshooting, incident analysis, and root cause analysis.
- Identify recurring operational issues and develop automation to reduce or eliminate them.
- Partner with infrastructure and operations teams on continuous reliability improvements.
- Develop automated infrastructure validation, recovery, and disaster recovery workflows.
- Support rapid and repeatable restoration of critical environments.
- Automate recovery testing to improve disaster recovery readiness and business continuity.
- Help establish infrastructure practices that improve system recoverability and cybersecurity resilience.
- Support testing and documentation of recovery procedures and platform dependencies.
- Develop detailed technical documentation, architecture information, operational procedures, and runbooks.
- Create knowledge-transfer materials supporting infrastructure and operations teams.
- Collaborate with infrastructure, operations, cybersecurity, database, and application teams.
- Participate in architecture and technical-design discussions.
- Document automation standards, deployment procedures, recovery processes, and troubleshooting guidance.
- Promote repeatable infrastructure-management practices across engineering and operational teams.
- Strong hands-on experience as an Infrastructure as Code Engineer, SRE Engineer, DevOps Engineer, Infrastructure Automation Engineer, or comparable infrastructure engineering professional.
- Advanced hands-on experience with Terraform for infrastructure provisioning and automation.
- Strong experience developing Ansible playbooks and configuration-management automation.
- Experience with HashiCorp Packer for automated image creation and standardized server builds.
- Strong experience with HashiCorp Vault for secrets management and secure credential handling.
- Hands-on experience with Microsoft Azure DevOps, including automated CI/CD pipelines.
- Strong understanding of Git, DevOps, CI/CD, and version-control practices.
- Hands-on experience with Nutanix infrastructure and virtualization technologies.
- Experience supporting enterprise data center virtualization environments.
- Strong Linux administration experience, including Red Hat Linux and Oracle Linux.
- Strong Linux shell scripting capabilities.
- Experience developing infrastructure and operational automation using Python.
- Experience with observability and monitoring technologies, including Grafana and Prometheus.
- Understanding of modern observability practices involving metrics, monitoring, logging, and alerting.
- Experience developing operational procedures, technical documentation, and runbooks.
- Experience applying SRE principles to infrastructure reliability and operational automation.
- Strong troubleshooting and root cause analysis capabilities.
- Ability to collaborate effectively across infrastructure, operations, cybersecurity, database, and application teams.
- Experience automating Nutanix AHV and VMware provisioning and configuration.
- Experience with Loki and Alertmanager.
- Experience supporting Windows infrastructure in addition to Linux environments.
- Experience creating and maintaining standardized Linux and Windows golden images.
- Experience automating disaster recovery, infrastructure restoration, and recovery validation.
- Experience supporting business continuity and cybersecurity-resilience initiatives.
- Experience developing reusable Terraform modules and enterprise Ansible frameworks.
- Experience integrating Vault secrets and certificate management into automated CI/CD workflows.
- Experience supporting highly controlled, security-sensitive, or Operational Technology environments.
- Experience implementing infrastructure standards designed to minimize configuration drift.
- Strong understanding of infrastructure security, auditability, recovery, and operational resilience.
The ideal candidate is a hands-on infrastructure automation engineer who thinks like an SRE. You should be equally comfortable writing Terraform modules and Ansible automation, building Packer images, developing Azure DevOps pipelines, scripting with Python or Linux shell, and troubleshooting Nutanix, VMware, and Linux infrastructure.
This position goes beyond traditional infrastructure administration. We are looking for someone who can turn manual operational procedures into reusable, version-controlled automation, establish repeatable infrastructure baselines, improve observability, and build automated recovery capabilities.
Success in this role means creating infrastructure that is not only easier to deploy, but also more reliable, secure, consistent, auditable, and recoverable.
group id: 10530356