Storage & Devops Engineer (Resident)

Indotronix International Corporation
United States
9 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$176,800.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Server Message Block Configuration Management Continuous Integration Data Centers Data Stores Python (Programming Language) Machine Learning NetApp Applications Network File Systems Data ONTAP (Server Appliance) Performance Tuning
+17 more
Windows PowerShell Ansible Runbook Virtualization Technology Large Language Models Snowflake Git Flow Infrastructure Automation Frameworks StorageGrid Hashicorp Restful APIs Terraform SnapMirror Software Version Control FlexClone Pagerduty Servicenow

Job description

Client is seeking a senior NetApp Resident Engineer to embed with the Storage Systems Group (SSG) and accelerate automation and AI-enabled operations across a large multi-site ONTAP fleet. This role pairs deep NetApp ONTAP expertise with hands-on Infrastructure-as-Code (Terraform/Ansible) skills and practical experience applying AI/ML tooling to storage operations - including anomaly detection, capacity forecasting, and AI-assisted troubleshooting. The resident engineer will work directly with SSG leadership and engineers on active initiatives spanning fleet lifecycle management and ServiceNow ITOM/AIOps integration., * Design, build, and maintain Terraform modules and Ansible playbooks/roles for ONTAP provisioning, SVM/volume lifecycle management, SnapMirror/SnapCenter operations, and fleet-wide configuration drift remediation.

  • Partner with SSG engineers to extend existing automation (e.g., self-service backup/restore workflows, capacity reclamation scripting) into standardized, version-controlled IaC pipelines.
  • Apply AI/ML capabilities - including NetApp BlueXP/AIOps tooling, anomaly detection, and LLM-assisted diagnostics - to reduce time-to-resolution and to support predictive capacity and health management across the fleet.
  • Support integration work between ONTAP telemetry, Metabase/Snowflake reporting, and ServiceNow ITOM Event Management webhook pipelines.
  • Support NetApp snapshot and clone technology (Snapshot, FlexClone) automation and NFS datastore lifecycle management across the fleet.
  • Document runbooks, automation architecture, and operational procedures; provide knowledge transfer and upskilling to SSG engineers on IaC and AI-assisted operations practices., * Contract engagement placed by NetApp to work embedded within Client’’s SSG team.
  • Remote/hybrid; occasional coordination across U.S. data center sites as fleet work requires.
  • Success will be measured by automation coverage delivered (Terraform/Ansible modules in production use), reduction in manual operational toil, and measurable AI-assisted improvements to incident response or capacity planning.

Requirements

  • 8+ years of experience with NetApp ONTAP administration in enterprise, multi-cluster environments (SVMs, FlexVols/FlexGroups, SnapMirror, SnapCenter, CIFS/NFS).
  • 3+ years hands-on experience writing and maintaining Terraform for infrastructure provisioning, including NetApp/ONTAP or adjacent storage/infrastructure providers.
  • 3+ years hands-on experience with Ansible for configuration management and operational automation (playbooks, roles, idempotent task design).
  • Demonstrated experience applying AI/ML or LLM-based tooling to infrastructure or storage operations (e.g., predictive analytics, anomaly detection, AI-assisted scripting or troubleshooting, AI-output verification practices).
  • Strong scripting ability in Python and/or PowerShell for automation glue code, API integration, and reporting.
  • Working knowledge of REST API-based automation against ONTAP and adjacent platforms.
  • Experience with version control and CI/CD practices (Git-based workflows) for infrastructure code.
  • Excellent written communication skills for runbook and architecture documentation., * Experience with NFS-backed datastore performance tuning (NFSv3 vs. NFSv4.1/4.2) across virtualized environments.
  • Familiarity with Dell PPDM/Data Domain, StorageGRID, or other backup/DR platforms.
  • Experience integrating storage/infrastructure telemetry into ITSM/ITOM platforms (ServiceNow Event Management, PagerDuty).
  • NetApp certifications (NCDA, NCIE) and/or HashiCorp Terraform Associate certification.
  • Prior experience in a vendor-resident or embedded consulting engagement model.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

5:02 min

Mapping Git flow branches to application tester segments

Majid Hajian · LIVE

4:45 min

Building careers inside distributed technology consulting environments

Oliver Zimmert · LIVE

3:53 min

Introduction to git flow and clean feature branches

Johannes Haux · World Congress 2022

3:19 min

Executing complex workflows using Ansible Automation Platform

Goetz Rieger Goetz Rieger · World Congress 2025

Videos

See all

Related articles

See all