Senior Windows Platform Engineer (Automation / SRE)

Page Michael International Inc
New York, NY, United States
7 days ago
Apply on www.michaelpage.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$250,000.0 - $300,000.0
Working hours
Regular working hours

Tech stack

Microsoft Windows Active Directory Application Programming Interfaces (APIs) Application Frameworks Build Automation Microsoft Azure Microsoft Online Services Cloud Engineering Configuration Management Continuous Integration Software Debugging Programming Tools
+23 more
Disaster Recovery Distributed Systems Github Identity and Access Management Nagios Operational Data Store Windows PowerShell Reliability Engineering Ansible Prometheus Datadog Data Logging Workspace ONE System Availability Grafana Reliability of Systems Microsoft InTune Deployment Automation Api Design Puppet Terraform Splunk Jenkins

Job description

We are seeking a Senior Windows Platform Engineer to join our Enterprise Engineering team. This role is focused on automation, infrastructure engineering, and platform reliability across a global Microsoft ecosystem.

This is not a traditional Windows administration position. We are looking for an engineer who leverages PowerShell as a development tool, applies Infrastructure-as-Code and Configuration-as-Code principles to manage systems at scale, and uses observability data to drive operational decisions.

You will partner with a globally distributed engineering team to design, build, and maintain automated solutions supporting Windows endpoints, server infrastructure, Microsoft Azure, identity platforms, and enterprise device management systems.

What You’ll Do

  • Build and maintain PowerShell-based automation, tooling, and reusable frameworks
  • Develop solutions that eliminate repetitive operational work through automation
  • Implement Infrastructure-as-Code and Configuration-as-Code practices across enterprise environments
  • Provision and manage Azure infrastructure using Terraform and related IaC technologies
  • Manage Windows systems through code-based configuration management platforms such as Ansible, Chef, or Puppet
  • Design and support CI/CD workflows that improve speed, consistency, and reliability of infrastructure deployments
  • Leverage monitoring, logging, and telemetry data to improve platform performance, reliability, and operational efficiency
  • Build automation and integrations using Microsoft Graph APIs
  • Engineer and support solutions across Azure, Entra ID, Active Directory, Intune, Autopilot, and Workspace ONE environments
  • Identify opportunities to improve resiliency, scalability, security, and operational maturity throughout the enterprise platform
  • Participate in troubleshooting and root cause analysis of complex infrastructure and automation issues

Requirements

  • Experience designing, writing, debugging, and maintaining production-grade PowerShell scripts and automation frameworks
  • Ability to build reusable tools, modules, and workflows rather than simply executing administrative commands
  • Strong understanding of scripting best practices, error handling, logging, testing, and code maintainability

Infrastructure & Automation

  • Experience implementing Infrastructure-as-Code using Terraform or equivalent tooling
  • Experience with Configuration-as-Code platforms including Ansible, Chef, Puppet, or similar technologies
  • Experience building automation pipelines using Jenkins, Azure DevOps, GitHub Actions, or equivalent CI/CD tools
  • Strong understanding of infrastructure lifecycle automation and platform engineering principles

Microsoft Ecosystem

  • Deep experience supporting enterprise Windows environments
  • Strong Azure infrastructure knowledge
  • Experience with Microsoft Graph API development and automation
  • Experience with Intune, Autopilot, Entra ID, and Active Directory
  • Understanding of endpoint management and enterprise identity services

Reliability Engineering & Observability

  • Experience leveraging monitoring, logging, and telemetry platforms to improve system reliability
  • Familiarity with tools such as Prometheus, Grafana, Splunk, Nagios, Datadog, or similar observability platforms
  • Ability to use metrics and operational data to drive troubleshooting, root cause analysis, and engineering decisions
  • Strong understanding of reliability, scalability, and operational excellence concepts

Preferred Experience

  • Site Reliability Engineering (SRE) experience
  • Experience operating large-scale Windows environments
  • Experience managing globally distributed infrastructure
  • Exposure to cloud-native operations and modern platform engineering practices
  • Strong understanding of high availability, disaster recovery, and infrastructure resiliency

Benefits & conditions

New York, New York Permanent USD250,000 - USD300,000 per year View Job Description In this role, success is measured by your ability to automate and engineer solutions rather than manually administer systems. The strongest candidates will have a software-oriented approach to infrastructure, demonstrate deep PowerShell development experience, embrace IAC and Configuration-as-Code principles, and use metrics, logging, and observability data to continuously improve platform reliability and operational efficiency.

  • Permanent opportunity with established firm.
  • Join a tech-forward financial services firm., * Competitive salary ranging from $250,000 to $300,000 per year.
  • Comprehensive benefits package.
  • 401(k) match for retirement planning.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.michaelpage.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

Videos

See all

Related articles

See all