System Administrator

Applied Systems
United States
1 day ago
Apply on careers-appliedsystems.icims.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Microsoft Azure Bash Shell Cloud Computing Cloud Engineering Configuration Management Data Centers VMware ESX Servers Monitoring of Systems Python (Programming Language) System Center Configuration Manager Windows Servers Windows PowerShell
+23 more
Reliability Engineering Ansible Systems Integration Virtual Machines VMware VSphere Datadog Google Cloud Cloud Platform System Grafana Software Troubleshooting Reliability of Systems HybridCloud Gitlab-ci Infrastructure Automation Frameworks Information Technology SolarWinds (Software) Vcenter Patch Management Terraform New Relic (SaaS) Appdynamics Wsus Vmware

Job description

Be an Early Applicant Remote or Hybrid Hiring Remotely in United States Mid level Remote or Hybrid Hiring Remotely in United States Mid level Administer hybrid Windows Server infrastructure across colocation, on-premises, Azure, and GCP environments. Manage virtual machines, cloud resources, patching, upgrades, backups, monitoring, alerts, incident response, vulnerability remediation, and root cause analysis. Develop automation using PowerShell, Python, Ansible, and infrastructure-as-code tools for provisioning, remediation, and self-healing workflows. Maintain documentation and runbooks while collaborating with engineering, IT support, security, and operations teams. The summary above was generated by AI, We’re seeking an experienced and highly skilled Systems Administrator to join our infrastructure team. In this role, you will support and maintain hybrid infrastructure spanning colocation data centers and cloud environments, including Azure and Google Cloud Platform (GCP). You will be a key contributor to the day-to-day operation and continuous improvement of our infrastructure - responsible for Windows Server administration, virtual machine lifecycle management, cloud platform operations, patch management, monitoring and observability, and infrastructure automation. You will help ensure our systems remain reliable, secure, and operationally efficient while identifying opportunities to eliminate repetitive manual work and turn recurring incidents into permanent solutions. The ideal candidate is a self-sufficient infrastructure operator with strong Windows Server experience, a solid understanding of hybrid cloud environments, strong troubleshooting skills, and an automation mindset. This is an individual contributor role focused on execution, reliability, operational excellence, and continuous improvement. What You’ll Do

  • Administer and support Windows Server environments across colocation and cloud infrastructure, including provisioning, configuration, upgrades, patching, and day-to-day operational support
  • Manage virtual machine lifecycles, including provisioning, resource management, right-sizing, operational health, and decommissioning across on-premises and cloud environments
  • Support and maintain infrastructure across colocation facilities and cloud environments, including Azure and GCP
  • Manage cloud resources including virtual machines, storage, networking, and core platform services
  • Perform system upgrades, patching, and backup operations across cloud and on-premises environments
  • Execute and maintain patch management processes across Windows Server environments, including maintenance windows, patching cadences, and pre- and post-patch validation
  • Respond to security vulnerability findings by triaging, scheduling, and remediating identified exposures in partnership with security teams
  • Monitor system performance, availability, and capacity across cloud and on-premises environments using monitoring and observability platforms
  • Build, tune, and maintain monitors, alerts, and dashboards to improve signal quality, reduce noise, and identify issues before they become incidents
  • Respond to infrastructure alerts and incidents, perform root cause analysis, and implement corrective actions
  • Identify recurring operational issues and use automation or permanent fixes to improve system reliability
  • Build and maintain automation for recurring operational processes such as patching, VM provisioning, health checks, remediation, and reporting
  • Develop scripts and automated solutions using PowerShell, Python, Ansible, or similar technologies to eliminate repetitive operational tasks
  • Integrate automation with monitoring and observability platforms to support automated responses, proactive checks, and self-healing workflows
  • Support Infrastructure as Code practices for repeatable and consistent environment provisioning
  • Create and maintain technical documentation, operational runbooks, and procedural guides
  • Collaborate with development, cloud engineering, IT support, and operations teams to support system integration and service delivery
  • Communicate clearly regarding system status, incident updates, and project progress
  • Contribute to post-incident reviews and continuous improvement of operational processes

Requirements

  • Bachelor’s degree in Engineering, Computer Science, or a related field, or equivalent practical experience
  • 3-6 years of experience in systems administration, infrastructure operations, or cloud operations
  • Strong hands-on experience administering Windows Server environments, including production patching, upgrades, and troubleshooting
  • Experience managing virtual machine infrastructure, including provisioning, resource management, and lifecycle operations across on-premises and/or cloud platforms
  • Experience with a modern monitoring and observability platform; Datadog preferred, with equivalent experience in SolarWinds, Grafana, AppDynamics, New Relic, or similar platforms considered
  • Proficiency in at least one scripting language such as PowerShell, Python, or Bash
  • Experience operating hybrid infrastructure across colocation or on-premises data centers and public cloud environments
  • Experience with patch management tooling such as Ansible, PatchMyPC, SCCM, or WSUS in a production environment
  • Strong troubleshooting and problem-solving skills across infrastructure, operating systems, and cloud environments
  • Experience with Azure and Google Cloud Platform (GCP)
  • Familiarity with Infrastructure as Code tools such as Terraform or ARM templates
  • Experience with Ansible for configuration management and patching automation
  • Familiarity with SRE concepts, including SLIs, SLOs, error budgets, toil reduction, and reliability engineering practices
  • Exposure to VMware vSphere, ESXi, and vCenter administration
  • Exposure to GitLab CI/CD or similar pipelines for infrastructure automation workflows
  • Experience with monitoring and observability technologies such as Datadog, OpenTelemetry, Signoz, or equivalent platforms
  • Strong written and verbal communication skills, with the ability to create clear technical documentation and operational runbooks
  • Relevant certifications such as AZ-104 (Azure Administrator), Google Cloud Associate, VMware VCP, or Ansible equivalent are a plus

Benefits & conditions

A culture that values who you are and recognizes that you aren’t just an employee; you are a teammate, and you matter. We thrive on the benefits of our different experiences and celebrate the uniqueness our teammates bring to work with them every day. We flex our time together, collaborating remotely and in-person to empower our teams to work in the ways that work best for them. A comprehensive benefits and compensation package that centers our teammates and helps them to bring their best to work every day: Medical, Dental, and Vision Coverage Holiday and Vacation Time Health & Wellness Days

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careers-appliedsystems.icims.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

7:31 min

Essential foundational skills and concepts for infrastructure roles

Megha Kadur · LIVE

1:40 min

Managing containerized infrastructure with Podman Desktop

Cedric Clyburn Cedric Clyburn +1 · World Congress 2025

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:41 min

Parallels between cloud and legacy infrastructure lock-ins

Björn Stahl Björn Stahl · World Congress 2024

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

3:46 min

The history of abstractions and hardware virtualization

Edoardo Dusi · LIVE

Videos

See all

Related articles

See all