VNOC Resiliency Specialist

Mount Indie
Gilbert, AZ, United States
2 days ago
Apply on www.clearancejobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Application Performance Management Systems Engineering JIRA Cloud Computing Monitoring of Systems Information Technology Operations Queue Management Systems Cloud Services Runbook Software Deployment Data Logging
+4 more
SC Clearance SolarWinds (Software) Oracle Cloud Infrastructure Servicenow

Job description

The VNOC Resiliency Specialist (Day Shift) supports day-to-day operations as the primary monitoring, routing, and triage authority alongside Incident Managers. Operating with high autonomy during off-hours, this team member serves as our first line of defense in maintaining on-premise and cloud infrastructure stability.

Our Ideal Candidate

We are seeking a highly reliable, process-driven professional who remains composed under pressure and thrives in a dynamic environment. This role does not require deep system engineering; instead, it demands vigilance and execution.

Core Focus

Your daily mission focuses on dashboard surveillance, rapid alert triage, and runbook execution to support organizational resiliency.

This role is a hybrid position located in Gilbert, AZ with 1-2 days onsite per 14 day period; mission dependent.

Core Functional Responsibilities

  • Dashboard Surveillance: Maintain high-vigilance monitoring of enterprise network, on-premise, and cloud infrastructure health utilizing SolarWinds, Elastic, and Application Performance Monitoring (APM) dashboards to immediately catch system degradation or outages.
  • Queue Management & Triage: Triage incoming outage calls, acknowledge automated system alerts, prioritize the queue, and route events to the correct technical engineering teams based on established routing rules.
  • First-Response & Escalation: Provide composed, immediate first-line response during Major Incidents by executing standard step-by-step runbooks. Coordinate closely with Incident Managers to escalate issues and engage technical Subject Matter Experts (SMEs) as required.
  • Cloud Health Monitoring: Monitor high-level system alerts and service health dashboards within AWS and OCI (Oracle Cloud Infrastructure) environments to identify cloud service disruptions and initiate standard escalation workflows.
  • Ticketing & SLA Compliance: Manage the administrative lifecycle of incidents within ServiceNow and Jira, ensuring precise documentation of event timelines, ticket updates, and strict adherence to established SLA response times.
  • Shift Turnover & Incident Logging: Maintain precise shift turnover logs and conduct formal, detailed handovers to incoming day-shift personnel. Assist Incident Managers by compiling chronological event timeline data for post-incident reviews and After Action Reports (AARs).
  • Change Window Monitoring: Support change management activities by observing dashboard health statuses via SolarWinds and APM for anomalies during scheduled maintenance and software deployment windows tracked in ServiceNow.
  • Runbook & Process Adherence: Strictly follow established VNOC runbooks, Tactical Techniques and Procedures (TTPs), and SOPs. Flag outdated documentation in Jira/ServiceNow repositories to ensure instructions remain accurate.

Requirements

To succeed in this functional role, you will be a disciplined, detail-oriented operator who excels at executing established procedures.

  • Process-Oriented & Composed: You respect established SOPs when critical alerts trigger. You can clearly follow step-by-step triage checklists during high-stress incidents.
  • Autonomous & Reliable: You are highly self-motivated and punctual, capable of maintaining focus during quiet night-shift hours without direct leadership supervision.
  • Strong Communicator: You possess excellent written communication skills, essential for drafting clear chronological turnover logs and logging precise ticket updates.

Required Experience:

  • Current Secret clearance or higher
  • 3 years of experience in an IT Operations Center (NOC/SOC) environment, Tier 1/2 IT Helpdesk, or relevant Military Operations (such as communications, cyber, or logistics).
  • Fundamental Cloud Knowledge: Basic understanding of cloud infrastructure concepts and exposure to AWS or OCI environments (such as AWS Certified Cloud Practitioner or Oracle Cloud Infrastructure Foundations level).

Desired Experience:

  • Familiarity with enterprise ticketing platforms (ServiceNow, Jira) and monitoring tools (such as SolarWinds, Elastic, APM) is highly preferred.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.clearancejobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Intersecting neurodivergent support and technical operations consulting

56 sec

Integrating automated approval workflows into the portal

Markus Eisele Markus Eisele · World Congress 2025

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

2:27 min

Establishing a simulated technical environment for the workflow demo

Tobias Dunn-Krahn · LIVE

1:32 min

Structuring automated incident workflows between runbooks and raw models

Aram Hakobyan Aram Hakobyan +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all