Lead Data Center NOC Engineer

Armada LTD
San Francisco, CA, United States
7 days ago
Apply on www.workingnomads.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Compensation
$126,000.0 - $157,000.0
Working hours
Shift work

Tech stack

LTE (Telecommunication) JIRA Bash Shell Data Centers Data Center Infrastructure Management (CIM) Noise Reduction Linux Distributed Data Store Monitoring of Systems IP Addressing Virtual Private Networks (VPN) Python (Programming Language)
+14 more
Logical Security Windows Servers Network Diagrams Windows PowerShell Runbook Virtual Local Area Networks Network Routers Scripting Grafana Mttr Break Fix Firewalls (Computer Science) SolarWinds (Software) Servicenow

Job description

We are seeking a Lead Level 2 Data Center NOC Engineer to provide technical leadership and operational ownership across modular, edge, and distributed data center environments. This lead role acts is responsible leading incident response, mentoring NOC staff, and ensuring uptime of critical infrastructure including power, cooling, and networking systems. The ideal candidate combines deep hands-on troubleshooting with strong leadership, communication, and process ownership., This position may be performed remotely in the U.S. The Greater Seattle Area is strongly preferred, given Armada’s operations and the collaborative nature of the role., * Act as the senior technical lead on shift, guiding and supporting L1 NOC technicians.

  • Serve as primary escalation point for incidents until resolution or handoff to L3 Engineering.
  • Lead shift handovers, prioritization, and operational decision-making.
  • Mentor and coach junior NOC staff and support onboarding.
  • Represent the NOC during cross-functional incident bridges.

Incident Command & Reliability:

  • Serve as Incident Commander for medium to high-severity incidents.
  • Own end-to-end incident lifecycle: detection, triage, mitigation, communication, and resolution.
  • Drive root cause analysis (RCA) and corrective actions for recurring issues.
  • Ensure high-quality, executive-ready incident communications.
  • Continuously improve MTTR and operational reliability.

Infrastructure & Data Center Operations:

  • Monitor and troubleshoot physical infrastructure using PLC, BMS, and DCIM platforms.
  • Perform health checks on UPS, PDUs, generators, CRAC/CRAH systems.
  • Troubleshoot MEP-related issues and coordinate maintenance and remote hands.
  • Support modular, containerized, and micro data center deployments.
  • Enforce physical and logical security policies.

Networking Operations & Troubleshooting (L2):

  • Monitor and troubleshoot switches, routers, firewalls, and edge connectivity.
  • Perform L2/L3 troubleshooting (VLANs, IP addressing, MTU, routing fundamentals).
  • Validate fiber/copper connectivity, optics, link status, and redundancy paths.
  • Troubleshoot VPNs and secure remote access solutions.
  • Escalate architecture or design issues to Network Engineering (L3).

Tooling, Runbooks & Automation:

  • Use observability and ITSM platforms (Grafana, Zenduty, ServiceNow, Jira, SolarWinds).
  • Own and maintain operational runbooks, SOPs, and escalation procedures.
  • Improve alerting quality, dashboards, and noise reduction.
  • Support automation, reporting, audits, and compliance requests.

Collaboration & Change Management:

  • Partner with Engineering and Product teams.
  • Participate in change planning, execution, and post-change validation.
  • Identify operational gaps and drive continuous improvement initiatives.

Requirements

  • 7+ years experience in NOC, data center, or infrastructure operations.
  • Demonstrated experience leading incidents or technical shifts.
  • Strong hands-on experience with DCIM/BMS platforms (Distech, Schneider, RadixIOT).
  • Advanced L2/L3 networking troubleshooting experience.
  • Solid understanding of power, cooling, and monitoring systems.
  • Ability to interpret electrical one-lines and network diagrams.
  • Comfortable with shift work and on-call rotations.

Preferred Experience and Skills

  • Certifications: CCNA, JNCIA, CDCTP, CDCP, or equivalent.
  • Experience with edge/remote connectivity (VSAT, LTE/5G, Starlink).
  • Familiarity with Linux/Windows server environments.
  • Basic scripting (Python, Bash, PowerShell)., * A go-getter with a growth mindset. You’re intellectually curious, have strong business acumen, and actively seek opportunities to build relevant skills and knowledge
  • A detail-oriented problem-solver. You can independently gather information, solve problems efficiently, and deliver results with a ‘get-it-done’ attitude
  • Thrive in a fast-paced environment. You’re energized by an entrepreneurial spirit, capable of working quickly, and excited to contribute to a growing company
  • A collaborative team player. You focus on business success and are motivated by team accomplishment vs personal agenda
  • Highly organized and results-driven. Strong prioritization skills and a dedicated work ethic are essential for you

Benefits & conditions

For U.S. Based candidates: To ensure fairness and transparency, the starting base salary range for this role for candidates in the U.S. are listed below, varying based on location experience, skills, and qualifications.

We use a geographic pay structure based on cost-of-labor markets.

  • Tier 1 (e.g., SF Bay Area, NYC, Seattle): $144,715 - $180,890
  • Tier 2 (most U.S. metro areas): $125,840 - $157,300
  • Tier 3 (other cities): $119,548 - $149,435

Final compensation will be determined by experience, scope, and level, and may vary from the posted range.

In addition to base salary, this role will also be offered equity and subsidized benefits (details available upon request)., * Competitive base salary and equity

  • Medical, dental, and vision (subsidized cost)
  • Health savings accounts (HSA), flexible spending accounts (FSA), and dependent care FSAs (DCFSA)
  • Retirement plan options, including 401(k) and Roth 401(k)
  • Unlimited paid time off (PTO)
  • 14 paid company holidays per year, $125,840-$157,300 USD

About the company

Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company’s Most Innovative Companies and to the CNBC Disruptor 50, Armada’s solutions are deployed in over 60 countries globally for organizations ranging from energy to defense.

With nearly $500 million in funding to date, Armada is backed by leading investors including Founders Fund, Lux, BlackRock and Microsoft (M12), alongside strategic partnerships with Microsoft, Dell, Palantir, NVIDIA, SpaceX, and Skydio. We are building the infrastructure layer for sovereign and edge AI - rugged, deployable compute for customers that cannot rely on centralized cloud.

Working at Armada means taking ownership, driving autonomy, and delivering impact. You’ll tackle challenges that haven’t been solved before and help build something transformative from the ground up. What you do here will not only define your career but help further Armada’s mission to bridge the digital divide for customers around the world.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.workingnomads.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:48 min

Daily responsibilities and alignment practices for technical engineering leadership

Edoardo Dusi · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · World Congress 2021

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

3:07 min

Establishing service level agreements directly for internal platforms

Pawel Piwosz · LIVE

Videos

See all

Related articles

See all