> Markdown version of [/jobs/ext/3053319-server-engineer](https://www.wearedevelopers.com/jobs/ext/3053319-server-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Server Engineer - **Company:** True North ITG, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $91,098.0 - $109,709.0 - **Contract:** Permanent contract - **Skills:** Proxmox, Link Aggregation (Ethernet), Active Directory, Artificial Intelligence, Backup Devices, Intelligent Platform Management Interface, Bash Shell, Border Gateway Protocol, Health Informatics, BIOS, Cisco PIX, Data Centers, Debian Linux, Linux, RAID, Domain Name System (DNS), Network Interface Controllers, Trunking, Firmware, Virtual Private Networks (VPN), Python (Programming Language), Kernel-Based Virtual Machine, Network Security, Network Layer, Linux System Administration, Windows Servers, Network Diagrams, Network Segmentation, Packet Analyzer, Open Shortest Path First (OSPF), PCI Express, Performance Tuning, Ansible, Virtual Local Area Networks, Virtualization Technology, Wide Area Networks, Zabbix, Ceph (Software), Scripting, Grafana, Caching, Firewalls (Computer Science), Backend, Bare Metal, PfSense, Fortinet, ZFS File System, Hardware Infrastructure, Restful APIs, Terraform, Network Server, Cisco, Docker, Nvme - **Published:** September 24, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=7177071488d7e8f2 ## About the Role You do not need to check every box in the lower tiers, but the first tier is essential. 1. Proxmox, Ceph, ZFS and KVM (essential, very high level) * 5+ years running Proxmox VE and KVM in production, including multi-node clusters, HA and live migration * Deep, hands-on Ceph experience: you have built clusters from scratch, tuned them under load and recovered them from real failures * Strong ZFS knowledge: pool design, ARC/L2ARC/SLOG, replication and performance tuning * Able to design and deploy a complete cluster alone, and to explain how compute, storage, network and backup layers interact * Strong Linux administration (Debian preferred) and Bash or Python scripting 2. Enterprise networking (strong) * Working knowledge of OSPF and BGP in production environments * VLAN design, trunking, LACP/bonding and MTU/jumbo frame configuration for storage networks * General SD-WAN experience * Ability to troubleshoot Layer 2 and Layer 3 problems with packet captures and switch/router CLI tools * Experience with Cisco, Meraki or similar enterprise platforms 3. Physical server hardware (strong) * Hands-on experience building, servicing and upgrading rack servers from Gigabyte, Supermicro, Dell, HP and other OEMs * Comfort with IPMI/iDRAC/iLO, firmware and BIOS updates, RAID/HBA and NVMe configuration, and hardware diagnostics * Physical data center skills: racking, cabling, power and labeling 4. Edge security (solid) * Experience administering Cisco ASA, Netgate/pfSense and FortiGate firewalls * Firewall policy, NAT, site-to-site and remote-access VPN, and network segmentation General requirements * Clear written and verbal communication, including documenting your own work * Ability to work independently in a fast-moving, multi-client environment * Willingness to travel and to join an on-call rotation, * Experience with Proxmox Backup Server at scale, including pool-based job design and tiered NVMe/HDD backup targets * Experience with mixed-media Ceph clusters (NVMe and HDD tiers) and cross-cluster or cross-site migration * PCIe and GPU passthrough experience, including GPU nodes for AI workloads * Windows Server, Active Directory and Group Policy administration * Automation and infrastructure-as-code experience (Ansible, Terraform, Python, REST APIs) * Monitoring experience with Zabbix, Grafana or similar tools * Prior MSP or multi-tenant hosting experience * Basic healthcare IT knowledge, such as HIPAA-aware operations and the uptime and data-integrity demands of clinical systems * Relevant certifications, such as CCNA/CCNP, Fortinet NSE, Linux (LPIC, RHCE) or Proxmox certification ## Description Virtualization and storage (primary focus) * Design, deploy, upgrade and troubleshoot Proxmox VE clusters end to end, from bare metal through corosync, HA and production cutover * Architect and operate Ceph: CRUSH maps, device classes and pools, OSD lifecycle, PG tuning, rebalancing, capacity planning and failure recovery * Manage ZFS pools, caching and tuning for local and backup storage * Plan and execute live migrations, cross-cluster moves, PCIe/GPU passthrough and version upgrades with minimal downtime * Build and maintain standard VM templates, automation and provisioning workflows Backup and disaster recovery * Operate Proxmox Backup Server, including per-customer pools, pool-based jobs, retention, verification and prune/GC schedules * Run regular restore tests and maintain documented, tested recovery procedures for each client * Monitor backup health and capacity, and resolve failures before they become data-loss events Backend stack and monitoring * Maintain the supporting stack: Active Directory and DNS, monitoring (Zabbix), Docker-based services and internal tooling * Respond to alerts, perform root cause analysis and write clear post-incident reports Hardware and data center operations * Travel to data centers for hardware installs, replacements and upgrades, including servers, drives, NICs, memory and power components * Diagnose and repair hardware faults across Gigabyte, Supermicro, Dell, HP and other OEM platforms, and manage vendor RMAs and support cases * Rack, cable, label and document new deployments to a consistent standard Networking and edge security * Configure and troubleshoot routed and switched networks using OSPF, BGP, VLANs, LACP and SD-WAN * Administer edge firewalls on Cisco ASA, Netgate/pfSense and FortiGate, including policy, VPN and segmentation changes * Work with network and security teammates on changes that touch cluster and storage traffic Documentation and collaboration * Keep runbooks, network diagrams and asset records accurate and current * Participate in change control, capacity planning and on-call rotation * Mentor junior staff and support other engineers on escalations, To our Tacoma, WA and Las Vegas, NV data centers as needed for hardware replacements, upgrades and new builds; company-paid On-call Shared rotation covering production alerts and after-hours maintenance windows Maintenance windows Some planned work is scheduled evenings and weekends to protect client uptime Physical demands Lift up to 50 lbs, work in data center environments and rack equipment Core Competencies * Ownership: you treat production as your own, from initial design to the post-incident review. * Systems thinking: you understand how a change in one layer, such as a switch config or a Ceph pool setting, affects the rest. * Calm under pressure: you troubleshoot methodically during outages and communicate clearly while doing it. * Independence: you can take a project from requirements to production with minimal supervision. * Documentation discipline: you leave systems better documented than you found them. * Client focus: you understand that behind every VM is a business, and in healthcare, often a patient. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [10M Data Records Lost, Underwater Computing, and Psychedelic Fish - Matthias Geniar](https://www.wearedevelopers.com/videos/1908-10m-data-records-lost-underwater-computing-and-psychedelic-fish-matthias-geniar) - [How to Submit CFPs and Get into Public Speaking - Moran Weber](https://www.wearedevelopers.com/videos/2112-how-to-submit-cfps-and-get-into-public-speaking-moran-weber) - [Starting business without breaking the bank: Self hosted OSS productivity ecosystem](https://www.wearedevelopers.com/videos/1219-starting-business-without-breaking-the-bank-self-hosted-oss-productivity-ecosystem) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)