> Markdown version of [/jobs/ext/2219365-devops-infrastructure-operations-engineer](https://www.wearedevelopers.com/jobs/ext/2219365-devops-infrastructure-operations-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps / Infrastructure Operations Engineer - **Company:** Spectraforce - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Temporary contract - **Skills:** Amazon Web Services, CentOS, Cyber Security, Continuous Integration, DevOps, Disaster Recovery, Domain Name System (DNS), Monitoring of Systems, Identity and Access Management, Virtual Private Networks (VPN), Lightweight Directory Access Protocols (LDAP), Linux System Administration, Networking Basics, OpenShift, Red Hat Enterprise Linux, Ansible, Prometheus, Security Assertion Markup Language (SAML), User Provisioning Software, Software Vulnerability Management, Datadog, Load Balancing, System Availability, Grafana, Gitlab, Containerization, Gitlab-ci, Kubernetes, Deployment Automation, Firewall Services Module, Terraform, Docker - **Published:** August 25, 2026 - **Apply:** http://leoforce.us/Careers/Spectraforce/JobDetails.html?jobid=fa6544aa-66b6-4314-b111-9a589afa1cf7&OrgId=1&UserId=4680 ## About the Role * 4+ years of experience in Platform Engineering, DevOps, or production platform operations * Strong GitLab administration experience - installation, configuration, upgrades, Geo replication, backup/restore at scale * Linux systems administration (RHEL/CentOS) * Infrastructure as Code proficiency - Ansible, Terraform, and CI/CD pipelines * Monitoring and observability experience - Prometheus, Grafana, or equivalent * Containerization and orchestration - Docker/Podman and Kubernetes/OpenShift * Incident management experience - on-call, incident response, root cause analysis * Networking fundamentals - DNS, load balancing, VPN, firewall rules * Strong documentation skills, * GitLab Geo replication operations and troubleshooting * Enterprise compliance frameworks (SOC2, ISO 27001, or equivalent; Red Hat ESS/PIA/SIA a plus) * IAM integration (SSO/SAML, LDAP) * High-availability architecture design and operations * Red Hat or IBM enterprise environment experience * Datadog monitoring platform * AWS infrastructure operations ## Description Client is seeking a Senior Platform Engineer to join the operations team for gitlab.cee - Client's self-managed GitLab instance. This is a C1 (Mission-Critical) service serving ~10,000+ engineers with high availability architecture across multiple AWS regions. The platform is mature and operational. Your primary focus will be maintaining reliability, performing upgrades, managing compliance, and improving automation. You will provide US timezone coverage alongside existing team members, ensuring round-the-clock operational resilience for this critical platform. What You'll Do * Operate and maintain production and pre-production GitLab environments * Perform GitLab version upgrades through the Stage-to-Production pipeline * Execute system patching, vulnerability remediation, and compliance tasks * Manage GitLab Shared Runner infrastructure * Manage GitLab Geo replication across primary and secondary sites * Conduct and maintain disaster recovery exercises and documentation * Automate secret rotation via Ansible Automation Platform (AAP) * Maintain and improve Infrastructure as Code (Ansible/Terraform + GitLab CI) * Handle SNOW tickets - access requests, pipeline issues, configuration changes * Monitor service health using Prometheus, Grafana, and Datadog * Participate in on-call rotation with peer engineers * Create and maintain runbooks, documentation, and post-incident reviews, Key Stakeholders * ALM/DEP - Platform ownership, priority alignment * InfoSec - SOC monitoring, incident response, vulnerability remediation * IT-IAM - User provisioning, SSO integration * Engineering teams - Thousands of users relying on platform availability * PCO/Ops - Infrastructure, networking, AWS account management Platform State * GitLab 10k reference architecture with high availability * Geo replication across multiple AWS regions * Automated deployment via Ansible/Terraform + GitLab CI * Monitoring: Prometheus + Grafana + Datadog * C1 Mission-Critical service This Role Is NOT * A build-from-scratch project - the infrastructure is mature and well-documented * A pure development role - this is infrastructure operations * A solo position - you join an existing team of engineers * A user support role - you manage the platform, not individual project workflows ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)