Senior Site Reliability Engineer

UNITED IT TECHNICAL SERVICES, INC.
Charlotte, NC, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

IBM AIX Bash Shell Unix Continuous Integration Relational Databases Linux Disaster Recovery Domain Name System (DNS) Middleware Monitoring of Systems IBM WebSphere MQ Python (Programming Language)
+30 more
PostgreSQL Enterprise Messaging Systems Microsoft SQL Server Networking Basics Oracle (Applications) Performance Tuning Systems Development Life Cycle RabbitMQ Red Hat Enterprise Linux Release Management Reliability Engineering Ansible Prometheus Message Oriented Middleware Solaris (Operating System) Scripting Transport Layer Security Enterprise Software Applications Load Balancing System Availability Grafana Gitlab-ci Infrastructure Automation Frameworks Deployment Automation Apache Kafka Puppet Splunk Appdynamics Dynatrace Jenkins

Job description

We are seeking a Senior Linux Production Support Engineer to support and enhance mission-critical enterprise applications running on Unix/Linux platforms. This role is ideal for someone who enjoys solving complex production issues, automating operational processes, improving platform reliability, and collaborating with cross-functional teams in a fast-paced environment., * Support and maintain high-availability Linux/Unix production environments.

  • Build, maintain, and improve CI/CD pipelines using Jenkins or GitLab CI/CD.
  • Automate deployment, configuration, and operational tasks using Bash, Python, and Ansible.
  • Monitor application health and troubleshoot production issues to ensure platform stability.
  • Perform incident triage, root cause analysis (RCA), and implement long-term solutions.
  • Optimize system performance, capacity, and application reliability.
  • Manage application releases, deployments, rollbacks, and change management activities.
  • Work closely with development, QA, business, and infrastructure teams throughout the SDLC.
  • Implement monitoring dashboards, alerts, and observability using tools such as Splunk or Dynatrace.
  • Support messaging platforms including IBM MQ or similar middleware technologies.
  • Create and maintain technical documentation, operational runbooks, and support procedures.
  • Participate in production support and scheduled maintenance activities as needed.

Requirements

  • 8+ years of experience supporting enterprise Linux or Unix production environments.
  • Strong experience with Red Hat Enterprise Linux (RHEL), Unix, AIX, or Solaris.
  • Hands-on experience with CI/CD tools such as Jenkins or GitLab CI/CD.
  • Strong scripting skills using Bash and Python.
  • Experience with configuration management tools such as Ansible (or Puppet/Chef).
  • Experience working with relational databases including Oracle, PostgreSQL, or SQL Server.
  • Experience supporting message-oriented middleware such as IBM MQ (Kafka or RabbitMQ is a plus).
  • Knowledge of networking fundamentals including DNS, load balancing, SSL/TLS, and certificates.
  • Experience with monitoring and observability tools such as Splunk, Dynatrace, ELK, Prometheus, Grafana, or AppDynamics.
  • Strong production support experience, including incident management, root cause analysis, and performance tuning.
  • Excellent troubleshooting, communication, and collaboration skills.

Preferred Qualifications

  • Experience supporting Core Payments or enterprise financial applications.
  • Experience working in banking or financial services environments.
  • Knowledge of ITIL processes, change management, and release management.
  • Experience with disaster recovery, high availability, and resiliency planning.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann +2 · LIVE

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all