Site Reliability Engineer

ElevaIT Solutions LLC
Charlotte, NC, United States
10 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Amazon Web Services Application Configuration Access Protocols Systems Engineering Microsoft Azure Common ISDN Application Programming Interface (CAPI) DevOps Linux System Administration Performance Tuning Scrum Methodology Reliability Engineering
+9 more
Software Engineering Data Logging Scripting Google Cloud Cloud Platform System High Performance Computing Kubernetes Information Technology Rancher

Job description

This is a global materials science leader with a 170 plus year history, operating dozens of manufacturing and R&D sites worldwide. This role sits within a research and development group supporting advanced computing infrastructure behind ongoing materials science innovation.

Top 3 Skills

  • Kubernetes cluster operations and management, including provisioning, upgrades, and troubleshooting across on premises and cloud environments
  • Rancher for Kubernetes platform management
  • Linux systems administration, including performance tuning, scripting, and networking

What You’ll Do

  • Maintain and enhance Kubernetes platforms across on premises and cloud environments
  • Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher
  • Provide deep Linux systems administration support, including performance tuning, troubleshooting, and automation
  • Develop and maintain infrastructure as code solutions to standardize and automate platform deployment
  • Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration
  • Collaborate with developers, scientists, and infrastructure teams to deliver reliable platform services
  • Identify opportunities to improve platform resilience, observability, security, and maintainability

Requirements

  • 5 plus years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering
  • Hands on experience operating Kubernetes platforms in production environments, both on premises and cloud based
  • Experience with Rancher for Kubernetes cluster management
  • Strong Linux systems administration skills, including troubleshooting, scripting, and system performance analysis
  • Experience implementing infrastructure as code solutions for platform provisioning and lifecycle management

Preferred/Bonus

  • Bachelor’s degree in Computer Science, Software Engineering, Information Technology, or related field
  • Experience with Cluster API (CAPI)
  • Experience with hybrid infrastructure spanning on premises and public cloud (AWS, Azure, Google Cloud Platform)
  • Familiarity with Kubernetes observability, logging, monitoring, and alerting tooling
  • Experience supporting scientific research, high performance computing, or computational science environments
  • Experience with Agile teams (Scrum, Kanban)

About the company

Elevait Solutions was founded by veterans who believe that how you treat people is the only thing that actually matters in this industry. We’re a team that stays in your corner before, during, and after placement. We show up for the communities we work in, we tell you the truth, and we work hard to make sure every placement is a good fit for both sides. If that sounds like the kind of team you want behind you, we’d like to talk.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

3:47 min

Comparing declarative GitOps tooling alternatives and interface priorities

Davide Imola Davide Imola · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:53 min

Evaluating traditional scripting languages for modern development tasks

Jens Knipper Jens Knipper · Europe 2026 Virtual

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

Videos

See all

Related articles

See all