Site Reliability Systems Engineer

Everforth Apex
St. Louis, United States
6 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) IBM AIX Apache Tomcat Systems Engineering Bash Shell Oracle WebLogic Server C++ (Programming Language) Information Systems System Configuration Linux Middleware Perl (Programming Language)
+19 more
Monitoring of Systems Integrated Development Environments Networking Hardware Mobile Application Software Python (Programming Language) Windows PowerShell Reliability Engineering Software Engineering PL-SQL System Testing Web Platforms Web Services Scripting Reliability of Systems Information Technology Operational Systems Api Design Splunk Dynatrace

Job description

The Rental Operations Products team is responsible for the development and support of mission-critical rental ticketing and operational systems. This role is responsible for implementing enhancements, monitoring system health, conducting capacity planning, resolving production issues, and designing solutions that support an API-first architecture strategy.

The team partners closely with multiple technology groups, including mobile applications, kiosks, web platforms, and customer experience solutions, to ensure consistent, reliable service delivery. This is an opportunity to contribute to the evolution of highly available systems and the implementation of new capabilities that support business growth and operational excellence.

The Senior Systems Engineer (Engineer II) is responsible for ensuring the availability, performance, efficiency, monitoring, capacity planning, change management, and incident response of highly available production systems. These systems consist of a hybrid environment spanning cloud platforms, web services, and legacy technologies.

The ideal candidate will have experience with Site Reliability Engineering (SRE) principles and a strong understanding of both software engineering and infrastructure operations. This role serves as a bridge between development and operations teams by applying engineering practices to system administration and operational support.

As a Senior Systems Engineer, you will be expected to proactively monitor solution health, ensure systems meet established service-level objectives, and assist in resolving performance issues, capacity concerns, and outage events. You will also be responsible for supporting maintenance activities such as upgrades and patching, while developing a comprehensive understanding of the overall solution ecosystem to drive reliability and continuous improvement.

This role requires a technical leader who can serve as a subject matter expert, represent the team on complex initiatives, evaluate technology effectiveness, and recommend enhancements that improve system reliability, scalability, and consistency. Responsibilities

  • Contribute to strategic capacity planning initiatives.
  • Focus on production infrastructure support and operational improvement efforts.
  • Monitor key performance metrics and proactively address issues.
  • Serve as a subject matter expert in multiple technical areas.
  • Provide technical guidance and leadership for projects and initiatives.
  • Support large-scale and complex assignments.
  • Operate with a high degree of autonomy while collaborating across teams.
  • Define, develop, communicate, and implement standards, processes, and procedures.
  • Build and maintain strong relationships with stakeholders across the organization.
  • Partner with architects to improve solution performance, scalability, reliability, and quality.
  • Create and maintain technical documentation.
  • Mentor and support less experienced team members.
  • Explore emerging technologies and contribute innovative solutions.
  • Participate in troubleshooting, root cause analysis, and incident resolution activities.
  • Support system health during maintenance windows, upgrades, and deployments.

Requirements

This position offers the opportunity to work fully remote within the United States (excluding Alaska and Hawaii). Candidates must be able to work within U.S. Central Time core business hours. Periodic travel to company locations for meetings, team events, or business needs may be required a few times per year., * Authorized to work in the United States without current or future sponsorship requirements.

  • Must reside within the United States (excluding Alaska and Hawaii).
  • Ability to work within U.S. Central Time core business hours.
  • 3+ years of experience supporting software development environments utilizing technologies such as Java, Web Services, C/C++, and PL/SQL.
  • 2+ years of experience configuring and supporting middleware technologies including Tuxedo, WebLogic, and Tomcat.
  • 2+ years of experience administering AIX and/or Linux systems.
  • 2+ years of experience developing scripts for automation, system administration, or application support using tools such as Shell/Bash, Python, Perl, PowerShell, or similar technologies.
  • 1+ year of experience with monitoring and observability platforms such as Splunk, Dynatrace, or comparable tools.
  • Strong commitment to incorporating security best practices into daily responsibilities and technical decisions., * Bachelor’s degree in Computer Science, Information Systems, Management Information Systems, or a related field.
  • General knowledge of network engineering concepts.
  • Experience with infrastructure and system testing methodologies.
  • Strong problem-solving and troubleshooting skills.
  • Familiarity with capacity planning techniques and methodologies.
  • Understanding of hardware and infrastructure concepts, including networking devices, resiliency patterns, and system topology.
  • Working knowledge of SQL.

About the company

Everforth Apex is a world-class IT services company that serves thousands of clients across the globe. When you join Everforth Apex, you become part of a team that values innovation, collaboration, and continuous learning. We offer quality career resources, training, certifications, development opportunities, and a comprehensive benefits package. Our commitment to excellence is reflected in many awards, including ClearlyRateds Best of Staffing in Talent Satisfaction in the United States and Great Place to Work in the United Kingdom and Mexico.

Everforth Apex uses a virtual recruiter as part of the application process. Click for more details. By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Everforth Apex and its affiliates, and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy at

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all