Site Reliability Engineer (SRE)

Merican Inc
Atlanta, GA, United States
25 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Automation of Tests Cloud Computing Configuration Management Continuous Integration DevOps Github Gradle Monitoring of Systems Hardware Security Module Identity and Access Management
+26 more
Python (Programming Language) Linux System Administration Apache Maven Cisco Nexus Switches OpenShift Performance Tuning Reliability Engineering Cloud Services Ansible Shell Script Software Deployment Software Engineering Data Streaming VersionOne Virtualization Technology Data Processing Enterprise Software Applications Load Balancing Firewalls (Computer Science) Gitlab Cloudformation Information Technology Sumo Logic (Software) Terraform Jenkins Servicenow

Job description

We are seeking a skilled Site Reliability Engineer (SRE) to support highly available, business-critical applications across on-premises and AWS cloud environments. The ideal candidate will have strong expertise in DevOps, cloud infrastructure, automation, CI/CD, monitoring, and troubleshooting complex production systems. This role offers the opportunity to work with modern cloud technologies and contribute to the reliability, scalability, and performance of enterprise applications., * Manage and optimize data streaming and API components in OpenShift (On-Premises) and AWS.

  • Review application APIs and processes to identify performance optimization opportunities.
  • Automate testing, including data quality validation, production deployments, and release processes.
  • Develop integrations between on-premises, AWS, and third-party tools such as ServiceNow, VersionOne, and Sumo Logic.
  • Collaborate with teams to define and implement SLIs and SLOs.
  • Monitor production environments, troubleshoot performance issues, conduct root cause analysis, and document findings.
  • Design, build, and maintain CI/CD pipelines for application artifacts, APIs, and data processing jobs.
  • Configure monitoring, alerting, and observability solutions to enable proactive issue detection.
  • Implement AWS security best practices, including IAM, HSM, encryption, and access controls.
  • Monitor cloud costs, generate usage reports, and recommend cost optimization strategies.
  • Design and implement solutions to address security vulnerabilities and compliance requirements.
  • Analyze infrastructure capacity and performance to support scalable and resilient systems.
  • Develop backup and disaster recovery strategies for critical applications and data.
  • Collaborate with architecture, infrastructure, and application teams to continuously improve system performance, reliability, and security.

Requirements

  • Strong experience with AWS cloud services and cloud operations.
  • Hands-on experience with OpenShift, CloudFormation, Terraform, Ansible, Shell scripting, and Python.
  • Experience with Linux administration and enterprise infrastructure.
  • Knowledge of virtualization, networking, load balancers, firewalls, storage, backup, and monitoring tools.
  • Experience with CI/CD tools such as GitLab, GitHub, Jenkins, Maven, Gradle, and Nexus.
  • Experience with Software Release Management.
  • Strong troubleshooting and incident management skills for mission-critical systems.
  • Experience with automation, infrastructure orchestration, and configuration management., * Bachelor’s degree in Computer Science or a related technical field (or equivalent experience).
  • 3+ years of DevOps/SysOps engineering experience with a focus on AWS.
  • 2+ years of application development experience involving data streaming and high-availability applications.
  • 1+ year of experience in a Site Reliability Engineering (SRE) environment preferred.
  • Overall 4 6 years of IT experience.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

4:54 min

Implementing geographic salary tiers for compensation equity and fairness

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

Videos

See all

Related articles

See all