Engineer III (Site Reliability)
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+22 more
Job description
The Site Reliability Engineer (SRE) is a critical part of our Mapfre USA On-Prem and Cloud platform strategy. In this role, you will be focused on ensuring MUSA’s development platform and processes enable our software engineers to focus more on innovation than infrastructure. This role will drive the adoption of observability best practices and develop automations for resolving recurring issues. You must be comfortable working with software engineering teams and supporting their demanding needs to ensure the security, availability, and performance of the platform. This engineer must be capable of triaging issues on the front line as well as framing strategic initiatives from leadership. Being hands-on keyboard is a must for this role with a focus on developing reliability engineering for MUSA Platforms.
Responsibilities:
-
You will set standards for the monitoring of MUSA on-prem and Cloud infrastructure and applications.
-
You will ensure the platform target SLAs are met and implement appropriate SLIs for supporting services.
-
As a key member of the Critical Incident Response team, use expert communication and troubleshooting skills to aid the team in an efficient resolution.
-
You will work with developers during service transition, evaluating reliability and operability of the applications and ensuring adequate monitoring, alerting and observability.
-
You will partner with peers within Operations & Infrastructure supporting ongoing maintenance and enhancement of the platform.
-
To be successful in this role, you must focus on setting standards for automating routine tasks and workflows in supporting Infrastructure and Engineering teams.
-
The right candidate must be capable of supporting multiple internal stakeholders with a variety of technical challenges. Excelling in this role requires the ability to analyze and discern patterns in the variety of issues that arise and propose solutions to these problems., + A site reliability engineer uses a service to monitor performance metrics and detect anomalous application behavior. If there are issues with the application, the SRE team submits a report to the software engineering team. The developers fix the reported cases and publish the updated application.
Requirements
-
6 or more years of work experience with a Bachelors Degree or 4 or more years of relevant experience with an Advanced Degree.
-
Master’s Degree in IT, CS or related field preferred and/or 5+ years relevant work experience.
-
Hands-on experience in Linux and Windows systems and good understanding of distributed computing environments.
-
Intermediate level programming and/or scripting in 3 or more of the following: Python, PowerShell, JavaScript, Terraform, Ansible, etc.
-
2+ years of experience managing CI/CD tooling such as Jenkins, Github, Bitbucket, DevOps in a large-scale environment.
-
3+ Years’ experience managing observability tooling such as Splunk, Dynatrace, etc. in a large-scale environment.
-
Advanced understanding of YAML, JSON, HTML, XML.
-
2+ years of work experience supporting relational and non-relational databases [MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scaling.
-
3 or more years leading a Platform, SRE or Production Engineering group for high availability/critical platforms/applications.
-
Experience managing a distributed platform including but not limited to deployment/release management, provisioning, capacity management, workload management.
Benefits & conditions
As a global insurance leader with a strong local presence, we offer more than a job - we provide a purpose-driven career where your growth, well-being, and impact truly matter.
Purpose & Culture: Join a company built on trust, collaboration, and inclusion. Our values guide everything we do, creating a workplace where people feel respected and empowered.
Comprehensive Benefits: Enjoy competitive health coverage, retirement plans, paid time off, flexible work options, and lifestyle perks like employee discounts.
Career Growth: Advance your skills through tuition reimbursement, leadership programs, and internal mobility opportunities. Your development is our priority.
Social Responsibility: Contribute to meaningful initiatives through Fundación Mapfre, supporting communities and sustainability worldwide.
Pay Philosophy: The typical starting salary range for this role is determined by several factors including skills, experience, education, ertifications, and location. Some roles at Mapfre are eligible for commission and/or bonus earnings, in addition to salary, calculated based upon factors set forth in the compensation plan for the role.
Salary Range
$125,000 - $160,000
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Top-Paying Tech Jobs (with Salaries)
Data Engineer Salary UK
Highest Paying Tech Companies for Developers
The Most Popular IT Jobs on the Market