Sr Site Reliability Engineer

Renaissance
Bloomington, MN, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$109,500.0 - $150,550.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) JavaScript (Programming Language) .NET Framework Software as a Service Configuration Management Information Systems Computer Programming Disaster Recovery Github Python (Programming Language) Systems Development Life Cycle Reliability Engineering
+17 more
Ansible Newrelic Workflow Management Systems Datadog Data Logging Scripting Grafana Gitlab Cloudformation Kubernetes Information Technology Hashicorp Cloudwatch Terraform AWS EKS Docker Pagerduty

Job description

Renaissance is looking for an experienced Sr Site Reliability Engineer to be part of the Engineering Enablement group’s Site Reliability Team with a focus on Application and Infrastructure Availability, Reliability, Observability & Security.

We are at the crossroads of evolving our current team and looking for someone who has been involved in the SRE implementation journey at other companies. We are looking for someone who influences our SRE philosophy and practices, who is a problem solver, self-motivated, great at communication, values teamwork. You will apply your technical expertise to build and scale our highly available distributed SaaS platform used by millions of K-12 students worldwide.

In this role as a Sr Site Reliability Engineer, you will

  • Work with engineering, security & governance teams to improve observability, reliability, resiliency, auditability of our systems and minimize/prevent downtime.
  • Contribute to infrastructure-as-code using Terraform & CloudFormation.
  • Support CI/CD pipelines which ensures the prompt release of high-quality software.
  • Collaborate with cross-functional teams to resolve infrastructure issues.
  • Perform Disaster Recovery exercises on our products.
  • Explore and integrate AI tooling into the SRE workflows.
  • Be part of an on-call rotation & support off hour incidents & deployments.
  • Demonstrates strong skills in giving constructive feedback through coaching even without direct reports.

Requirements

Do you have experience in Tooling?, Do you have a Bachelor’s degree?, * 5+ years of experience focused on SRE.

  • Experience in managing & monitoring containerized cloud environments in production, preferably AWS EKS.
  • Experience with IaC, Configuration Management and Orchestration Tools like Terraform/Docker/Ansible.
  • Hands-on experience in any of the programming or scripting languages like .NET/Java, Python, Javascript etc .,
  • On Call experience & willingness to be on call during non-work hours and weekends.
  • Experience working in an agile environment.

Bonus points for:

  • BS in Information Systems or Computer Science, related field experience, or both.
  • Managing Kubernetes Clusters, EKS at Scale using Helm.
  • Setting up Gitlab & Github pipelines & workflows.
  • Experience setting up Monitoring, Logging, Alerting & Observability in tools such as NewRelic, Datadog, Grafana. CloudWatch, PagerDuty.
  • Experience w/Teleport, Hashicorp Boundary etc.,
  • Experience w/RedShift, OpenSearch/ZeroETL.
  • Experience running Disaster Recovery exercises.
  • Implementing service level objectives (SLO/SLI/SLA’s) & error budgets.
  • Experience using ClaudeCode using agentic coding, agentic SDLC, enabling/rolling-out agentic DX., Applicants must be authorized to work for any employer in the United States. We are unable to sponsor or take over sponsorship of an employment Visa at this time.

Benefits & conditions

3.33.3 out of 5 stars Bloomington, MN Remote $109,500 - $150,550 a year, Pulled from the full job description

  • Tuition reimbursement
  • Parental leave
  • 401(k)
  • Health insurance
  • 401(k) matching
  • Vision insurance
  • Dental insurance, Benefits for eligible US employees include:
  • World Class Health Benefits: Medical, Prescription, Dental, Vision, Telehealth
  • Health Savings and Flexible Spending Accounts
  • 401(k) and Roth 401(k) with company match
  • Paid Vacation and Sick Time Off
  • 12 Paid Holidays
  • Parental Leave (20 total weeks with 14 weeks paid) & Milk Stork program
  • Tuition Reimbursement
  • Life & Disability Insurance
  • Well-being and Employee Assistance Programs

About the company

When you join Renaissance®, you join a global leader in pre-K-12 education technology! Renaissance’s solutions help educators analyze, customize, and plan personalized learning paths for students, allowing time for what matters-creating energizing learning experiences in the classroom. Our fiercely passionate employees and educational partners have helped drive phenomenal student growth, with Renaissance solutions being used in over one-third of US schools and in more than 100 countries worldwide.

Every day, we are connected to our mission by exemplifying our values: trust each other, win together, strive for the best, own our actions, and grow and evolve.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

6:14 min

Structuring CI/CD pipelines with integrated security and quality checks

Christoph Ruggenthaler · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all