Sr SRE

Insight Global
Bellevue, WA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Bash Shell Relational Databases DevOps Monitoring of Systems Python (Programming Language) Routing Performance Tuning Reliability Engineering Scripting Reliability of Systems Firewalls (Computer Science) Kubernetes
+1 more
Build Tools

Job description

An employer is looking for an SRE to join their enterprise level SRE team. They are building a specialized team of Senior Site Reliability Engineers to act as embedded technical experts across their IT organization. This team will be responsible for solving complex production issues, guiding development teams, and building tools that improve system resilience and observability.

This is not a traditional SRE role. You will be a technical leader, coach, and hands-on problem solver who thrives in ambiguity and drives results across organizational boundaries.

Responsibilities

  • Investigate and resolve high-impact production issues across infrastructure and applications.

  • Embed with dev teams to guide them through performance, reliability, and architectural challenges.

  • Participate in incident response bridges as a technical expert.

  • Build tools and scripts to detect vulnerabilities, automate checks, and improve system visibility.

  • Conduct post-incident audits and ensure follow-through on remediation.

  • Collaborate with DBAs, network engineers, and platform teams to unblock and resolve issues.

  • Proactively identify issues and drive them to resolution without waiting for direction.

Requirements

10+ years of experience in SRE or DevOps roles.

Deep expertise in Kubernetes (deployment, troubleshooting, performance tuning), Networking (firewalls, routing, connectivity issues), Relational Databases (patching, auditing, performance tuning)

Strong scripting skills (e.g., Python, Bash) for tooling and automation.

Proven ability to lead through influence and solve problems across teams.

Comfortable navigating organizational blockers and driving issues to resolution.

Experience with incident response and postmortem processes.

Familiarity with monitoring and observability tools.

Ability to mentor and coach other engineers and development teams.

Strong communication, and the ability to explain complex technical issues clearly to both technical and non-technical audiences.

Ability to work cross functionally with DBAs, network engineers, developers, and leadership.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:04 min

Enhancing network privacy with routing fees and onion routing

Andreas M Antonopoulos · LIVE

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

2:33 min

Advocating for SRE practices within agency environments

Martin Beránek · LIVE

Videos

See all

Related articles

See all