Principal Site Reliability Engineer in Beverly Hills

Energy Jobline
Beverly Hills, CA, United States
23 days ago
Apply on www.energyjobline.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Cloud Computing Code Review DevOps Global Distribution Systems Performance Tuning Reliability Engineering Prometheus Web Services System Availability Grafana Reliability of Systems
+5 more
Containerization Kubernetes Infrastructure Automation Frameworks Data Analytics Terraform

Job description

KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating client.

Are you on the lookout for a unique career opportunity that offers leadership, responsibility, and the chance to make a significant impact? If you’re eager to contribute to a thriving and stable organization while maintaining your confidentiality, continue reading.

The Opportunity

An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms that deliver advanced 3D/4D spatial content to global users across AR/VR environments.

This is a high-impact role focused on ensuring system reliability at scale, requiring deep expertise in observability, multi-tenant architectures, and data-driven operational decision-making. You will play a key role in designing and maintaining infrastructure that supports high-volume streaming workloads while meeting enterprise-grade security and compliance standards.

This role partners closely with web services and platform engineering teams to implement SRE best practices, establish robust monitoring, and build infrastructure capable of supporting rapid growth and global distribution.

What You’ll Do

  • Design, configure, and maintain cloud infrastructure using infrastructure-as-code tools (e.g., Terraform), with a focus on optimizing content delivery and CDN performance
  • Develop and execute capacity planning strategies and performance optimization initiatives for large-scale streaming platforms
  • Instrument services to monitor system health, building dashboards and alerting systems that provide actionable insights into performance and user experience
  • Define and implement observability strategies, including SLI/SLO frameworks and error budget management
  • Establish escalation protocols and participate in on-call rotations to ensure 24/7 system availability
  • Lead incident response efforts and conduct post-incident reviews to drive continuous improvement
  • Implement and promote reliability engineering practices, including deployment safety, code review standards, and operational readiness
  • Mentor engineering teams on best practices for reliability, scalability, and production operations

Requirements

  • 7+ years of experience in Site Reliability Engineering, DevOps, or related roles, with a track record of improving system reliability and operational maturity
  • Strong expertise in cloud platforms and modern infrastructure environments (e.g., AWS, containerized workloads, or similar ecosystems)
  • Experience with infrastructure automation and container orchestration (e.g., Terraform, Kubernetes or equivalent technologies)
  • Deep understanding of multi-tenant architecture, security principles, and data protection practices
  • Hands-on experience with observability tools and monitoring frameworks (e.g., Prometheus, Grafana or similar)
  • Experience implementing automated compliance and governance practices (e.g., SOC 2, GDPR, ISO 27001 or similar standards)
  • Strong leadership and mentoring capabilities, with the ability to influence engineering teams and drive adoption of reliability-focused practices

About the company

About KēSTA I.T.:

Our name says it all; KēSTA I.T. (Keys-to-I.T.) AND our people are our keys to our success!

KēSTA I.T. is a premier Utah-based technical staffing and consulting services firm. We specialize in temporary and permanent placement of Software, Hardware, Network, Cloud, CRM/ERP, Data, End-User support, Web and Executive / leadership-based positions on a full time and consulting basis. If you’re interested in a role where top performance is rewarded, personal time is valued, and excellence is demanded at every level we want to talk to you today!

Where do you want to go? We’ve got the keys! ~ KēSTA I.T.

WWW.KeSTAIT.COM

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.energyjobline.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

1:07 min

Architecting the availability stack with Prometheus and Grafana

Gabriel Labachelerie · World Congress 2023

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:46 min

Introduction to the speaker and engineering background

Llywelyn Griffith-Swain · World Congress 2023

3:27 min

Defining DevOps through its historical origins and foundational texts

Sonal Patil · LIVE

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

Videos

See all

Related articles

See all