Site Reliability Engineer

ReqRoute View all jobs
Arizona City, AZ, United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$141,440.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Artificial Intelligence Application Performance Management Confluence BigQuery Cloud Computing Databases Continuous Integration Linux Distributed Systems Domain Name System (DNS) Github
+31 more
Monitoring of Systems Hypertext Transfer Protocols (HTTP) Python (Programming Language) PostgreSQL Microsoft SQL Server MongoDB Network Protocols Oracle Databases Redis Reliability Engineering Ansible Prometheus TCP/IP Rust (Programming Language) Load Balancing Istio System Availability Grafana Containerization Kubernetes Rancher Hashicorp Graphql Vertica Api Gateway Terraform Splunk Appdynamics Dynatrace Golang Programming Languages

Job description

Job title: Site Reliability Engineer Bill rate: $68/hr Role is onsite based out of Scottsdale, AZ Years of experience required: 7+ Please share the profiles only if the associ…

  • 1 day ago
  • Apply easily +

Requirements

Years of experience required: 7+ Please share the profiles only if the associate agrees to following requirements.

  1. 1 level of Internal evaluation
  2. 3 Levels of Client Interviews (2 Telephonic and 1 In person). Last round in person interview can be either on Richardson, Texas or Scottsdale, Arizona. Travel cost for in person interview will not be reimbursed Required Skills Service reliability/operation experience running large-scale, high-performance applications in a hybrid environment (on-prem and cloud). Experience in writing automation scripts and building dashboards for Application Performance management to manage Transaction journeys. Experience working with Programming languages such as Go, Python, Java, Rust etc. Working knowledge on with one or more databases- Oracle, SQL Server, Redis, Clickhouse, postgres, Mongo or any time-series databases Experience in transitioning platforms to the cloud and Containerization GCPand Rancher Experience maintaining containerized app in GKE/RKE/AKE environments. Experience Implementing Cloud observability using OTEL to enable real-time monitoring, distributed tracing and incident resolution. Experience working with specific GraphQL Framework (Apollo, Prisma, Hasura etc.). Experience using knowledge of networking protocols such as TCP/IP, HTTP, DNS, Load balancing and service mesh to troubleshoot issues in high pressure situations. Preferred Skills: Proven experience managing Application availability, building creative solutions to manage repetitive activities, improving gating and detect for applications at every touchpoint for a 24 x 7 High availability platform exposed to critical clients and customers. Working knowledge of Monitoring tools - Splunk, App-dynamics, grafana/Prometheus and Dynatrace. Experience with tools like Rally, Confluence and other CI/CD extenders. Hands-on experience with implementing in-memory caching solutions. Experience on Redis DB is a plus. Excellent debugging skills across variety of integrated technical platforms on API gateway. Hands-on with GCS, Cloud SQL, Spanner and Firestore. Extensive experience in Enterprise level Infrastructure and Operations. Experience in High Availability and distributed systems, Linux and Windows administration, troubleshooting and support. Monitor and troubleshoot HashiCorp Vault environments, ensuring minimal downtime and rapid recovery from incidents. Working knowledge on Vertex AI, Gen AI and Bigquery Google Cloud Platform (GCP) Containerization, Kubernetes Infrastructure as Code (Terraform), CI/CD (GitHub Actions), and Helm Automation and scripting using Python, Ansible, and Node.js Monitoring and observability with

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · WWC Europe 2026

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

7:15 min

Installing Istio programmatically with bash scripts

Thomas Südbröcker · LIVE

Videos

See all

Related articles

See all