Site Reliability Engineer (SRE)

Robert Half
United States
8 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Performance Management Microsoft Azure Cloud Computing DevOps Github Log Analysis Reliability Engineering Prometheus Software Engineering Datadog Data Logging Cloud Platform System
+7 more
Cloud Monitoring Grafana Kubernetes Infrastructure Automation Frameworks Information Technology Terraform Dynatrace

Job description

We are looking for an experienced Site Reliability Engineer (SRE) to strengthen observability and operational resilience across a Microsoft Azure environment. This long-term Contract role will work closely with DevOps and engineering teams to establish monitoring standards, expand telemetry coverage, and improve service reliability across cloud-based platforms. The ideal candidate brings deep expertise in Azure operations, modern observability tooling, and production support, with the ability to turn data into actionable insight for faster troubleshooting and stronger system performance.

Responsibilities:

Create and advance an observability framework for Azure-hosted systems and integrated third-party platforms, ensuring scalable monitoring coverage.

Develop meaningful dashboards, alerting rules, log analysis views, and distributed tracing to provide actionable insight into application and infrastructure behavior.

Utilize Azure services such as Azure Monitor, Log Analytics, Application Insights, Managed Prometheus, and Azure Managed Grafana to expand end-to-end visibility.

Work alongside DevOps and software engineering teams to strengthen platform stability, incident response readiness, and service performance.

Assess existing monitoring practices to uncover blind spots, reduce unnecessary alert volume, and support quicker root-cause identification.

Improve insight into the health of applications, infrastructure components, and dependent services across production environments.

Support reliability-focused engineering efforts by applying SRE principles such as service measurement, alert strategy refinement, and operational readiness improvements.

Requirements

At least 5 years of hands-on experience working with Microsoft Azure in production environments. Strong practical knowledge of observability platforms such as Datadog, Dynatrace, or comparable monitoring solutions. Demonstrated experience with Azure Monitor, Log Analytics, Application Insights, Prometheus, and Grafana. Background using GitHub Actions, Terraform, Kubernetes, and Infrastructure as Code practices. Solid understanding of metrics, logging, tracing, alerting, and reliability engineering for cloud-based systems. Ability to independently lead observability initiatives from technical design through implementation and optimization. Experience supporting live production systems, troubleshooting operational issues, and improving incident response processes. Bachelor’s degree in Computer Science, Information Technology, or a related discipline., All applicants applying for U.S. job openings must be legally authorized to work in the United States. Benefits are available to contract/temporary professionals, including medical, vision, dental, and life and disability insurance. Hired contract/temporary professionals are also eligible to enroll in our company 401(k) plan. Visit roberthalf.gobenefits.net for more information.

Benefits & conditions

Robert Half works to put you in the best position to succeed. We provide access to top jobs, competitive compensation and benefits, and free online training. Stay on top of every opportunity - whenever you choose - even on the go. Download the Robert Half app and get 1-tap apply, notifications of AI-matched jobs, and much more.

About the company

Robert Half is the world’s first and largest specialized talent solutions firm that connects highly qualified job seekers to opportunities at great companies. We offer contract, temporary and permanent placement solutions for finance and accounting, technology, marketing and creative, legal, and administrative and customer support roles.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

2:33 min

Advocating for SRE practices within agency environments

Martin Beránek · LIVE

Videos

See all

Related articles

See all