Site Reliability Engineer III

Robert Half
Mount Laurel Township, NJ, United States
26 days ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Cloud Computing Cloud Engineering Data Structures DevOps Distributed Systems Elasticsearch Monitoring of Systems Python (Programming Language) MySQL Reliability Engineering
+17 more
Cloud Services Prometheus Azure Machine Learning Datadog Scripting Google Cloud Cloud Platform System System Availability Grafana Infrastructure as Code (IaC) Solid Principles Kubernetes Information Technology Deployment Automation Apache Kafka Terraform Docker

Job description

  • Support Site Reliability Engineering initiatives across AI/ML platform environments.
  • Deploy, maintain, and optimize cloud infrastructure across AWS and Google Cloud Platform.
  • Build, manage, and maintain Infrastructure as Code (IaC) solutions using Terraform.
  • Improve platform reliability, scalability, security, and operational efficiency.
  • Administer and support Kubernetes and Amazon EKS environments.
  • Monitor and troubleshoot system performance using observability and monitoring tools including Prometheus, Grafana, Datadog, and Elasticsearch.
  • Automate operational processes, workflows, and routine administrative tasks using Python and related tooling.
  • Support and enhance CI/CD pipelines and deployment automation.
  • Troubleshoot complex distributed systems and production issues in highly available environments.
  • Collaborate with engineering teams to improve platform performance, monitoring, and operational resiliency.
  • Work with technologies including Kubernetes, Docker, AWS, Google Cloud Platform, EKS, Terraform, Prometheus, Grafana, Datadog, Elasticsearch, MySQL, Kafka, and Python.

Requirements

  • 4-8 years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Cloud Engineering, or related disciplines.
  • Strong hands-on experience with AWS cloud services.
  • Hands-on Kubernetes administration and support experience.
  • Expertise with Terraform and Infrastructure as Code (IaC) practices.
  • Experience designing, supporting, and improving CI/CD pipelines.
  • Strong observability and monitoring experience, particularly with Prometheus.
  • Experience with Grafana, Datadog, Elasticsearch, or similar monitoring platforms.
  • Python programming and scripting proficiency.
  • Experience supporting distributed systems and large-scale, highly available production environments.
  • Knowledge of algorithms, data structures, software design principles, and system troubleshooting.
  • Experience with Docker containers and cloud-native platforms.

Preferred Qualifications:

  • Bachelor’s degree in Computer Science or a related technical discipline.
  • Experience supporting AI/ML platforms or infrastructure., All applicants applying for U.S. job openings must be legally authorized to work in the United States. Benefits are available to contract/temporary professionals, including medical, vision, dental, and life and disability insurance. Hired contract/temporary professionals are also eligible to enroll in our company 401(k) plan. Visit roberthalf.gobenefits.net for more information.

Benefits & conditions

Robert Half works to put you in the best position to succeed. We provide access to top jobs, competitive compensation and benefits, and free online training. Stay on top of every opportunity - whenever you choose - even on the go. Download the Robert Half app and get 1-tap apply, notifications of AI-matched jobs, and much more.

About the company

Robert Half is the world’s first and largest specialized talent solutions firm that connects highly qualified job seekers to opportunities at great companies. We offer contract, temporary and permanent placement solutions for finance and accounting, technology, marketing and creative, legal, and administrative and customer support roles.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

57 sec

Addressing developer job apprehension through an Indeed platform integration

Prashanth Chandrasekar Prashanth Chandrasekar · WWC 2024

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · WWC Europe 2026

Videos

See all

Related articles

See all