> Markdown version of [/jobs/ext/3044747-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3044747-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Reliability Engineer - **Company:** Robert Half - **Location:** Avon, MN, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Performance Management, Distributed Systems, Monitoring of Systems, Supervisory Control and Data Acquisition (SCADA), Information Technology Operations, Reliability Engineering, Cloud Services, Ansible, Scripting, Cloud Platform System, Reliability of Systems, Infrastructure Automation Frameworks, Terraform, Splunk, Dynatrace - **Published:** September 24, 2026 - **Apply:** https://www.juju.com/job/14_288a3e2f ## About the Role We are looking for a Senior Reliability Engineer to strengthen the stability, scalability, and performance of technology systems that support construction and field operations. This position plays a key role in building dependable infrastructure across remote job sites, field offices, and cloud environments so teams can work efficiently with minimal disruption. The ideal candidate brings a strong SRE mindset, combines technical depth with practical problem-solving, and partners effectively with operational teams to maintain business continuity and system resilience., * Relevant experience in site reliability, infrastructure engineering, or a closely related IT operations role. * Strong understanding of cloud platforms, enterprise networking, and edge computing concepts in distributed environments. * Hands-on experience with Infrastructure as Code, including tools such as Terraform and Ansible. * Proficiency with monitoring and observability platforms such as Splunk and Dynatrace, along with scripting or automation capabilities. * Familiarity with construction, industrial, or field-based technology environments is strongly preferred. * Knowledge of systems used in operational technology settings, including SCADA, IoT-connected devices, or similar enterprise platforms. * Excellent written and verbal communication skills with the ability to collaborate effectively across technical and operational teams. * Ability to handle sensitive information with discretion while working independently and contributing positively in a team setting. ## Description * Build and enhance highly available infrastructure that supports office locations, remote field environments, networking needs, cloud services, and edge-based systems. * Direct incident response efforts for service disruptions, coordinate restoration activities, and lead root cause investigations to prevent repeat issues. * Create and maintain monitoring, alerting, and observability capabilities that improve visibility into system health, uptime, and application performance. * Work closely with construction, engineering, and field personnel to ensure technology reliability aligns with project schedules, operational demands, and safety expectations. * Implement automated infrastructure deployment and recovery processes using Infrastructure as Code and configuration management tools such as Terraform and Ansible. * Establish service reliability targets, manage service level objectives, and use error budgets to guide operational decisions and continuous improvement. * Strengthen the security posture of remote and field-deployed systems by improving hardening practices and secure access methods. * Provide guidance to less experienced engineers and help foster a culture centered on reliability, accountability, and operational excellence. * Identify process improvements that reduce inefficiencies, simplify support efforts, and improve overall service delivery. ## Related Videos - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Eclipse Che for Infrastructure Automation](https://www.wearedevelopers.com/videos/1611-eclipse-che-for-infrastructure-automation) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Mastering Remote Work: Tips for Developers](https://www.wearedevelopers.com/magazine/558-mastering-remote-work-tips-for-developers)