> Markdown version of [/jobs/ext/1972258-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1972258-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** The Dignify Solutions LLC - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Big Data, Ubuntu (Operating System), Software as a Service, Cloud Computing, Cloud Foundry, Databases, Shard (Database Architecture), Linux, DevOps, Disaster Recovery, Distributed Systems, Amazon DynamoDB, Identity and Access Management, Internet Protocol, Performance Tuning, Query Optimization, Reliability Engineering, Cloud Services, Prometheus, Web Services, Apache Zookeeper, SUSE Linux, Data Ingestion, Cloud Monitoring, Grafana, Indexer, Amazon Virtual Private Cloud (VPC), Git, Concourse, Amazon Relational Database Service, Kubernetes, Apache Kafka, Azure AKS, Route53, Cloudwatch, Terraform, Jenkins - **Published:** August 7, 2026 - **Apply:** https://www.dice.com/job-detail/f1d269a0-ef9f-47b6-878e-af7b6f20f413 ## About the Role * 10+ years of experience in Site Reliability Engineer - OpenSearch to help ensure the highest levels of availability, performance, scalability, and Quality of Service (QoS) for mission-critical cloud services. * Deep experience in site reliability engineering, DevOps, cloud operations, automation, observability, and distributed systems, with proven hands-on expertise architecting, building, deploying, operating, and optimizing high-performance OpenSearch clusters and platforms from the ground up in production environments. * Expert with Kubernetes, including troubleshooting, operations, management, and configuration of complex Kubernetes services. * Proven hands-on expertise designing, building, deploying, supporting, and maintaining OpenSearch clusters and platforms from scratch in production environments * Strong experience with OpenSearch administration, cluster architecture, performance tuning, scaling, upgrades, and troubleshooting * Experience with index design, shard and replica strategy, cluster sizing, node management, snapshot/restore, backup, and disaster recovery * Strong understanding of distributed systems, search platforms, indexing pipelines, query optimization, and high-availability architectures * Expertise with Git * Expertise with Concourse, including setup, management, and troubleshooting of new pipelines * Expertise with Linux, specifically SUSE and Ubuntu * Expertise with Kafka, Zookeeper, and Big Data technologies * Expert in development of automation for testing, deployment, scalability, and management of cloud services * Expertise with building, implementing, and/or supporting cloud monitoring tools * Expert knowledge of cloud computing, infrastructure operations, and databases * Expert understanding of web services, networking, virtualization, and internet protocols * Ability to multitask and handle various projects, deadlines, and changing priorities * Excellent communication and prioritization skills * Expertise with security fundamentals as they pertain to SaaS multi-tenant application systems * Experience with AWS services including Route 53, EC2, S3, CloudWatch, DynamoDB, RDS, IAM, ACM, KMS, and VPC * Experience deploying and operating OpenSearch in AWS-based environments * Experience with Cloud Foundry-based environments * Experience with Jenkins, Chef, and/or Terraform * Exposure to and understanding of troubleshooting IP networks and application stacks * Experience with observability tools such as Prometheus and Grafana * Experience with log ingestion pipelines, index lifecycle management, retention strategies, and search platform security controls ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Add Location-based Searching to Site with ElasticSearch](https://www.wearedevelopers.com/videos/77-add-location-based-searching-to-site-with-elasticsearch) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)