> Markdown version of [/jobs/ext/1007318-site-reliability-engineer-job](https://www.wearedevelopers.com/jobs/ext/1007318-site-reliability-engineer-job). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer job - **Company:** Tekmetric LLC - **Location:** United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Java (Programming Language), JavaScript (Programming Language), Amazon Web Services, Bash Shell, Cloud Computing, DevOps, Disaster Recovery, Monitoring of Systems, Python (Programming Language), Performance Tuning, Reliability Engineering, Prometheus, Data Logging, Scripting, Cloud Platform System, System Availability, Grafana, Reliability of Systems, Containerization, Kubernetes, Infrastructure Automation Frameworks, Terraform, Docker, Elk Stack, Golang - **Published:** June 1, 2026 - **Apply:** https://jobs.diversity.com/career/2362312/site-reliability-engineer ## About the Role * Experience: 3+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related field, with deep knowledge of cloud environments (preferably AWS or GCP.). * Cloud Infrastructure: Hands-on experience with AWS (or similar cloud providers) and infrastructure as code (Terraform, etc.). * Automation: Strong experience in automation tools * Containerization: Expertise in working with containerized environments like Docker and orchestration tools such as Kubernetes. * Monitoring and Logging: Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack). * Scripting: Proficiency in scripting languages like Python, Bash, or similar. * CI/CD pipelines: Experience with designing and optimizing Continuous Integration and Continuous Deployment (CI/CD) pipelines. * Collaboration and Communication: Strong communication skills and ability to work cross-functionally, solving complex technical challenges in a collaborative manner. * Problem-solving mindset: Ability to troubleshoot and resolve critical issues in high-pressure environments, maintaining composure and professionalism. Bonus Points: * Experience with Infrastructure as Code tools like Terraform. * Familiarity with monitoring tools like Prometheus, Grafana, or the ELK stack. * Exposure to compliance and security best practices in cloud environments. * Experience coding in one or multiple programming languages such as Go, Java, Javascript. ## Description * Design and implement scalable infrastructure: Architect and maintain reliable, scalable, and secure cloud infrastructure that supports positive user experiences and measurable business growth. * Monitor and optimize system performance: Develop and maintain monitoring, alerting, and incident response practices to ensure system reliability and performance at scale. * Automate everything: Create automated pipelines for deployment, testing, and infrastructure management to improve speed, consistency, and reliability across the organization. * Ensure high availability and disaster recovery: Implement and manage solutions for backup, disaster recovery, and failover processes to ensure business continuity. * Security and compliance: Apply best practices in security, monitoring, and compliance, ensuring that systems meet necessary requirements and regulations. * Collaboration: Work cross-functionally with development, data, product, and QA teams to improve application reliability and scalability. * Leadership and mentorship: Provide technical leadership, mentorship, and guidance to junior DevOps team members, fostering a culture of continuous learning and improvement. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025)