> Markdown version of [/jobs/ext/318007-principal-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/318007-principal-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Site Reliability Engineer - **Company:** Kharidle Online Enterprise - **Location:** Zaragoza, Spain - **Salary:** €72,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Microsoft Azure, Backup Devices, Bash Shell, Cloud Computing, Computer Networks, DevOps, Disaster Recovery, Distributed Systems, Python (Programming Language), Linux System Administration, Reliability Engineering, Software Engineering, Data Logging, Scripting, Google Cloud, Reliability of Systems, Cloudformation, Infrastructure Automation Frameworks, Information Technology, Terraform, Docker, Golang, Programming Languages - **Published:** June 19, 2026 - **Apply:** https://es.indeed.com/viewjob?jk=2cf0367088cc8594 ## About the Role Do you have experience in Disaster recovery?, Do you have a Bachelor's degree?, * Bachelor's degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience. * 8+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Infrastructure Engineering. * Strong experience with cloud platforms such as AWS, Azure, or Google Cloud Platform. * Expertise in Linux system administration and networking concepts. * Experience with Infrastructure as Code tools such as Terraform or CloudFormation. * Strong knowledge of containerization and orchestration technologies, including Docker and Kubernetes. * Experience with CI/CD pipelines and automation tools. * Proficiency in scripting or programming languages such as Python, Go, Bash, or Java. * Excellent troubleshooting, communication, and leadership skills., * Experience leading large-scale distributed systems. * Professional cloud certifications. * Experience with observability platforms and advanced monitoring solutions. * Knowledge of security best practices and compliance standards. ## Description As a Principal Site Reliability Engineer, you will be responsible for ensuring the availability, scalability, performance, and security of mission-critical systems. You will provide technical leadership, establish reliability standards, and collaborate closely with engineering teams to build resilient services that support our business growth., * Lead the architecture, implementation, and optimization of highly available cloud infrastructure. * Define and maintain Service Level Objectives (SLOs), Service Level Indicators (SLIs), and reliability metrics. * Drive automation initiatives to improve operational efficiency and reduce manual intervention. * Design and maintain monitoring, logging, alerting, and incident response systems. * Lead root cause analysis and post-incident reviews to prevent recurring issues. * Collaborate with software engineering teams to improve system reliability throughout the development lifecycle. * Develop disaster recovery, backup, and business continuity strategies. * Mentor and guide engineers on reliability engineering best practices. * Evaluate and implement new technologies that improve platform stability and performance. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers)