> Markdown version of [/jobs/ext/2174363-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2174363-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Restb - **Location:** Barcelona, Spain (Remote available) - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Cloud Computing, Cloud Engineering, Continuous Integration, DevOps, Distributed Systems, Github, Monitoring of Systems, Python (Programming Language), Network Security, Software Architecture, Reliability Engineering, Software Engineering, Data Logging, Scripting, Load Balancing, Cloud Platform System, Delivery Pipeline, Amazon Virtual Private Cloud (VPC), Amazon Relational Database Service, Containerization, Gitlab-ci, Cloudwatch, Docker, Jenkins - **Published:** August 22, 2026 - **Apply:** https://www.jobleads.com/es/job/e182a52bff87d78004c410e9c47d31686 ## About the Role We are looking for a Site Reliability Engineer who is passionate about technology and always looking for new ways to tackle complex issues. As our SRE, you would help maintain our production infrastructure, focusing on reliability, performance, observability, and cost-efficiency. You'll be a key contributor to ensuring our state-of-the-art AI solutions are delivered with high reliability, supporting a culture of automation and infrastructure-as-code. This role offers the opportunity to grow your skills and experience within a hands-on SRE and Platform Engineering environment. The usual day-to-day tasks include supporting cloud architecture, automating deployments (CI/CD), improving system monitoring and alerting, and troubleshooting production issues. The role will be completely in English, and CVs/resumes in other languages will not be considered., A strong candidate will ideally possess solid experience in the following areas: * Experience: 3+ years of Site Reliability or DevOps engineering experience, with hands-on experience maintaining production infrastructure and cloud environments. * AWS Proficiency: Hands-on experience deploying and managing production workloads on AWS. * Python Skills: Knowledge of Python for scripting, automation, and building SRE tools and services. * Containerization: Hands-on experience with Docker in a production environment. * CI/CD Pipelines: Experience working with automated delivery pipelines (e.g., GitHub Actions, GitLab CI, Jenkins, AWS CodePipeline). * Monitoring & Observability: Experience with monitoring stacks and centralized logging (e.g. OpenSearch). * Networking and Security: Understanding of cloud networking (VPC, security groups, load balancers) and security best practices. * Troubleshooting: Strong ability to diagnose and resolve issues across distributed systems. ## Description By joining our team, you will play an important role in supporting the design, implementation, and maintenance of our highly available, scalable, and secure production systems on AWS. If you thrive on solving operational challenges, driving automation through code, and have a passion for operational excellence, we'd love to meet you!, * Support AWS Cloud Infrastructure: Help design, implement, and maintain reliable, scalable, and cost-efficient services utilizing core AWS tools (e.g., EC2, ECS/EKS, Lambda, RDS, S3, CloudWatch). * Contribute to Operational Excellence: Build and maintain CI/CD pipelines, automating infrastructure deployment and configuration. * Improve System Observability: Help establish monitoring, logging, and alerting strategies to identify and resolve performance and reliability issues. * SRE/DevOps Collaboration: Work closely with the Software Development teams to help define and uphold Service Level Objectives and improve the service lifecycle, from design through deployment. * Technical Contribution: Participate in discussions around software architecture and infrastructure scaling, bringing ideas and technical input to the team. * Incident Response: Participate in incident response and root cause analysis (RCA), helping drive continuous improvement to minimize downtime and prevent recurrence. ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) ## Related Articles - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [The 12 Best Jobs for Software Engineers](https://www.wearedevelopers.com/magazine/401-the-12-best-jobs-for-software-engineers)