Senior Site Reliability Engineer (Spain)

Parser
Granada, Spain
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Microsoft Azure Bash Shell Cloud Computing Configuration Management Couchbase Servers Distributed Systems Fault Tolerance Java Virtual Machine (JVM) Python (Programming Language) Nagios
+25 more
NoSQL Parsing Performance Tuning Systems Development Life Cycle Reliability Engineering Prometheus Software Engineering Data Streaming Web Application Frameworks Data Logging Scripting Data Storage Technologies Cloud Monitoring Grafana Spring-boot Kubernetes Infrastructure Automation Frameworks Apache Kafka Azure AKS Terraform Splunk Docker Elk Stack Jenkins Microservices

Job description

At Parser, our vision is to become every customer’s favourite way to shop, whether they are at home or out on the move. Our core purpose is ‘Serving our customers, communities and planet a little better every day’. This means more than a transactional relationship with our customers; it means acting as a responsible and sustainable business for all stakeholders, for the communities we are part of and for the planet.The Quote team is a core team within the Parser transactional APIs, enabling baskets to be priced for all purchases at Parser, both online and through our tills. Our Engineering Enablement sub-team is at the forefront of this transformation, building the foundational tools and practices that empower our engineering squads to deliver value faster and more reliably. We operate on Microsoft Azure, leveraging a modern stack that includes Java, Kafka, and Couchbase.The OpportunityWe are seeking a highly skilled and passionateSenior Site Reliability Engineerto join our Engineering Enablement team. This is a critical role within a large, complex, and high-impact initiative focused on deconstructing our monolithic architecture, revitalising our technology stack, and embedding quality and resilience into every stage of our development lifecycle.You will play a pivotal role in shaping our future-state platform, driving operational excellence, and fostering a culture of continuous improvement.Why Join Parser?This isn’t just another SRE role; it’s an opportunity to be a true architect of change within one of the UK’s leading retailers. You will be solving challenging distributed systems problems, and significantly contribute to our strategic objective of accelerating product development and drastically reducing incidents.We are proud to have an inclusive culture at Parser where everyone truly feels able to be themselves. We not only celebrate diversity, but recognise the value and opportunity it brings. We’re committed to creating a workplace where differences are valued, and make sure that all colleagues are given the same opportunities. We’re proud to have been accredited Disability Confident Leader and are committed to providing a fully inclusive and accessible recruitment process.What You’ll DoArchitect and Implement Reliability:Design, build, and maintain highly scalable, resilient, and performant systems on Azure, focusing on our Java, Kafka, and Couchbase stack.Drive Modernisation:Work hands-on as part of the team spearheading the adoption of Micronaut, standardising application templates, and transitioning to managed cloud services.Enhance Operational Excellence:Develop and implement strategies for improving system observability (standardised logging, metrics, tracing), alerting, and on?call practices.Automate Everything:Champion automation across the software development lifecycle (SDLC), from CI/CD pipelines to infrastructure provisioning, focusing on accelerating delivery and de?risking deployments.Incident Management & Learning:Contribute to our mature, blameless post?incident review process, identifying root causes and implementing preventative measures to reduce incident hours.Tooling & Standards:Develop, maintain, and drive the adoption of shared, standardised SRE tooling and best practices across engineering teams, including containerisation (e.g., Docker, Kubernetes on Azure), infrastructure as code (e.g., Terraform), and configuration management.Mentorship & Collaboration:Provide technical leadership and mentorship to junior engineers, fostering a culture of SRE principles and operational excellence across the wider engineering organisation.Strategic Input:Contribute to the overall technical strategy and roadmap for our SRE and platform initiatives, ensuring alignment with business objectives.What You’ll BringDeep SRE Expertise:Proven experience as a Senior Site Reliability Engineer or a similar role, with a strong understanding of SRE principles (error budgets, SLOs/SLIs, toil reduction).Azure Cloud Proficiency:Extensive hands?on experience designing, deploying, and operating highly available and scalable applications on Microsoft Azure.Azure Kubernetes Service (AKS) Expertise:Mandatory extensive hands?on experience with AKS for container orchestration, including deployment, scaling, monitoring, and troubleshooting.Java Ecosystem Mastery:Expert?level proficiency with Java, including experience with modern frameworks (ideally Micronaut, Spring Boot, or similar) and JVM performance tuning.Distributed Systems Knowledge:Solid understanding and practical experience with distributed systems, microservices architecture, and associated challenges (e.g., consistency, fault tolerance).Messaging & Database Expertise:Hands?on experience with an event streaming platform (ideally Kafka) and NoSQL data storage (ideally Couchbase), including operational best practices.Automation First Mindset:Strong scripting skills (e.g., Python, Bash) and experience with Infrastructure as Code tools (e.g., Terraform, ARM templates) and CI/CD pipelines (e.g., Azure DevOps, Jenkins).Observability Tools:Experience with monitoring, logging, and alerting tools (e.g., Azure Monitor, Prometheus, Grafana, ELK Stack, Splunk).Problem?Solving Acumen:Exceptional analytical and troubleshooting skills, with a methodical approach to diagnosing and resolving complex production issues.Communication & Collaboration:Excellent communication skills, with the ability to articulate complex technical concepts to diverse audiences and collaborate effectively with cross?functional teams.Continuous Improvement:A proactive and innovative mindset, always seeking ways to improve systems, processes, and team efficiency.Bonus Points If You HaveExperience in a large?scale enterprise modernisation effort.Familiarity with consumer?driven contract testing.Experience implementing canary releases and feature flagging strategies.Our Commitment to YouWe are dedicated to creating an inclusive and supportive work environment where everyone can thrive. We offer competitive compensation, comprehensive benefits, and opportunities for continuous learning and professional growth.Ready to make an impact?Join Parser and help us build the future of retail technology.#J-*****-Ljbffr

Requirements

Proven experience as a Senior Site Reliability Engineer or a similar role, with a strong understanding of SRE principles (error budgets, SLOs/SLIs, toil reduction). Azure Cloud Proficiency: Extensive hands?on experience designing, deploying, and operating highly available and scalable applications on Microsoft Azure. Azure Kubernetes Service (AKS) Expertise: Mandatory extensive hands?on experience with AKS for container orchestration, including deployment, scaling, monitoring, and troubleshooting. Java Ecosystem Mastery: Expert?level proficiency with Java, including experience with modern frameworks (ideally Micronaut, Spring Boot, or similar) and JVM performance tuning. Distributed Systems Knowledge: Solid understanding and practical experience with distributed systems, microservices architecture, and associated challenges (e.g., consistency, fault tolerance). Messaging & Database Expertise: Hands?on experience with an event streaming platform (ideally Kafka) and NoSQL data storage (ideally Couchbase), including operational best practices. Automation First Mindset: Strong scripting skills (e.g., Python, Bash) and experience with Infrastructure as Code tools (e.g., Terraform, ARM templates) and CI/CD pipelines (e.g., Azure DevOps, Jenkins). Observability Tools: Experience with monitoring, logging, and alerting tools (e.g., Azure Monitor, Prometheus, Grafana, ELK Stack, Splunk). Problem?Solving Acumen: Exceptional analytical and troubleshooting skills, with a methodical approach to diagnosing and resolving complex production issues. Communication & Collaboration: Excellent communication skills, with the ability to articulate complex technical concepts to diverse audiences and collaborate effectively with cross?functional teams. Continuous Improvement: A proactive and innovative mindset, always seeking ways to improve systems, processes, and team efficiency. Bonus Points If You Have Experience in a large?scale enterprise modernisation effort. Familiarity with consumer?driven contract testing. Experience implementing canary releases and feature flagging strategies.

Benefits & conditions

We are dedicated to creating an inclusive and supportive work environment where everyone can thrive. We offer competitive compensation, comprehensive benefits, and opportunities for continuous learning and professional growth. Ready to make an impact? Join Parser and help us build the future of retail technology. #J-*****-Ljbffr

About the company

At Parser, our vision is to become every customer’s favourite way to shop, whether they are at home or out on the move. Our core purpose is ‘Serving our customers, communities and planet a little better every day’. This means more than a transactional relationship with our customers; it means acting as a responsible and sustainable business for all stakeholders, for the communities we are part of and for the planet.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all