Production Support Engineer

A Place For Rover, Inc.
Barcelona, Spain
18 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
2 years minimum
Compensation
€42,840.0 - €51,927.0
Working hours
Regular working hours

Tech stack

Data Analysis JIRA Bash Shell Databases Database Queries Software Debugging Monitoring of Systems Mobile Application Software Python (Programming Language) NoSQL Reliability Engineering Prometheus
+9 more
SQL Databases Datadog Data Logging Scripting Freeform SQL Grafana Splunk Pagerduty Elk Stack

Job description

Meet the Site Reliability and Production Support TeamLa siguiente información tiene como objetivo proporcionar a los posibles candidatos una mejor comprensión de los requisitos para este puesto.The support engineer would be joining the Site Reliability (SRE) Team.The SRE team is responsible for the Rover platform’s overall performance, scalability, and reliability.We use observability tools, investigative skills, and data to help our engineering teams deliver a stable, reliable, and robust experience for our owners and sitters.The SRE team is part of the Platform Engineering division.As such, we work broadly across the company, serving many partners and requiring exceptional partner communication and relationship building.What we are looking forWe are seeking a highly motivated and technically skilled Support Engineer to join our team.This role is critical in ensuring the stability, performance, and availability of our production systems and maintaining a positive customer experience.The ideal candidate will serve as the primary technical liaison between our Customer Experience (CX) team and our Product Development teams, focusing on rapid response, in-depth troubleshooting, and effective resolution of customer?impacting issues.This is an ideal position for a technical, customer?focused individual familiar with large?scale consumer websites and mobile applications.We are looking for someone who can thrive in a collaborative environment, have a true passion for Rover’s mission and values and are passionate about building a deep and wide understanding of Rover’s systems and using those to improve our systems.Note that this role is part of a cross?regional team and may require some flexibility in working hours to ensure good alignment with the rest of the team.We follow a hybrid model (Mondays & Thursdays in our office in Poblenou, Barcelona).Your Responsibilities:Act as the primary technical point of contact when CX escalates customer?impacting issues, translating business impact into clear technical problem statements.Triage incoming incidents, assess severity and urgency, and communicate status updates across stakeholders in a clear and timely manner.Manage the incident lifecycle from initial report through resolution, communicating status updates clearly to relevant stakeholders (CX, Product, Engineering).Develop and maintain comprehensive runbooks and knowledge base articles for common issues and standard operational procedures.Troubleshoot and debug complex production issues utilizing logging platforms (e.g., Splunk, ELK stack), monitoring tools (e.g., Datadog, Prometheus, Grafana), and database query tools (SQL, NoSQL) to diagnose the root cause of problems.Perform code?level analysis when necessary to pinpoint defects or architectural weaknesses contributing to production instability.Collaborate effectively with Product Development teams to prioritize, document, and hand off confirmed bugs and large?scale systemic issues for permanent resolution.Act as the escalation point for the CX team when issues require deeper technical investigation or coordination with engineering teams.Your Qualifications and Skills:2+ years of experience in a Production Support, Application Support, Technical Operations, Site Reliability Engineering (SRE), Support Helpdesk or Engineering role focused on production system operations.Hands?on experience using monitoring and observability platforms to investigate live incidents (e.g., Splunk, Datadog, ELK).Solid experience with database systems, including the ability to write and execute complex SQL queries for data analysis and issue resolution.Experience coordinating between CX or non?technical teams and engineering, comfortable with technical and non?technical communication.Proficiency in at least one scripting language (e.g., Python, Bash) for automation and ad?hoc analysis.Bonus: Experience with incident management frameworks (e.g., PagerDuty, OpsGenie) and platforms such as Jira Service Management or Zendesk.Benefits of working at Rover:Long?term incentive plan with a company performance?based cash payoutPension planPrivate medical insurance25 days PTOMeal allowance and flexible compensation plan (transport and nursery)Gym membership€450 to cover the costs associated with the adoption of a petAnnual €150 wellness reimbursementFlexible work hours, sometimes you’ll need to be in at certain times, but on the whole, we’re pretty flexible when it comes to managing workload and timeGrab snacks, fresh fruit, in our kitchen to keep yourself goingRegular team activities, events, game nights, and moreDog?friendly officeCompensation:In the greater Barcelona area the first?year salary range is €42,840 - €51,927.Additionally, Rover offers a long?term incentive plan with a company performance?based cash payout and benefits to full?time employees.The cash compensation offered for this role will be dependent on the candidate’s experience, qualifications, skills, and abilities as demonstrated in the interview and hiring process.Rover is an equal?opportunity employer committed to promoting a diverse, inclusive, and inventive environment with the best employees.We’re driven by seeing our people succeed and grow, and we work to ensure everyone contributes to their fullest potential.xqbhyrx We consider all qualified applicants without regard to age, race, color, ancestry, national origin, religion, disability, protected veteran status, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable laws, regulations, and ordinances.#J-*****-Ljbffr

Requirements

As such, we work broadly across the company, serving many partners and requiring exceptional partner communication and relationship building.What we are looking forWe are seeking a highly motivated and technically skilled Support Engineer to join our team., Develop and maintain comprehensive runbooks and knowledge base articles for common issues and standard operational procedures.Troubleshoot and debug complex production issues utilizing logging platforms (e.g., Splunk, ELK stack), monitoring tools (e.g., Datadog, Prometheus, Grafana), and database query tools (SQL, NoSQL) to diagnose the root cause of problems.Perform code?level analysis when necessary to pinpoint defects or architectural weaknesses contributing to production instability.Collaborate effectively with Product Development teams to prioritize, document, and hand off confirmed bugs and large?scale systemic issues for permanent resolution.Act as the escalation point for the CX team when issues require deeper technical investigation or coordination with engineering teams.Your Qualifications and Skills:2+ years of experience in a Production Support, Application Support, Technical Operations, Site Reliability Engineering (SRE), Support Helpdesk or Engineering role focused on production system operations.Hands?on experience using monitoring and observability platforms to investigate live incidents (e.g., Splunk, Datadog, ELK). Solid experience with database systems, including the ability to write and execute complex SQL queries for data analysis and issue resolution.Experience coordinating between CX or non?technical teams and engineering, comfortable with technical and non?technical communication.Proficiency in at least one scripting language (e.g., Python, Bash) for automation and ad?hoc analysis.Bonus: Experience with incident management frameworks (e.g., PagerDuty, OpsGenie) and platforms such as Jira Service Management or Zendesk.Benefits of working at Rover:Long?term incentive plan with a company performance?based cash payoutPension planPrivate medical insurance25 days PTOMeal allowance and flexible compensation plan (transport and nursery)Gym membership€450 to cover the costs associated with the adoption of a petAnnual €150 wellness reimbursementFlexible work hours, sometimes you’ll need to be in at certain times, but on the whole, we’re pretty flexible when it comes to managing

Benefits & conditions

workload and timeGrab snacks, fresh fruit, in our kitchen to keep yourself goingRegular team activities, events, game nights, and moreDog?friendly officeCompensation:In the greater Barcelona area the first?year salary range is €42,840 - €51,927. Additionally, Rover offers a long?term incentive plan with a company performance?based cash payout and benefits to full?time employees.The cash compensation offered for this role will be dependent on the candidate’s experience, qualifications, skills, and abilities as demonstrated in the interview and hiring process.Rover is an equal?opportunity employer committed to promoting a diverse, inclusive, and inventive environment with the best employees. We’re driven by seeing our people succeed and grow, and we work to ensure everyone contributes to their fullest potential. xqbhyrx We consider all qualified applicants without regard to age, race, color, ancestry, national origin, religion, disability, protected veteran status, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable laws, regulations, and ordinances. #J-*****-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all