Site Reliability Engineer

Fpt Software
Coslada, Spain
2 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
Chinese

Tech stack

Batch Processing Business Software Cloud Computing Databases Continuous Delivery Continuous Integration Linux Middleware Monitoring of Systems Python (Programming Language) Network Protocols Nginx
+11 more
Queueing Systems RabbitMQ Reliability Engineering Prometheus TCP/IP Virtual Machines Scripting Delivery Pipeline Grafana Kubernetes Apache Kafka

Job description

Responsibilities and DutiesCompruebe que cumple con los requisitos de habilidades para este puesto, así como con la experiencia asociada, y luego envíe su CV a continuación.Own end-to-end operation, maintenance support and architecture governance of overseas public cloud infrastructures, covering core cloud resources including containers, cloud virtual machines, storage and networks.Also oversee the operational stability and architectural standardization of overseas databases and middleware systems.Design, implement and maintain automated CI/CD pipelines to support continuous integration, continuous delivery and standardized release workflows for overseas business applications.Undertake daily on-call rotation responsibilities.Proactively troubleshoot and resolve functional defects, resource bottlenecks and performance anomalies of cloud infrastructure and business applications, and drive continuous optimization of overall resource performance and service stability.Manage daily operation, maintenance and production release of overseas containerized and cloudnative applications, ensuring reliable and smooth online iteration of business services.Collaborate closely with domestic technical teams to formulate and implement unified containerized application management specifications across the platform, and steadily promote standardized governance, architecture optimization and business migration initiatives for overseas applications.Continuously analyze the usage status of overseas cloud and application resources, drive resource scheduling optimization and efficiency improvement, and maximize overall resource utilization and costeffectiveness of overseas business environments.We are looking for you whoSolid mastery of Linux operating system principles and core network protocols, including TCP/IP and HTTP, with proficient hands-on operational capabilities.In-depth understanding of the Kubernetes ecosystem and the working mechanisms of core components; proficient in operating and maintaining Kubernetes Operators in production environments.Proficient in the principles and practical usage of mainstream observability tools including Prometheus and Grafana, capable of building and maintaining complete monitoring systems for cloud-native environments.Familiar with mainstream CI/CD toolchains represented by ArgoCD, with solid capabilities to build, configure and maintain automated continuous integration and delivery pipelines.Familiar with the architecture, operational principles, deployment and daily O&M of common cloudnative gateway systems including Nginx, APISIX and Envoy, as well as mainstream message queue middleware such as RocketMQ, RabbitMQ and Kafka.Proficient in at least one mainstream scripting language (Python / Shell) to support daily automation operation, batch processing and operational tool development.Have a clear understanding of overseas data compliance specifications and privacy protection regulations such as GDPR, able to carry out cloud operation and maintenance work in compliance with regional regulatory requirements.Given that these positions will collaborate closely with engineering and operations teams located in China, strong Chinese and English communication skills are highly desirable.To facilitate efficient communication, collaboration, and alignment with China-based stakeholders, Chinese-speaking candidates with the legal right to work in Spain are strongly preferred.xqysrnh However, qualified international candidates meeting the required competencies will also be considered.

Requirements

In-depth understanding of the Kubernetes ecosystem and the working mechanisms of core components; proficient in operating and maintaining Kubernetes Operators in production environments. Proficient in the principles and practical usage of mainstream observability tools including Prometheus and Grafana, capable of building and maintaining complete monitoring systems for cloud-native environments. Familiar with mainstream CI/CD toolchains represented by ArgoCD, with solid capabilities to build, configure and maintain automated continuous integration and delivery pipelines. Familiar with the architecture, operational principles, deployment and daily O&M of common cloudnative gateway systems including Nginx, APISIX and Envoy, as well as mainstream message queue middleware such as RocketMQ, RabbitMQ and Kafka. Proficient in at least one mainstream scripting language (Python / Shell) to support daily automation operation, batch processing and operational tool development. Have a clear understanding of overseas data compliance specifications and privacy protection regulations such as GDPR, able to carry out cloud operation and maintenance work in compliance with regional regulatory requirements. Given that these positions will collaborate closely with engineering and operations teams located in China, strong Chinese and English communication skills are highly desirable. To facilitate efficient communication, collaboration, and alignment with China-based stakeholders, Chinese-speaking candidates with the legal right to work in Spain are strongly preferred. xqysrnh However, qualified international candidates meeting the required competencies will also be considered.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

7:28 min

Constructing a new Docker layer from scratch

Oliver Seitz Oliver Seitz · World Congress 2026 Europe

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:50 min

Queues in TCP stacks and continuous network connections

Clemens Vasters Clemens Vasters · World Congress 2022

5:02 min

Manual port forwarding configuration using network address translation

Oliver Seitz Oliver Seitz · World Congress 2025

Videos

See all

Related articles

See all