Site Reliability Engineer

GENESIS GROUPE
Paris, France
19 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Distributed Systems Fault Tolerance Software Engineering Cloud Platform System Kubernetes

Job description

  • Take on ambiguous reliability, scalability, and efficiency challenges and drive solutions across SRE and development teams.
  • Build and run large-scale, massively distributed, fault-tolerant systems that keep Genesis platform reliable and performant for our customers.
  • Optimize existing systems, build infrastructure, and eliminate toil through automation to continuously improve uptime and rate of change.
  • Cultivate a culture of reliability throughout the organization, guiding technical decisions that balance system health with fast-moving product priorities.
  • Ensure the long-term health, maintainability, and reliability of services through capacity planning, performance analysis, and proactive incident prevention.

Requirements

  • Strong software engineering skills (e.g., in Python, Go, or similar) with extensive experience designing, analyzing, and troubleshooting distributed systems.
  • Deep expertise with cloud computing platforms (e.g., Kubernetes, Cloud Functions) and Non-Abstract Large Systems Design (NALSD).
  • Experience leading complex, large-scale technical projects and providing technical leadership across teams.
  • Ability to apply coding, algorithms, and complexity analysis to solve ambiguous problems at scale with minimal disruption.
  • A collaborative, intellectually curious mindset - comfortable working across a wide variety of backgrounds and bringing cross-team perspective to build robust, reusable solutions.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

1:05 min

Practical Byzantine Fault Tolerance in distributed computing systems

Jonan Scheffler · World Congress 2022

1:29 min

Overcoming challenges in AI-assisted distributed system development

Przemysław Ładyński Przemysław Ładyński · World Congress 2026 Europe

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

Videos

See all

Related articles

See all