Site Reliability Engineer

Algolia
Paris, France
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
€69,768.0 - €96,900.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Cloud Computing Program Optimization Computer Programming Continuous Integration Distributed Systems Github Python (Programming Language) Platform as a Service (PAAS) Reliability Engineering Ruby
+10 more
Circleci System Availability Reliability of Systems Build Management Kubernetes Data Analytics Api Management Legacy Systems Golang Microservices

Job description

The Platform as a Service (PaaS) team is dedicated to empowering development teams by creating toolchains, guidelines, and standards. Our focus is on enabling seamless automation and CI/CD, comprehensive observability, and unwavering reliability in a secured cloud-native environment.

The Opportunity

The Senior Site Reliability Engineer (IC4) position within the Platform As a Service team presents an exciting opportunity for a seasoned professional to enhance scalable infrastructure with a focus on CI/CD, Observability, and application hosting. In this role, you will bridge the gap between our junior and senior staff, playing a critical role in ensuring the reliability, scalability, and performance of Algolia’s Search Products.

As a senior contributor, you will be responsible for building and optimizing systems that ensure the platform’s efficiency and reliability, while also mentoring junior engineers and collaborating across teams. Your work will be pivotal in improving infrastructure, enhancing observability standards, and streamlining CI/CD processes. You will play a significant role in transitioning legacy systems to a modern Kubernetes-based architecture, contributing to long-term infrastructure strategies, and ensuring alignment with business needs.

Your role will consist of:

  • CI/CD Development and Maintenance: Contribute to the design, optimization, and maintenance of the CI/CD pipelines to improve the speed, reliability, and efficiency of the development lifecycle. Assist in driving standardization across various services hosted on the platform.
  • Observability Enhancement: Lead efforts to improve the observability of critical systems, working closely with cross-functional teams to ensure actionable monitoring and alerting frameworks are in place. Help troubleshoot complex issues and optimize system reliability.
  • Kubernetes and Cloud Management: Contribute to the development and operation of our Kubernetes-based architecture. Ensure systems are resilient, scalable, and optimized for performance. Actively participate in enhancing cloud-based solutions for API management and microservices.
  • System Optimization and Scaling: Collaborate with team members to ensure system scalability, operability, and performance. Lead initiatives to optimize resource utilization, focusing on cost efficiency while maintaining high system availability.
  • Mentorship and Knowledge Sharing: Mentor mid-level engineers (IC3) by providing guidance on technical challenges and SRE best practices. Support team growth by fostering knowledge-sharing sessions and helping establish processes that drive operational excellence.
  • Cross-Team Collaboration: Work closely with product, software, and other SRE teams to ensure that platform goals align with broader business objectives. Drive initiatives aimed at enhancing platform stability, security, and scalability.

Requirements

  • Strong Programming Skills: Proficient in Golang and Python with a solid understanding of software craftsmanship. Knowledge of Ruby is a plus.
  • Experience in CI/CD Pipelines: Hands-on experience in building and maintaining CI/CD pipelines using tools like GitHub Actions, CircleCI, or alternatives. Familiarity with best practices for ensuring build and deployment reliability.
  • Observability: Experience designing and implementing monitoring, alerting, and observability frameworks that provide actionable insights. Strong troubleshooting skills in production environments.
  • Kubernetes and Cloud Infrastructure: Proven experience in managing and optimizing Kubernetes-based architectures and working with public cloud providers such as GCP, AWS, or Microsoft Azure.
  • Distributed Systems Expertise: Experience in designing, building, and operating distributed systems at scale, with a focus on reliability, availability, and performance.
  • Mentorship and Leadership: Experience mentoring junior engineers and helping them grow. Ability to collaborate with cross-functional teams and contribute to strategic initiatives.
  • Problem-Solving Skills: Ability to independently solve complex technical problems with minimal supervision while collaborating effectively with other team members.
  • Excellent Communication and Organizational Skills: Strong ability to communicate complex technical issues to both technical and non-technical audiences. Ability to organize and prioritize multiple projects., * GRIT - Problem-solving and perseverance capability in an ever-changing and growing environment.
  • TRUST - Willingness to trust our co-workers and to take ownership.
  • CANDOR - Ability to receive and give constructive feedback.

Benefits & conditions

The annual base salary compensation range for this role reflects market pay data within this location. The exact compensation offered for this role may vary depending on specific location and job-related knowledge, technical skills, and experience; and is only one part of our Total Rewards philosophy to compensate and recognize employees for their work. Base Salary Pay Range €69.768-€96.900 EUR

About the company

Algolia is the retrieval intelligence layer that turns intent into trusted, decision-grade outcomes. Powering more than 1.7 trillion queries a year for over 18,000 customers with millisecond latency and 99.999% reliability, we are the recognized leader for Search and Product Discovery by top industry analyst firms. The Algolia platform turns a company’s products, content, and business rules into data that humans, applications and AI agents can act upon. The result is trusted customer experiences with stronger conversions for measurable business impact.

Algolia is set to enable every company to create world-class Search and Discovery experiences with an API-first approach. Performance and Scalability is at the heart of our mission: we power 1.5 trillion searches a year, for 10K+ customers all over the world.

If you’re a problem solver, able to think outside the box and eager to nurture others and learn from them, then this is your challenge!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all