Staff Site Reliability Engineer, Energy Software

Tesla
Amsterdam, Netherlands
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Data Centers Linux Distributed Systems Github PostgreSQL Reliability Engineering Prometheus Scala (Programming Language) Software Systems Systems Architecture Rust (Programming Language)
+5 more
Reliability of Systems Kubernetes Influxdb Apache Kafka Terraform

Job description

Tesla is looking for a Site Reliability Engineer to build, enhance, and scale the infrastructure that underpins our Energy IoT applications. These applications provide real-time monitoring, optimization, and control for Tesla’s industry-leading energy products, including Powerwall, Megapack, Solar Roof, Supercharger, Wall Connector, Autobidder, and Virtual Power Plants.

We are a high-impact team that values curiosity, learning, mentorship, open discourse, and making disciplined decisions by weighing trade-offs. Our work supports over 50 engineers and directly affects millions of customers.

If you enjoy thinking in systems and tackling challenges related to the availability, reliability, scalability, and security of distributed software, this role is for you.

You’ll work with and deepen your expertise in Linux, Networking, Kubernetes, on-premises data centers, AWS, Terraform, Prometheus, Helm, GitHub Actions, PostgreSQL, CloudNativePG, Kafka, InfluxDB, Scala, and Rust.

Join us in accelerating the world’s transition to sustainable energy. What You’ll Do

  • Envision and implement changes that improve system reliability
  • Conduct deep investigations into new technologies and resolve unexpected issues that arise during operation
  • Provide guidance on system architecture and security best practices
  • Review, digest, and distill complex code and technical topics to ensure clarity and accessibility for all engineers
  • Provide technical leadership, foster collaboration, and drive key initiatives to completion
  • Uphold team values, including engineering excellence, curiosity, bias for action, self-awareness, inclusivity, and openness

Requirements

Do you have experience in Terraform?, * Minimum 3+ years of relevant industry experience

  • Experience in developing, scaling, and maintaining infrastructure for distributed systems, including IoT applications
  • Proficiency in many of the following: Linux, Networking, Kubernetes, on-premises data centers, AWS, Terraform, Prometheus, Helm, GitHub Actions, PostgreSQL, and Kafka
  • Strong understanding of system design principles and the challenges of ensuring availability, reliability, scalability, and security in distributed software systems
  • Effective verbal and written communication skills
  • Ability to navigate uncertainty and loosely defined problem statements
  • Strong analytical and problem-solving skills, with the ability to evaluate trade-offs and make well-reasoned decisions
  • Collaborative mindset with a willingness to learn, mentor, and engage in open discussions

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:58 min

Building engineering communities and finding technical inspiration

1:49 min

Handling metrics pipelines, geo-routing, and Kubernetes workload isolation

Josip Stuhli Josip Stuhli · WWC 2023

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:42 min

Creating model transparency with continuous operational metrics logging

Hauke Brammer · WWC 2021

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

Videos

See all

Related articles

See all