Lead Software Engineer, Platform Engineering

Klaviyo
Boston, MA, United States
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Cloud Engineering Databases Software Debugging Linux Distributed Systems Python (Programming Language) Networking Basics Klaviyo Email and SMS Marketing Kubernetes Api Gateway
+2 more
Terraform Golang

Job description

Experteer Overview In this role you will set the technical direction for Klaviyo’s core platform services and ensure platform primitives scale with growth while staying simple, reliable, and cost-efficient. You will guide cross-team engineering needs, shape multi-year roadmaps, and own foundational primitives that power product teams. You’ll drive improvements in availability, latency, and cost, while mentoring engineers and promoting best practices. This is a highly visible, impact-focused leadership role that shapes how thousands of engineers ship products. You’ll work with Python/Go, AWS, and Kubernetes to advance the platform and enable rapid, safe delivery for the business. Compensation / Benefits * Guide design and development of platform primitives (infrastructure, databases, queuing, observability) * Set technical vision and multi-year roadmap for core platform services * Own design, development, and evolution of foundational primitives used by product teams * Lead cross-team initiatives to improve availability, scalability, latency, and cost-efficiency * Identify architectural bottlenecks and drive solutions * Leverage Python/Go, AWS, Kubernetes to advance the platform * Champion design reviews, configuration as code, and defensive programming * Drive cost-optimization and build tooling/guardrails for cost-efficiency * Mentor senior and mid-level engineers to raise technical quality * Participate in and improve on-call practices with root-cause focus * Experiment with AI tools to improve efficiency Tasks * 10+ years of distributed systems design, build, and operation with technical leadership * Hands-on production software development in Python and/or Go * Cloud-native experience with infrastructure as code (Terraform) and Kubernetes * Strong Linux/networking fundamentals and production-debugging skills * Experience across API gateways, observability, asynchronous processing, or database/storage platforms * Proven ability to own full lifecycle of cross-team platform initiatives * Mentoring track record and ability to influence technical direction beyond own team * Excellent technical design docs and RFCs authoring and stakeholder alignment * Comfort handling outages and driving root-cause resolution Key requirements * health benefits * welfare benefits * wellbeing benefits * equity * sign-on payments * annual bonus plan

Requirements

multi-year to improve availability, scalability, latency, and cost-efficiency * Identify architectural bottlenecks and drive solutions * Leverage Python/Go, AWS, Kubernetes to advance the platform * Champion design reviews, configuration as code, and defensive programming * Drive cost-optimization and build tooling/guardrails for cost-efficiency * Mentor senior and mid-level engineers to raise technical quality * Participate in and improve on-call practices with root-cause focus * Experiment with AI tools to improve efficiency Tasks * 10+ years of distributed systems design, build, and operation with technical leadership * Hands-on production software development in Python and/or Go * Cloud-native experience with infrastructure as code (Terraform) and Kubernetes * Strong Linux/networking fundamentals and production-debugging skills * Experience across API gateways, observability, asynchronous processing, or database/storage platforms * Proven ability to own full lifecycle of cross-team platform initiatives * Mentoring track record and ability to influence technical direction beyond own team * Excellent technical design docs and RFCs authoring and stakeholder alignment * Comfort handling outages and driving root-cause resolution Key requirements * health benefits * welfare benefits * wellbeing benefits * equity * sign-on payments * annual bonus plan

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:18 min

Prioritizing communication and structural awareness over strict tool mastery

Liam Hurrel +1 · WWC 2021

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all