SRE (Terminal)
MLabs
London, UK
2 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source
Tech stack
Amazon Web Services
Cloud Computing
Cloud Engineering
Digital Assets
Network Topologies
Identity and Access Management
Python (Programming Language)
PostgreSQL
Data Streaming
Systems Integration
Kubernetes
Apache Kafka
+2 more
Web3.js
Terraform
Job description
- Own Foundation & Architecture: Design, scale, and maintain highly available, multi-region, or active-active cloud infrastructure patterns.
- Incident Response & Reliability: Lead critical incident response efforts, participate in real on-call rotations, and drive comprehensive, blameless post-mortems to continuously harden the system.
- Automation & Tooling: Write clean, production-grade automation code (Python, Go, or similar) for infrastructure tooling, operators, and seamless systems integration.
- Risk & Security Management: Exercise sharp judgment regarding system risks, balancing rapid deployment velocity with robust infrastructure safety and stability.
- Operational Excellence: Raise the engineering and operational bar across the organization through the implementation of rigorous standards, modern tooling, and technical mentorship.
Requirements
Do you have experience in Terraform?, * Core SRE & Infrastructure Focus: Deep expertise in infrastructure-as-code (Terraform/OpenTofu), network topology, high-availability architecture, and system internals.
- Proven Track Record: Experience building foundational infrastructure (ideally from 01) and running high-availability environments where reliability is treated with financial-system levels of seriousness.
- Cloud-Native Fluency: Advanced proficiency with modern cloud providers (AWS, GCP) and container orchestration platforms (Kubernetes).
- Pragmatic Problem Solver: Strong capacity to operate independently in high-stakes environments, deciding when to gather consensus versus when to execute autonomously., * Security & Compliance: Experience with infrastructure security hardening, IAM architecture, or compliance mapping (e.g., SOC2, ISO).
- Data & Streaming Infrastructure: Hands-on experience managing and scaling high-throughput, low-latency data backbones and event streaming systems (Kafka, Redpanda, PostgreSQL).
- Digital Asset Exposure: A working understanding of Web3/crypto infrastructure patterns and comfort operating within them.
Benefits & conditions
- Competitive Base Salary
About the company
Our client is a high-growth software development organization and a key contributor to one of the largest and fastest-growing decentralized crypto social networks globally. The platform has achieved massive scale, generating significant revenue and global attention since its inception.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
CH
Chris Heilmann
Dev Digest 138 - Are you secure about this?
almost 2 years ago
CH
Chris Heilmann
Dev Digest 121 - AI goes offline
about 2 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago