Principal Platform Engineer

Adoptioninfluence
London, UK
3 days ago
Apply on www.apply4u.co.uk
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Amazon Web Services Cloud Computing Continuous Integration Distributed Systems Fault Tolerance Productivity Software Reliability Engineering Software Deployment Cloud Platform System

Job description

Define and evolve the technical strategy, architecture, and roadmap for an AWS cloud platformDesign and build scalable, secure, resilient, and reusable platform capabilitiesDevelop self-service tooling and services to improve developer productivityDrive automation across infrastructure, application delivery, operational processes, and platform managementEstablish platform standards, engineering patterns, and best practicesImprove platform reliability, scalability, performance, observability, and operational efficiencyReduce engineering friction and accelerate software deliveryEstablish engineering principles and guardrails for security, reliability, and governanceLead complex platform initiatives from discovery and architectural design through implementation and operational adoptionInfluence engineering teams and senior stakeholders on platform strategy and technology choicesMentor and develop engineers and promote strong platform engineering practicesChampion automation, ownership

Requirements

continuous improvement, and operational excellenceRequirementsSignificant experience designing and operating internal developer platforms at scaleExperience building reusable platform services, tooling, and self-service capabilitiesExperience applying platform-as-a-product principlesExperience establishing standards and paved-road approachesExperience designing highly available, scalable, and resilient cloud-native architecturesStrong understanding of cloud infrastructure, distributed systems, networking, identity, and securityExperience making architectural trade-offs across reliability, scalability, performance, cost, and developer experienceExperience with Infrastructure as Code and automated cloud provisioningExperience with CI/CD and automated software deliveryExperience automating operational processes and platform lifecycle managementExperience establishing repeatable, standardised engineering workflowsExperience designing for resilience, fault tolerance, observability, and operational excellenceExperience applying SRE principles and practicesExperience in performance analysis, capacity planning, scalability engineering, and proactive reliability improvementExperience establishing service-level objectives, monitoring, alerting, and operational practicesExperience providing technical leadership across multiple engineering teams or a wider engineering organisationAbility to influence architecture and technical direction without relying solely on organisational authorityExperience leading complex technical initiatives and driving decisions through ambiguityExperience mentoring senior and developing engineers and creating a culture of knowledge sharing and technical excellenceCore CompetenciesDemonstrates expertise in designing and operating scalable, secure, and resilient AWS cloud platforms while applying platform-as-a-product principles. Proven ability to lead complex technical initiatives, mentor engineers, and establish best practices for operational excellence and automation.Highest-signal resume keywordsAWS Cloud Platform ArchitectureInfrastructure as CodeCI/CD AutomationPlatform-as-a-Product PrinciplesSite Reliability Engineering (SRE) PracticesATS Optimization KeywordsHard SkillsCloud InfrastructureDistributed SystemsNetworkingSecurityPerformance AnalysisCapacity PlanningScalability EngineeringOperational ExcellenceMonitoring and AlertingFault ToleranceSoft SkillsTechnical LeadershipMentoringInfluencing StakeholdersKnowledge SharingContinuous ImprovementIndustry KeywordsPlatform EngineeringCloud-Native ArchitecturesEngineering WorkflowsService-Level ObjectivesOperational PracticesTools & TechnologiesAutomated Cloud ProvisioningSelf-Service ToolingEngineering StandardsOperational ProcessesDeveloper Productivity Tools #J-18808-Ljbffr

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.apply4u.co.uk
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:49 min

Container hosting options available on Amazon Web Services

Federico Fregosi · World Congress 2022

1:05 min

Practical Byzantine Fault Tolerance in distributed computing systems

Jonan Scheffler · World Congress 2022

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · World Congress 2024

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:49 min

Enhancing system resilience and fault tolerance

Michael Eder +1 · LIVE

2:11 min

Evolving devops towards platform engineering and self-service principles

Bruno Amaro Almeida · World Congress 2022

Videos

See all

Related articles

See all