> Markdown version of [/jobs/ext/449551-director-of-platform-reliability-engineering](https://www.wearedevelopers.com/jobs/ext/449551-director-of-platform-reliability-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Director of Platform & Reliability Engineering - **Company:** Forge Global - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $235,000.0 - $245,000.0 - **Contract:** Permanent contract - **Skills:** Cloud Computing, Cloud Engineering, Continuous Integration, Disaster Recovery, Distributed Systems, Reliability Engineering, Software Engineering, Containerization, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Cloud Migration - **Published:** June 6, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=7ce2353129098493 ## About the Role Do you have experience in Technology management?, Do you have a Bachelor's degree?, * 8+ years of software engineering experience, including significant time leading infrastructure, platform, cloud, or reliability-focused teams. * 5+ years of people leadership experience, including leading managers and building high-performing engineering organizations. * Deep experience with cloud infrastructure, infrastructure as code, observability, incident response, and modern platform engineering practices. * Strong technical judgment in distributed systems, production operations, service reliability, and scalable engineering architecture. * Experience defining engineering strategy, driving cross-functional alignment, and translating business priorities into platform and infrastructure roadmaps. * Bachelor's degree in Computer Science or a closely related field, or equivalent practical experience. * Excellent communication and stakeholder management skills, with the ability to influence technical and non-technical leaders., * Experience in FinTech, financial services, or another regulated industry. * Experience leading organizations through cloud modernization, platform standardization, or large-scale reliability transformations. * Strong familiarity with Kubernetes, container platforms, CI/CD systems, and infrastructure automation tooling. * Experience building developer platforms and self-service infrastructure capabilities that improve engineering productivity. * Experience at growth-stage companies where balancing scale, speed, and reliability is essential. * Physical requirements: operate a computer for 8 hours per day; give and receive detailed information through verbal and written communication For residents of San Francisco, CA the annual salary range for this role is $235,000 - $245,000+ annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, bonus, and other factors Upon offer, we conduct background checks that include employment and education verification, state, and county criminal history searches. ## Description The Director of Platform & Reliability Engineering will lead a critical engineering organization responsible for the systems, services, and operational practices that enable Forge to build and run secure, scalable, and highly reliable products. This leader will oversee Platform Engineering, Cloud Operations, and Site Reliability Engineering, setting the vision for how internal platforms, cloud infrastructure, developer enablement, and production operations evolve to support the company's growth., This leader will be responsible for building and scaling the capabilities, teams, and operating model needed to deliver resilient infrastructure and strong internal engineering platforms across Forge. * Lead and develop the Platform Engineering, Cloud Engineering, and Site Reliability Engineering teams, including organizational design, hiring, coaching, and performance management. * Define and execute the strategy for internal platforms, cloud infrastructure, reliability engineering, observability, and developer enablement. * Drive improvements in availability, performance, scalability, security, and operational maturity across production systems. * Establish and evolve standards for incident management, service ownership, operational readiness, disaster recovery, and post-incident learning. * Partner with product and software engineering leaders to create paved-road solutions that improve delivery speed, reliability, and developer experience. * Lead cloud capacity, cost, and architecture planning to ensure infrastructure investments align with business priorities and engineering demand. * Create and monitor meaningful service level objectives (SLOs), operational metrics, and executive-level reporting for reliability and platform health. * Guide technical architecture decisions for cloud platforms, CI/CD, infrastructure automation, and runtime environments with a focus on resilience and maintainability. * Promote a culture of accountability, continuous improvement, automation, and operational excellence across the engineering organization. * Collaborate with security, compliance, and risk partners to ensure platform and infrastructure practices meet the needs of a regulated business. ## Related Videos - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Green Cloud Computing](https://www.wearedevelopers.com/videos/592-green-cloud-computing) - [Convincing Product teams to Adopt Gitops in a Large Org](https://www.wearedevelopers.com/videos/1936-convincing-product-teams-to-adopt-gitops-in-a-large-org) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [The journey from developer to devops - what i've learnt along the way](https://www.wearedevelopers.com/videos/238-the-journey-from-developer-to-devops-what-i-ve-learnt-along-the-way) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)