> Markdown version of [/jobs/ext/1400612-principal-cloud-kubernetes-engineer](https://www.wearedevelopers.com/jobs/ext/1400612-principal-cloud-kubernetes-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Cloud Kubernetes Engineer - **Company:** InterSystems - **Location:** Boston, MA, United States - **Experience:** Expert - **Salary:** $135,000.0 - $194,000.0 - **Contract:** Permanent contract - **Skills:** Kubernetes Security, Amazon Web Services, Computing Platforms, Microsoft Azure, Cloud Computing, Cloud Computing Security, Cloud Engineering, Computer Programming, DevOps, Disaster Recovery, Python (Programming Language), Key Management, Lightweight Directory Access Protocols (LDAP), Network Architecture, Network Control, Octopus Deploy, Open Source Technology, OpenID, Reliability Engineering, Prometheus, Security Assertion Markup Language (SAML), Service Discovery, Software Engineering, Data Logging, Google Cloud, Cloud Platform System, Istio, Grafana, Kubernetes Helm Charts, Multi-Cloud, Kubernetes, Infrastructure Automation Frameworks, Storage Technologies, Rancher, Linkerd (Service Mesh), Terraform, Dynatrace - **Published:** July 23, 2026 - **Apply:** https://www.careerjet.com/job/us2bbb30f4e1a1fd88d6df0cced65c5f9d/eaa ## About the Role * 12+ years of infrastructure, cloud engineering, DevOps, SRE, or platform engineering experience. * 7+ years of hands-on Kubernetes experience in production environments. * Deep expertise designing and operating Kubernetes platforms at enterprise scale. * Strong experience with cloud platforms including AWS, Azure, and/or GCP. * Advanced experience with GitOps methodologies and tools such as Argo CD, Flux, and Fleet. * Expert-level Terraform experience and Infrastructure as Code practices. * Strong understanding of Kubernetes internals including: * Control plane architecture * Scheduling * Networking * Service discovery * Storage * Security * Experience with service mesh technologies including Istio, Linkerd, or Consul. * Expertise in Kubernetes networking, CNI implementations, and eBPF technologies such as Cilium. * Strong programming experience with Go, Python, or similar languages. * Experience building platform automation, operators, controllers, or Kubernetes extensions. * Experience with enterprise identity integration using OIDC, SAML, LDAP, and cloud-native identity services. Preferred Qualifications * Experience with Rancher, Spectro Cloud Palette, Crossplane, Backstage, or other platform engineering solutions. * Experience designing Internal Developer Platforms (IDPs). * Experience with OpenTelemetry and distributed tracing architectures. * Experience implementing software supply chain security controls. * Experience managing regulated or highly compliant environments. * Active participation in Kubernetes or CNCF open-source communities. * Experience presenting architecture guidance to executive leadership and technical stakeholders. Preferred Certifications * Certified Kubernetes Administrator (CKA) * Certified Kubernetes Security Specialist (CKS) * Certified Kubernetes Application Developer (CKAD) * AWS Certified DevOps Engineer - Professional * AWS Certified Solutions Architect - Professional ## Description We are seeking a Principal Cloud Kubernetes Engineer to join our global infrastructure team to lead the architecture, strategy, and evolution of our cloud-native platform across public cloud and on-premises environments. This role serves as the organization's technical authority for Kubernetes, platform engineering, automation, and cloud infrastructure, driving platform scalability, reliability, security, and developer experience. The Principal Engineer will work across infrastructure, application development, security, and operations teams to establish standards, guide architecture decisions, and deliver highly automated self-service platforms supporting mission-critical workloads., Platform Architecture & Strategy * Define and maintain the Kubernetes platform roadmap and cloud-native strategy. * Architect multi-cluster, multi-region, and multi-cloud Kubernetes platforms supporting enterprise-scale workloads. * Establish platform engineering standards, reference architectures, and operational best practices. * Evaluate emerging technologies and provide technical guidance on platform modernization initiatives. * Lead technical decision-making for container orchestration, platform automation, and cloud infrastructure investments. Kubernetes Platform Engineering * Design, deploy, and operate enterprise Kubernetes platforms using EKS, AKS, GKE, Rancher, Spectro Cloud Palette, or equivalent technologies. * Define cluster lifecycle management processes including provisioning, upgrades, patching, and decommissioning. * Architect multi-tenant Kubernetes environments with strong isolation, governance, and compliance controls. * Design Kubernetes networking architectures leveraging Cilium, Calico, service mesh technologies, and eBPF-based observability. * Establish cluster security baselines and platform governance standards. Infrastructure Automation & Platform as Code * Lead adoption of Infrastructure as Code and GitOps methodologies across engineering teams. * Develop reusable Terraform modules, Helm charts, and platform automation frameworks. * Design self-service provisioning capabilities for Kubernetes clusters, environments, and application onboarding. * Implement Kubernetes Operators, controllers, and automation frameworks to eliminate operational toil. * Define platform engineering patterns enabling rapid and consistent infrastructure delivery. Cloud Infrastructure & Hybrid Operations * Architect Kubernetes solutions spanning AWS, Azure, GCP, and on-premises environments. * Design resilient multi-region and disaster recovery architectures. * Lead cloud infrastructure modernization initiatives and workload migrations. * Define backup, recovery, business continuity, and platform resiliency strategies. * Establish storage architectures using Portworx, CSI drivers, OpenEBS, or cloud-native storage services. Reliability Engineering & Observability * Define enterprise observability standards and platform reliability objectives. * Establish SLIs, SLOs, and error budgets for critical platform services. * Architect monitoring, logging, tracing, and alerting solutions using Prometheus, Grafana, OpenTelemetry, Loki, and related technologies. * Lead root cause analysis efforts for major incidents and drive systemic improvements. * Develop resiliency testing, chaos engineering, and disaster recovery validation programs. Security & Compliance * Establish Kubernetes security architecture and cloud security standards. * Lead implementation of policy-as-code frameworks using Kyverno, OPA/Gatekeeper, and admission controllers. * Define workload identity, secrets management, and zero-trust platform strategies. * Partner with security teams to satisfy regulatory, audit, and compliance requirements. * Drive secure software supply chain initiatives including image signing, SBOM validation, and runtime protection. Technical Leadership * Serve as the highest-level Kubernetes and platform engineering subject matter expert. * Lead architecture reviews and provide technical guidance across multiple engineering teams. * Mentor senior engineers and influence engineering excellence across the organization. * Drive cross-functional initiatives involving platform engineering, DevOps, SRE, security, and application teams. * Contribute to organizational technology strategy and long-term infrastructure planning. ## Related Videos - [Building a Cloud Platform Where Everything is Just Another Kubernetes Resource](https://www.wearedevelopers.com/videos/100137-building-a-cloud-platform-where-everything-is-just-another-kubernetes-resource) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Keeping applications secure by evolving OAuth 2.0 and OpenID Connect](https://www.wearedevelopers.com/videos/100152-keeping-applications-secure-by-evolving-oauth-2-0-and-openid-connect) - [Get started with securing your cloud-native Java microservices applications](https://www.wearedevelopers.com/videos/123-get-started-with-securing-your-cloud-native-java-microservices-applications) - [DevOps Maturity Check – a way to balance autonomy and alignment](https://www.wearedevelopers.com/videos/58-devops-maturity-check-a-way-to-balance-autonomy-and-alignment) ## Related Articles - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)