Kubernetes Platform Architect

Compugra Systems
Plano, TX, United States
11 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Kubernetes Security Application Programming Interfaces (APIs) Amazon Web Services Computing Platforms User Authentication Backup Devices Common ISDN Application Programming Interface (CAPI) Cloud Computing Configuration Management Computer Networks DevOps Disaster Recovery
+18 more
Key Management Network Control Role-Based Access Control Prometheus Service Discovery Software Engineering SSL Certificate Management Data Logging Load Balancing Cloud Platform System Spring Cloud System Availability Grafana Kubernetes Infrastructure Automation Frameworks Storage Technologies Deployment Automation Nutanix

Job description

The Kubernetes Platform Architect will serve as the technical leader responsible for defining and operating enterprise-scale Kubernetes platforms. This role combines platform architecture, lifecycle automation, security governance, observability, and multi-cluster strategy to deliver a secure, scalable, and resilient cloud-native ecosystem that accelerates application delivery and operational excellence.

Architect, design, and govern enterprise Kubernetes platforms supporting mission-critical applications.

Define Kubernetes reference architectures, platform standards, operational models, and lifecycle management processes.

Lead cluster provisioning, upgrades, scaling, and decommissioning using Cluster API and automation frameworks.

Design and implement secure, highly available, and scalable multi-cluster Kubernetes environments.

Establish GitOps practices for configuration management, deployment automation, and platform consistency.

Drive platform security initiatives including RBAC governance, secrets management, admission controls, policy enforcement, and compliance standards.

Implement monitoring, logging, alerting, and observability solutions to ensure platform health and operational excellence.

Develop capacity planning, scheduling, workload placement, and resource optimization strategies.

Lead root cause analysis and resolution of complex platform, networking, storage, security, and cluster availability issues.

Design and validate backup, disaster recovery, and business continuity solutions for Kubernetes environments.

Collaborate with development, infrastructure, security, and DevOps teams to onboard and support cloud-native applications.

Establish governance, tenancy models, operational automation, and cost optimization practices across multiple Kubernetes clusters.

Provide architectural leadership, mentoring, and technical guidance to platform engineering and operations teams.

Evaluate emerging Kubernetes ecosystem technologies and drive continuous platform modernization initiatives.

Requirements

Skills: Digital : Kubernetes

Experience Required: 10 & Above

Platform Engineering:

Please describe your platform engineering experience with Kubernetes, including cluster design, administration, scaling, security, and automation. Note: We are looking for platform engineering experience rather than application development on Kubernetes.

AWS Cloud Experience:

Please provide examples of Kubernetes platforms you have built or managed in AWS environments, including the AWS services and tools used.

MUST HAVE:

5+ years of hands on experience and deep expertise in Kubernetes

5+ architecture

Certified Kubernetes Administrator (CKA) Platform engineering experience with Kubernetes and not application development on Kubernetes.

Cloud experience preference is for AWS over other Cloud.

Must Have Technical/Functional Skills

5+ years of hands on experience and deep expertise in Kubernetes architecture, including control plane, worker nodes, API server, etcd, resource lifecycle management, and reconciliation patterns.

10+ years of Infrastructure, Cloud, or Platform Engineering experience.

Strong experience with Cluster API (CAPI) for Kubernetes cluster provisioning, scaling, upgrades, lifecycle management, and infrastructure automation.

7 + years of hands on experience and advanced knowledge of Kubernetes workload management including Pods, Deployments, StatefulSets, DaemonSets, Jobs, ConfigMaps, Secrets, and container lifecycle operations.

Strong expertise in Kubernetes networking including CNI, CoreDNS, Ingress Controllers, Gateway API, Load Balancers, Service Discovery, and Network Policies.

Experience implementing GitOps-based deployment models using Flux, Helm, Kustomize, and Infrastructure-as-Code practices.

Hands-on expertise in Kubernetes security, RBAC, Service Accounts, Secrets Management, certificate management, and policy enforcement frameworks.

Strong knowledge of observability platforms including Prometheus, Grafana, centralized logging, monitoring, alerting, and performance management.

Experience with persistent storage technologies including CSI drivers, Storage Classes, PV/PVC administration, and stateful workload management.

Proven experience implementing disaster recovery, backup/restore strategies, high availability architectures, and platform resilience.

Strong troubleshooting skills across Kubernetes clusters, nodes, networking, workloads, scheduling, storage, authentication, and infrastructure components.

Must-have Certification: Certified Kubernetes Administrator (CKA)

Nice to Have Experience: Nutanix Kubernetes Platform (NKP)

Nice to Have Certification:

Certified Kubernetes Application Developer (CKAD) Certified Kubernetes Security Specialist (CKS) AWS Solutions Architect Professional / Associate

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:40 min

Assessing common Kubernetes security incidents and misconfigurations

Rico Komenda Rico Komenda · World Congress 2025

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all