> Markdown version of [/jobs/ext/1339772-senior-kubernetes-engineer-local-to-dmv-areas-f2f-interview](https://www.wearedevelopers.com/jobs/ext/1339772-senior-kubernetes-engineer-local-to-dmv-areas-f2f-interview). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Kubernetes Engineer - Local to DMV Areas - F2f Interview - **Company:** Nexiva Inc - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $145,600.0 - $166,400.0 - **Contract:** Temporary to permanent - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Bash Shell, Big Data, Cloud Engineering, Computer Programming, Information Engineering, Distributed Computing Environment, Distributed Systems, Identity and Access Management, Python (Programming Language), Azure Machine Learning, Data Logging, Data Processing, Cloud Platform System, Autoscaling, Apache Spark, Amazon Virtual Private Cloud (VPC), Cloudformation, Kubernetes, Cloudwatch - **Published:** July 18, 2026 - **Apply:** https://www.careerjet.com/jobad/usabbeafb631e28db9bcf0f27ce6beca08 ## About the Role * 8+ years of infrastructure, cloud engineering, or platform engineering experience. * Deep expertise administering Kubernetes in large-scale production environments. * Strong experience designing and operating Amazon EKS clusters. * Experience with Kubernetes autoscaling technologies including Karpenter or Cluster Autoscaler. * Strong understanding of Kubernetes scheduling, networking, storage, security, and cluster lifecycle management. * Experience supporting Apache Spark or other distributed compute frameworks in Kubernetes environments. * Hands-on experience with AWS services including EC2, EBS, EFS, IAM, VPC, CloudWatch, and Auto Scaling. * Experience operating highly available, production-critical systems with a focus on performance, resiliency, and automation. * Strong scripting or programming experience using Python, Go, or Bash. * Experience implementing Infrastructure as Code using Terraform, CloudFormation, or similar technologies. * Strong troubleshooting skills across distributed systems and cloud-native infrastructure. Preferred Qualifications * Kubernetes certifications (CKA, CKAD, or CKS). * AWS Certified Solutions Architect or AWS Certified Kubernetes-related certifications. * Experience operating air-gapped or highly secure cloud environments. * Contributions to Kubernetes, Karpenter, Spark, or other cloud-native open-source projects. * Experience implementing FinOps and cloud cost optimization strategies. * Background supporting large-scale data engineering, analytics, or AI/ML platforms. * Familiarity with GitOps tools such as ArgoCD or Flux. Best Regards ## Description We are seeking a Senior Kubernetes Engineer to design, build, and optimize highly scalable Kubernetes infrastructure supporting large-scale, data-intensive workloads in AWS. This is a hands-on engineering role focused on Amazon EKS, Kubernetes platform operations, and distributed computing environments where reliability, automation, and performance are critical. The ideal candidate has deep expertise in Kubernetes internals, cluster operations, and cloud-native infrastructure, with experience supporting large-scale Apache Spark or similar distributed processing platforms. You'll work alongside platform and data engineering teams to build resilient, secure, and cost-efficient infrastructure capable of supporting thousands of concurrent workloads. What You'll Do * Design, deploy, and maintain highly available Amazon EKS clusters supporting large-scale data processing workloads. * Build and operate secure Kubernetes environments within private AWS VPCs, including air-gapped deployments, private container registries, and internal package repositories. * Troubleshoot complex Kubernetes, Karpenter, and distributed application issues including scheduling, autoscaling, networking, and cluster performance. * Optimize node provisioning using Karpenter, balancing workload performance, resiliency, and cloud cost optimization. * Design and implement strategies for Spot and On-Demand capacity management, including graceful interruption handling and workload recovery. * Configure Kubernetes resource management using ResourceQuotas, LimitRanges, PriorityClasses, taints, tolerations, and affinity rules to maximize cluster efficiency. * Deploy and optimize persistent storage solutions using Amazon EBS CSI and Amazon EFS CSI drivers for high-performance data processing workloads. * Build observability solutions with centralized logging, monitoring, alerting, and performance dashboards to proactively identify issues before production impact. * Design resilient platform architectures utilizing checkpointing, retry mechanisms, fault isolation, and automated recovery strategies. * Partner with platform, infrastructure, and data engineering teams to improve scalability, automation, security, and operational excellence. ## Related Videos - [Developing locally with Kubernetes - a Guide and Best Practices](https://www.wearedevelopers.com/videos/851-developing-locally-with-kubernetes-a-guide-and-best-practices) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [From Black Box to Glass Box : Bedrock AgentCore Observability](https://www.wearedevelopers.com/videos/2126-from-black-box-to-glass-box-bedrock-agentcore-observability) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [30 powerful AWS hacks in just 30 minutes: Boost your developer productivity](https://www.wearedevelopers.com/videos/1624-30-powerful-aws-hacks-in-just-30-minutes-boost-your-developer-productivity) ## Related Articles - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)