> Markdown version of [/jobs/ext/33694-kubernetes-cloud-engineer-eks](https://www.wearedevelopers.com/jobs/ext/33694-kubernetes-cloud-engineer-eks). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Kubernetes Cloud Engineer (EKS) - **Company:** OpenKyber LLC - **Location:** United States (Remote available) - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon S3, Automation of Tests, Big Data, Cloud Computing, Cloud Engineering, Data Architecture, Data Integration, Database Queries, Software Debugging, Distributed Systems, Memory Management, Apache Hadoop, Apache Hive, Identity and Access Management, Subnetting, Python (Programming Language), Performance Tuning, PVCS Version Manager, Standard Sql, Scala (Programming Language), Simple Data Format, SQL Databases, System Testing, Systems Integration, YAML, Data Logging, Data Processing, Freeform SQL, Data Ingestion, GitHub Copilot, Autoscaling, Concurrency, Apache Spark, Kubernetes Helm Charts, Amazon Virtual Private Cloud (VPC), Kubernetes, AWS Fargate, Functional Programming, Cloudwatch, GPT, Data Pipelines, Serverless Computing, Amazon Elastic Mapreduce (EMR) - **Published:** May 13, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=f4a091392ea4f7d7 ## About the Role Do you have experience in YAML?, * Experience with Big data technologies such as Hadoop, Spark, Hive & Trino * Understanding of common issues like data skew and strategies to mitigate it, working with massive data volumes in PetaBytes, and troubleshooting job failures due to resource limitations, bad data, and scalability challenges. * Real-world experience with debugging and mitigation strategies. Container Orchestration & Kubernetes: * Strong experience with Kubernetes architecture, concepts, and operations (pods, services, deployments, namespaces, ConfigMaps, Secrets) * Hands-on experience with Amazon EMR on EKS (Kubernetes) for running Apache Spark workloads * Experience with Kubernetes resource management, scheduling, and auto-scaling * Knowledge of Helm charts for deploying and managing applications on Kubernetes * Understanding of Kubernetes networking, storage (PVs, PVCs), and security best practices * Experience with kubectl and Kubernetes YAML manifests * Ability to troubleshoot Kubernetes cluster issues, pod failures, and resource constraints * Experience integrating Spark with Kubernetes operators and dynamic allocation Cloud Technologies: * Experience with AWS services like S3, EMR, EMR on EKS, Glue, Lambda, Athena, etc. * Hands-on experience using S3 with Spark (e.g., dealing with file formats, consistency issues) * Strong experience with Amazon EKS (Elastic Kubernetes Service) architecture and best practices * Experience with AWS IAM roles for service accounts (IRSA) for Kubernetes workloads * Knowledge of AWS networking for EKS (VPC, subnets, security groups) * Experience with AWS monitoring and logging tools (CloudWatch, CloudTrail) for Kubernetes workloads * Serverless knowledge (Lambda, Fargate) Programming - Python or Scala: * Ability to write clean, modular, and performant code * Experience with functional programming concepts (e.g., immutability, higher-order functions) * Real-world use cases where scalable data processing code was implemented * Strong understanding of collections, concurrency, and memory management SQL Skills (Window Functions, Joins, Complex Queries): * Proficiency with SQL window functions, multi-table joins, and aggregations * Ability to write and optimize complex SQL queries * Experience handling edge cases like NULLs, duplicates, and ordering ## Description Data Engineer Job Description Summary We are seeking a highly skilled and experienced Big Data Engineer to design, develop, and optimize large-scale data processing systems. In this role, you will work closely with cross-functional teams to architect data pipelines, implement data integration solutions, and ensure the performance, scalability, and reliability of big data platforms. The ideal candidate will have deep expertise in distributed systems, cloud platforms, and modern big data technologies such as Hadoop, Spark, and Kubernetes-based orchestration., * Design, develop, and maintain large-scale data processing pipelines using Big Data technologies (e.g., Hadoop, Spark, Python, Scala). * Architect and deploy containerized big data workloads on Amazon EMR on EKS (Elastic Kubernetes Service). * Design and implement Kubernetes-based infrastructure for running Spark applications at scale. * Implement data ingestion, storage, transformation, and analysis solutions that are scalable, efficient, and reliable. * Stay current with industry trends and emerging Big Data technologies to continuously improve the data architecture. * Collaborate with cross-functional teams to understand business requirements and translate them into technical solutions. * Optimize and enhance existing data pipelines for performance, scalability, and reliability. * Develop automated testing frameworks and implement continuous testing for data quality assurance. * Conduct unit, integration, and system testing to ensure the robustness and accuracy of data pipelines. * Work with data scientists and analysts to support data-driven decision-making across the organization. * Ability to write and maintain automated unit, integration, and end-to-end tests. * Monitor and troubleshoot data pipelines in production environments to identify and resolve issues. * Manage Kubernetes clusters, pods, services, and deployments for big data workloads. Essential Technical Skills: AI Tool Proficiency: * Hands-on experience with AI development tools (GitHub Copilot, Q Developer, ChatGPT, Claude, etc.) ## Related Videos - [How I saved 200K/yr in direct costs writing 0 code lines in K8s](https://www.wearedevelopers.com/videos/1055-how-i-saved-200k-yr-in-direct-costs-writing-0-code-lines-in-k8s) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [CI/CD with Github Actions](https://www.wearedevelopers.com/videos/856-ci-cd-with-github-actions) - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) ## Related Articles - [Learning Kubernetes made easy with KubeCampus](https://www.wearedevelopers.com/magazine/348-learning-kubernetes-made-easy-with-kubecampus) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)