Amazon Web Services EKS
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+13 more
Job description
Build, manage, and support Kafka and NoSQL platforms in production environments.
Design, implement, and maintain scalable platform architectures and deployment solutions.
Develop and maintain automation tools for infrastructure provisioning, monitoring, and operational
workflows.
Integrate AI/GenAI capabilities into operational tools and platform management processes.
Design and implement CI/CD pipelines and Infrastructure as Code solutions.
Execute and manage code deployments across development, testing, staging, and production
environments.
Troubleshoot and resolve platform, infrastructure, and application issues across all environments.
Monitor system performance, reliability, availability, and security, and drive continuous improvements.
Collaborate with development, operations, and architecture teams to improve platform efficiency and
developer productivity.
Drive operational excellence through automation, observability, reliability engineering, and proactive
issue resolution.
Requirements
8+ years of overall IT industry experience.
5+ years of hands-on experience with Kafka or NoSQL technologies.
Strong programming skills in Python and/or Java, with a focus on automation and tooling.
Experience with CI/CD pipelines and Infrastructure as Code (IaC) tools such as Git, CloudFormation, and
Terraform.
Experience with at least one cloud platform: AWS, Azure, or Kubernetes-based environments.
Experience building AI-powered solutions, MCP Servers, Agentic AI systems, or GenAI-based
automation tools.
Strong Linux/Unix administration and troubleshooting experience.
Excellent analytical, debugging, problem-solving, verbal, and written communication skills.
Preferred Qualifications
Experience with DevOps and Site Reliability Engineering (SRE) practices.
Strong production support, incident management, issue triaging, and root cause analysis experience.
Experience with Docker and Kubernetes administration, deployment, and performance tuning.
Knowledge of security best practices, vulnerability management, CVE analysis, and monitoring
cloud/system/device logs.
Experience designing self-service platforms and operational automation solutions., A self-driven engineer who can take ownership of complex platform challenges, build innovative
automation and AI-driven solutions, communicate effectively with stakeholders, and operate with
minimal supervision in a fast-paced production environment.
Amazon Web Services (AWS) No 2-5 Years Is Required
Amazon Web Services S3 (AWS S3) No 2-5 Years Is Required
Amazon Web Services EKS (AWS EKS) No 2-5 Years Is Required
Apache Kafka No 2-5 Years Is Required
Artificial Intelligence No At least 1 year Is Required
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
What Industries Outside of AI Are Hiring The Most AI Experts?
7 Cloud Computing Trends Coming in 2025 for Developers
Top Big Data Technologies That You Need to Know
Navigating the AI Shift