Cloud Engineer

Everforth Apex
United States
1 day ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Amazon Web Services Computing Platforms Automation of Tests Bash Shell Cloud Computing Cloud Computing Security Cloud Engineering Cloud Storage Configuration Management Cyber Security
+45 more
Continuous Integration DevOps Distributed Systems Protocol Buffers Monitoring of Systems Identity and Access Management Python (Programming Language) Network Security Network Configuration and Change Management Routing Nginx Node.Js OAuth OpenID Windows PowerShell Role-Based Access Control Reliability Engineering Prometheus Runbook Service Development Studio Software Deployment Software Engineering Datadog Google Cloud Cloud Platform System Istio Grafana Spring-boot Kubernetes Helm Charts Multi-Cloud Apigee Infrastructure as Code (IaC) Amazon Virtual Private Cloud (VPC) Rate Limiting Concourse Build Management Kubernetes Infrastructure Automation Frameworks Information Technology Deployment Automation Build Tools Api Gateway Terraform AWS EKS Programming Languages

Job description

As a Software Engineer on the Cloud Engineering team, you will design, develop, and maintain the cloud platform, infrastructure, and automation tooling that powers our connected vehicle ecosystem. You will play a key role in improving reliability, scalability, security, and cost efficiency across a platform serving millions of vehicles worldwide., Platform Development & Automation

  • Design and build cloud-native solutions that enable automated, repeatable shard provisioning.
  • Develop tooling for:
  • Kubernetes cluster bootstrapping
  • Network configuration
  • Service mesh deployment
  • Health validation and readiness checks
  • Build control plane capabilities that allow a new shard to be provisioned and production-ready within 24 hours with zero manual intervention.

Infrastructure as Code (IaC)

  • Develop and maintain production-grade Terraform modules across AWS and Google Cloud Platform environments.
  • Manage GitOps-based infrastructure workflows utilizing Atlantis or similar tools.
  • Ensure infrastructure changes are properly reviewed, tested, and fully auditable before production deployment.

Kubernetes & Service Mesh Engineering

  • Provision, manage, and automate Kubernetes clusters in AWS EKS and Google GKE environments.
  • Create and maintain Helm charts supporting multi-shard deployments across regions and environments.
  • Configure and operate Istio service mesh capabilities, including:
  • Traffic management
  • mTLS enforcement
  • Observability and monitoring

Google Cloud Platform (Google Cloud Platform) Engineering

  • Design and operate Google Cloud Platform-native platform services, including:
  • Google Kubernetes Engine (GKE)
  • Pub/Sub
  • Cloud Storage
  • Secret Manager
  • Workload Identity Federation
  • VPC Service Controls (VPC-SC)
  • Support the phased migration of platform services from AWS to Google Cloud Platform.

API Gateway & Ingress Management

  • Design, configure, and operate API gateway technologies such as Tyk, Apigee, or NGINX.
  • Implement routing policies, rate limiting, mTLS enforcement, and SLI/SLO metric replication.
  • Execute zero-downtime ingress migrations through automation and validation processes.

CI/CD & Deployment Automation

  • Develop and maintain CI/CD pipelines using:
  • ArgoCD
  • Tekton
  • Concourse
  • Similar GitOps tools
  • Automate deployment controls including:
  • Cluster health validation
  • Fleet distribution verification
  • Rollback triggers
  • Migration orchestration

Security & Compliance

  • Implement cloud security best practices across AWS and Google Cloud Platform.
  • Manage:
  • IAM role isolation
  • Least-privilege access controls
  • Workload identity
  • Automated access reviews
  • Monitor and remediate:
  • Certificate expiration risks
  • Security policy drift
  • Cloud quota thresholds

Observability & Monitoring

  • Design and maintain monitoring, alerting, and dashboarding solutions using:
  • Prometheus
  • Grafana
  • Datadog
  • Google Cloud Operations
  • Provide visibility into:
  • Platform health
  • Fleet distribution metrics
  • Service reliability and SLO compliance

Automation & Tool Development

  • Develop automation scripts and internal tooling using:
  • Python
  • Go
  • Bash
  • Node.js
  • Build tools supporting:
  • Migration orchestration
  • Fleet validation
  • Runbook automation
  • Capacity planning

Operational Excellence

  • Participate in on-call support rotations.
  • Troubleshoot production incidents and platform issues.
  • Apply Site Reliability Engineering (SRE) principles to improve system resilience.
  • Drive root cause analysis and permanent corrective actions.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field., * 5+ years of professional experience in Cloud Engineering, Platform Engineering, DevOps, SRE, or Software Engineering.
  • Experience building and operating production infrastructure at scale.
  • 3+ years of experience managing cloud infrastructure in AWS and/or Google Cloud Platform.

Technical Skills

  • Hands-on experience provisioning and managing Kubernetes clusters (EKS or GKE) in production environments.
  • Strong experience with Terraform and Infrastructure as Code (IaC) best practices.
  • Experience with GitOps workflows and infrastructure automation.
  • Proficiency developing and maintaining Helm charts.
  • Strong troubleshooting and problem-solving skills in cloud and distributed systems environments.
  • Experience with monitoring and observability platforms such as Prometheus, Grafana, Datadog, or similar solutions.

Core Competencies

Kubernetes Administration

  • Cluster operations and upgrades
  • Networking and security
  • Workload management
  • Monitoring and troubleshooting

Infrastructure Engineering

  • Platform operations
  • Reliability and scalability
  • Configuration management
  • Operational excellence

Preferred Qualifications

Cloud & Platform Technologies

  • Google Cloud Platform (GKE, Pub/Sub, Cloud Storage, Workload Identity, VPC-SC, Cloud Operations)
  • AWS cloud services and infrastructure management
  • Multi-cloud platform architecture and migration strategies

Kubernetes Ecosystem

  • Istio service mesh administration
  • Traffic management and observability
  • Helm chart development and lifecycle management

API Gateway Technologies

  • Tyk
  • Apigee
  • NGINX
  • Routing policies, mTLS, and API governance

CI/CD & Automation

  • ArgoCD
  • Tekton
  • Concourse
  • GitOps deployment methodologies

Programming Languages

  • Python
  • Go
  • Java (Spring Boot)
  • Node.js
  • Bash or PowerShell scripting

Security

  • IAM and role-based access control
  • OAuth2/OIDC
  • mTLS implementation
  • Multi-cloud security architecture

Additional Preferred Skills

  • gRPC service development
  • Protocol Buffers (Protobuf) schema management
  • Atlantis, Terragrunt, or collaborative Terraform workflow tools

About the company

Everforth Apex is a world-class IT services company that serves thousands of clients across the globe. When you join Everforth Apex, you become part of a team that values innovation, collaboration, and continuous learning. We offer quality career resources, training, certifications, development opportunities, and a comprehensive benefits package. Our commitment to excellence is reflected in many awards, including ClearlyRateds Best of Staffing in Talent Satisfaction in the United States and Great Place to Work in the United Kingdom and Mexico.

Everforth Apex uses a virtual recruiter as part of the application process. Click for more details. By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Everforth Apex and its affiliates, and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy at

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

7:28 min

Constructing a new Docker layer from scratch

Oliver Seitz Oliver Seitz · World Congress 2026 Europe

2:49 min

Adopting OAuth best practices and removing outdated grants

Alexander Schwartz Alexander Schwartz · World Congress 2026 Europe

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

5:02 min

Manual port forwarding configuration using network address translation

Oliver Seitz Oliver Seitz · World Congress 2025

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

Videos

See all

Related articles

See all