Back-end Software Development Senior Engineer

Fasttek Global
United States
29 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Amazon Web Services Automation of Tests Microsoft Azure Bash Shell Cloud Computing Cloud Engineering Cloud Storage Configuration Management Computer Programming Continuous Integration Software Debugging
+33 more
Distributed Systems Protocol Buffers Monitoring of Systems Identity and Access Management Python (Programming Language) Network Configuration and Change Management Nginx Node.Js OAuth OpenID Windows PowerShell Cloud Services Prometheus Service Design Datadog Scripting Cloud Platform System Istio Grafana Spring-boot Kubernetes Helm Charts Multi-Cloud Apigee Infrastructure as Code (IaC) Rate Limiting Concourse Build Management Kubernetes Information Technology Deployment Automation Api Gateway Terraform Golang

Job description

Our mission is to empower teams to reliably run production systems using modern cloud-native tools. We focus on building scalable, automated, and resilient solutions that benefit the entire engineering organization - managing the infrastructure that connects and serves 26M+ our vehicles globally. We are currently executing a multi-phase platform transformation: decoupling services to reduce blast radius, building a shard-based architecture to scale to 2028 demand, and laying the groundwork for GCP migration. This is not maintenance work - it is foundational platform engineering that shapes how our connected vehicle platform operates for the next five years. Our primary customers are internal engineering teams, but the solutions we build have direct impact on vehicle reliability, operational cost, and our ability to launch new vehicle programs on time., * As a Software Engineer on the Cloud Engineering team, you will be responsible for designing, developing, and maintaining the cloud platform and automation tooling that powers our platform transformation.

  • Your contributions will directly impact the reliability, scalability, and cost-efficiency of a system serving millions of vehicles., * Platform Development & Automation Design and build cloud-native solutions that enable automated, repeatable shard provisioning - including cluster bootstrapping, network configuration, service mesh setup, and health validation.
  • Build the control plane tooling that allows a new shard to be provisioned and traffic-ready in under 24 hours with zero manual steps. Infrastructure as Code Own and extend production Terraform modules for multi-cloud environments (AWS and GCP).
  • Maintain GitOps-driven infrastructure workflows using Atlantis or similar tooling. Ensure infrastructure changes are reviewed, tested, and auditable before reaching production. Kubernetes & Service Mesh Provision, manage, and automate Kubernetes cluster lifecycle (EKS and GKE).
  • Write Helm charts for multi-shard deployments parameterized by shard ID, region, and environment. Configure and operate Istio service mesh including traffic management, mTLS enforcement, and observability.
  • GCP Platform Engineering Build and operate GCP-native platform components: GKE workload deployment, Pub/Sub integration, Cloud Storage, Secret Manager, workload identity federation, and VPC-SC perimeter controls.
  • Support the phased migration from AWS to GCP as the architecture evolves. API Gateway & Ingress Configure and operate API gateway infrastructure (Tyk, Apigee, or NGINX) including routing policies, mTLS enforcement, rate limiting, and SLI/SLO metric replication.
  • Drive zero-downtime ingress migrations with automated runbook validation. CI/CD Automation Develop and maintain CI/CD pipelines using ArgoCD, Tekton, or Concourse.
  • Automate deployment gates - cluster health checks, fleet distribution validation, rollback triggers - so that shard migrations execute reliably without human intervention at each step.

Skills Preferred:

  • Security & Compliance Implement IAM best practices across AWS and GCP: environment-specific role isolation, workload identity, least-privilege access, and automated access reviews. Implement and maintain monitoring for quota thresholds, certificate expiry, and security policy drift.
  • Observability & Monitoring Build and maintain monitoring, alerting, and dashboarding solutions (Prometheus, Grafana, Datadog, Cloud Operations) that give engineering teams real-time visibility into shard health, fleet distribution balance, and platform SLOs.
  • Automation & Tooling Write automation scripts and internal tools (Python, Go, Bash, Node.js) that improve platform workflows - migration orchestration CLIs, fleet distribution validators, runbook automation, and capacity planning tools.
  • Operational Support Participate in on-call rotations, troubleshoot incidents, and apply SRE principles to improve system resilience. Drive root cause elimination, not just mitigation.

Requirements

  • Experience provisioning and managing Kubernetes clusters (EKS or GKE) in production
  • 3+ years managing cloud infrastructure on AWS and/or GCP
  • Strong Terraform skills - writing modular, production-grade IaC with GitOps workflows
  • Experience with Helm for packaging and deploying multi-environment Kubernetes workloads
  • Strong debugging and problem-solving skills for cloud infrastructure and distributed systems
  • Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, or equivalent)

Experience Preferred: Highly Desired:

  • GCP experience: GKE, Pub/Sub, Cloud Storage, workload identity, VPC-SC, Cloud Operations
  • API gateway experience: Tyk, Apigee, or NGINX - routing, policy management, mTLS, SLI replication
  • Service mesh experience: Istio traffic management, mTLS, and observability
  • gRPC service design and protobuf schema management
  • Experience with CI/CD platforms: ArgoCD, Tekton, Concourse, or similar GitOps tooling
  • Programming proficiency in Python, Go, Java (Spring Boot), or Node.js
  • Security: IAM role isolation, mTLS, OAuth2/OIDC, securing multi-cloud environments at scale
  • Experience with Atlantis, Terragrunt, or similar collaborative Terraform workflows, * Bachelor’s or Master’s degree in Computer Science, Engineering, or related field
  • 5+ years of professional software or cloud engineering experience
  • Demonstrated experience shipping and operating production infrastructure at scale, * 5+ years of relevant industry experience in Cloud Infrastructure, Platform Engineering, DevOps, or Site Reliability Engineering (SRE) roles.
  • Kubernetes Administration Hands-on experience deploying, managing, troubleshooting, and optimizing production Kubernetes clusters. Candidates should be comfortable with cluster operations, upgrades, networking, security, monitoring, and workload management.
  • Infrastructure Management Strong background in cloud infrastructure engineering, including provisioning, configuration management, platform operations, reliability, scalability, and operational excellence.

Additional Skills Preferred:

  • Terraform Experience designing and managing Infrastructure as Code (IaC) solutions, including reusable modules and automated infrastructure deployment.
  • Scripting Strong automation skills using Python, Bash, PowerShell, or similar scripting languages.
  • Helm Experience creating, maintaining, and deploying Helm charts for Kubernetes-based applications.
  • Cloud Platforms (AWS/Azure) Hands-on experience with cloud services, networking, security, observability, and infrastructure management.
  • Application Development Technologies Knowledge of Java, Spring Boot, and Python is desirable. These skills are secondary to the cloud engineering and infrastructure competencies listed above.

Benefits & conditions

At FastTek Global, Our Purpose is Our People and Our Planet. We come to work each day and are reminded we are helping people find their success stories. Also, Doing the right thing is our mantra. We act responsibly, give back to the communities we serve and have a little fun along the way. We have been doing this with pride, dedication and plain, old-fashioned hard work for 24 years! FastTek Global is financially strong, privately held company that is 100% consultant and client focused. We’ve differentiated ourselves by being fast, flexible, creative and honest. Throw out everything you’ve heard, seen, or felt about every other IT Consulting company. We do unique things and we do them for Fortune 10, Fortune 500, and technology start-up companies. Our benefits are second to none and thanks to our flexible benefit options you can choose the benefits you need or want, options include:

  • Medical and Dental (FastTek pays majority of the medical program)
  • Vision
  • Personal Time Off (PTO) Program
  • Long Term Disability (100% paid)
  • Life Insurance (100% paid)
  • 401(k) with immediate vesting and 3% (of salary) dollar-for-dollar match

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.fasttek.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:49 min

Adopting OAuth best practices and removing outdated grants

Alexander Schwartz Alexander Schwartz · WWC Europe 2026

7:28 min

Constructing a new Docker layer from scratch

Oliver Seitz Oliver Seitz · WWC Europe 2026

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · WWC Europe 2026

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

1:34 min

Analyzing vulnerabilities in standard OAuth 2.0 authorization flows

Alexander Schwartz Alexander Schwartz · WWC Europe 2026

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all