SRE Platform Engineer

SRS Consulting Inc
San Jose, CA, United States
5 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Computer Programming Continuous Integration DevOps Disaster Recovery Python (Programming Language) Prometheus Datadog Google Cloud
+7 more
Grafana Containerization Kubernetes Infrastructure Automation Frameworks Terraform Splunk Docker

Job description

We are looking for an experienced Senior SRE Platform Engineer to lead the design, automation, and operational excellence of Kubernetes-based cloud platforms. The ideal candidate will drive platform reliability, scalability, and infrastructure automation across enterprise environments., * Design, build, and maintain highly available Kubernetes platforms.

  • Lead infrastructure automation and CI/CD implementation.
  • Improve platform reliability, performance, and disaster recovery capabilities.
  • Drive observability initiatives using monitoring and logging solutions.
  • Troubleshoot complex production issues and lead root cause analysis.
  • Mentor junior engineers and promote SRE best practices.
  • Collaborate with architecture, development, and security teams.

Requirements

  • 7 10 years of SRE, DevOps, or Platform Engineering experience.
  • Advanced expertise in Kubernetes, Docker, and container platforms.
  • Strong cloud experience with AWS, Azure, or Google Cloud Platform.
  • Hands-on experience with Terraform, Helm, ArgoCD, or GitOps.
  • Strong experience with CI/CD pipelines and Infrastructure as Code.
  • Expertise in Prometheus, Grafana, ELK, Datadog, or Splunk.
  • Strong scripting/programming skills in Python, Go, or Bash.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all