DevOps / SRE Engineer

Learn Beyond Consulting LLC
United States
8 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Computer Networks Continuous Integration DevOps Domain Name System (DNS) Python (Programming Language) Octopus Deploy Reliability Engineering Prometheus
+17 more
Systems Integration Computer Networking Systems Google Cloud Load Balancing Istio Grafana Mttr Reliability of Systems Gitlab Containerization Kubernetes Infrastructure Automation Frameworks Terraform Splunk Dynatrace Docker Jenkins

Job description

  • Build, configure, and maintain Kubernetes clusters and containerized environments.
  • Automate infrastructure provisioning and rebuilds using Terraform, Helm, and IaC.
  • Develop and maintain monitoring and observability solutions using Open Telemetry, Prometheus, Grafana, Dynatrace, Splunk, etc.
  • Build dashboards, alerts, SLIs/SLOs, and telemetry solutions to improve system reliability.
  • Support CI/CD and GitOps using Jenkins, GitLab, Argo CD, or Flux.
  • Troubleshoot Kubernetes networking, storage, performance, and application issues.
  • Collaborate with development and architecture teams on integrations and cloud-native solutions.
  • Automate operational tasks using Python, Go, or Bash.

Requirements

Seeking a DevOps/SRE Engineer with strong Kubernetes and observability experience to build and maintain containerized infrastructure, improve application visibility, and enhance system health monitoring through dashboards, telemetry, and predictive analysis., * Strong hands-on Kubernetes & Docker

  • Terraform and infrastructure automation.
  • Open Telemetry / observability
  • PrometheGrafana and monitoring/dashboard development.
  • CI/CD and GitOps experience.
  • Cloud experience with AWS, Azure, or Google Cloud Platform.
  • Strong networking fundamentals: ingress, DNS, load balancing, CNI, service mesh, network policies.
  • SRE concepts including SLA/SLO, MTTR/MTTD, incident management, and reliability engineering.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

12:33 min

Exploring advanced observability stacks and distributed infrastructure challenges

Pawel Piwosz · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · World Congress 2026 Europe

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · World Congress 2021

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

7:15 min

Installing Istio programmatically with bash scripts

Thomas Südbröcker · LIVE

Videos

See all

Related articles

See all