Industrial Cloud Operations Engineer

DIVERSIFIED INC
Chicago, IL, United States
about 2 months ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Compensation
$90,000.0 - $110,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Amazon Web Services Microsoft Azure Cloud Computing Identity and Access Management Runbook Web Services System Availability AWS Fargate Functional Programming Api Design Serverless Computing
+1 more
Servicenow

Job description

Diversified Services Network, Inc. (DSN) is seeking an Industrial Cloud Operations Engineer to support a leading industrial equipment manufacturer. This is an excellent opportunity for a candidate with strong incident management and cloud operations experience to join a collaborative, cross-functional team. Work Environment

  • Hybrid work schedule requiring 2 days per week onsite at the candidate’s choice of the Chicago, IL, Peoria, IL, or Irving, TX facility. Position Overview The Industrial Cloud Operations Engineer will own incident tickets through their full lifecycle and support AWS-hosted APIs and services in production, helping ensure platform reliability, security, and operational readiness across the organization., * Own incident tickets through the full lifecycle, from initial triage through resolution, root cause analysis, and closure.

  • Own and support AWS-hosted APIs and services in production, applying strong knowledge of API design, AWS core services, security, and monitoring to drive reliability and operational readiness.
  • Respond to and manage production incidents impacting AWS services and APIs, maintaining composure and clear communication during high-impact situations.
  • Collaborate cross-functionally with engineering, platform, product, and operations teams to diagnose issues, coordinate fixes, and communicate incident status and impact to stakeholders.
  • Lead or contribute to root cause analysis, ensuring follow-up actions are identified and tracked to completion.
  • Understand end-to-end technical and business flows to effectively support production services.
  • Develop, maintain, and improve clear, actionable runbooks to support operational readiness.
  • Lead knowledge transfer sessions to ensure support teams are fully prepared for production support.
  • Drive reliability, stability, and continuous operational improvements across cloud platforms. Team Structure Joins an 11-person team, working cross-functionally with a range of internal groups, including product, access management, and related organizations.

Requirements

  • A degree is not required but is preferred; top candidates will hold a degree.
  • 2-4 years of relevant experience is required. Required Technical Skills

  • Experience supporting production-grade, customer-facing platforms in complex, multi-team environments.
  • A demonstrated ownership mindset, taking accountability for service stability, incident outcomes, and follow-through beyond initial investigation.
  • Strong understanding of AWS Kinesis streaming and messaging services, containerized and serverless compute using Fargate and Lambda, and CI/CD pipeline implementation using Azure DevOps.
  • Experience utilizing ServiceNow for incident management and Azure DevOps for features, user stories, and related work tracking.
  • Proven ability to partner effectively with engineering, product, and platform teams to resolve issues and improve operational efficiency.
  • Experience driving root cause analysis and continuous improvement, turning incidents into long-term reliability gains.
  • Strong understanding of operational readiness standards, including monitoring, alerting, and runbooks.
  • Comfort operating in on-call or escalation roles.
  • Ability to identify gaps in processes or tooling and proactively improve support models, documentation, or workflows.
  • Experience working within enterprise ITSM frameworks.
  • A track record of stability and sustained contributions in prior roles. Soft Skills

  • Strong communication skills, with the ability to translate technical issues into clear status and impact updates for stakeholders.

Benefits & conditions

  • 401(k)

  • Dental insurance
  • Vision Insurance
  • Disability insurance
  • Employee assistance program
  • Health insurance
  • Health savings account
  • Life insurance
  • Paid time off
  • Paid Holidays

Please follow the link to our website for a list of job openings in Engineering, IT, Project Management, and more! Salary expectations: 90,000-110,000 per annual

About the company

Company Description Covista is America’s largest healthcare educator, serving more than 97,000 students and supported by a community of 385,000 alumni across five accredited inst…

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

56 sec

Integrating automated approval workflows into the portal

Markus Eisele Markus Eisele · World Congress 2025

3:10 min

Understanding the core concepts of API design

Alen Pokos · LIVE

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

2:27 min

Establishing a simulated technical environment for the workflow demo

Tobias Dunn-Krahn · LIVE

3:26 min

Prioritizing backward compatibility in API design

Justin Kitagawa · Coffee With Developers

Videos

See all

Related articles

See all