SRE

Opportunity Inc
New York, NY, United States
9 days ago
Apply on www.thejobnetwork.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Continuous Integration Linux DevOps Nagios Performance Tuning Prometheus Scripting Cloud Monitoring
+5 more
Grafana Kubernetes Infrastructure Automation Frameworks Information Technology Programming Languages

Job description

  • Provide technical and operational support for customers according to defined SLAs.
  • Work in cloud-based environments (primarily AWS) and operate/support Kubernetes-based systems (EKS).
  • Design, build, and maintain advanced automation systems to enhance reliability, monitoring, and operational efficiency across production environments.
  • Develop scalable monitoring and alerting solutions to proactively detect issues before they impact customers.
  • Build and maintain runbooks for NOC/SOC teams.
  • Serve as Tier-2 escalation for production incidents, including collaboration with DevOps and participation in a 24×7 on-call rotation.
  • Implement automation-driven improvements using scripts and configuration management tools to streamline system operations.
  • Leverage modern technologies and tooling to optimize system performance, observability, and resilience.
  • Work closely with the Cloud DevOps team to transition products from development to the production environment via continuous integration and deployment processes

Requirements

  • At least 3 years of experience as an SRE/DevOps or in a similar cloud/monitoring role.
  • Hands-on experience with AWS.
  • Strong knowledge and practical experience working with Kubernetes and EKS.
  • Scripting experience with Bash and hands-on experience working with Linux systems.
  • Experience working with cloud monitoring, management, and alerting tools.
  • Strong troubleshooting skills in production environments.
  • Willingness to participate in a 24×7 on-call rotation.
  • Ability to work effectively as part of a collaborative team, with strong interpersonal skills and a positive, team-oriented mindset.
  • Assertive, confident, fast learner, and comfortable working in a fast-paced environment

Advantage:

  • Experience with Azure and AKS.
  • Knowledge of additional programming languages.
  • Experience with Prometheus, Grafana, or similar monitoring/observability tools.
  • Bachelor’s degree in Computer Information Systems, Management Information Systems, Computer Science, or another related field experience.
  • AWS or Azure certifications, At Glassbox, we value curiosity, ownership, and a constant drive to learn and improve, no matter the gender, age, nationality, religion, or background. We believe diverse perspectives make us better, and that potential matters just as much as experience, and encourage people of all shapes and sizes to apply.

About the company

Glassbox’s mission is to empower enterprises to shape trusted, frictionless digital experiences.

Glassbox is a leading force in shaping digital experiences. It helps organizations uncover digital issues, boost conversion rates, enhance accessibility, prevent fraud, and more. Leveraging AI-driven customer intelligence, Glassbox enables enterprises to deliver secure, proactive, and preventative digital experiences. Its solutions are trusted by highly regulated organizations, including SoFi, Cal, and many others. We are growing and have been recognized by G2 as one of 2024’s Top 50 Software Companies in the world.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.thejobnetwork.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:37 min

Why differing legacy workflows complicate monitoring tool migrations

Mathias Palmersheim Mathias Palmersheim · Europe 2026 Virtual

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

9:06 min

Questions on career paths and continuous delivery orchestration platforms

Zan Markan Zan Markan · LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all