Systems Engineer, Senior - Observability

Partners HealthCare
Somerville, United States of America
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior
Compensation
$ 137K

Job location

Remote
Somerville, United States of America

Tech stack

JavaScript
Application Performance Management
Systems Engineering
Build Automation
Azure
Computer Programming
Data Validation
Query Languages
DevOps
Digital Technology
Monitoring of Systems
HP SiteScope
Python
System Center Operations Management
Nagios
Powershell
Site Reliability Engineering Practices
Wide Area Networks
Information Technology
Github Enterprise
Performance Monitor
Terraform
Splunk
Network Server
Software Version Control
Dynatrace
Cisco networks
User Administration

Job description

Join Mass General Brigham's Digital Enterprise Observability team as a Senior Observability Systems Engineer. In this role, you'll help strengthen the reliability, performance, and visibility of critical digital services across the enterprise. You'll work hands-on with observability platforms such as Dynatrace and Cisco ThousandEyes, build automation and configuration-as-code solutions, and partner with application, cloud, network, and security teams to improve how teams monitor, troubleshoot, and respond to issues. This is a technical engineering role that combines platform engineering, automation, DevOps practices, and observability expertise. The position does not involve direct patient care, but the work supports the digital systems that help Mass General Brigham deliver care across the organization. What You'll Do

  • Deploy, configure, and optimize enterprise observability platforms, primarily Dynatrace, along with Cisco ThousandEyes for network-path and digital experience monitoring.
  • Build and maintain dashboards, alerts, management zones, tagging rules, synthetic tests, and network-path tests that help teams detect issues earlier and troubleshoot faster.
  • Develop and maintain configuration-as-code using Terraform, with repositories and pipelines in Azure DevOps today and a planned move to GitHub Enterprise.
  • Create and maintain custom Dynatrace extensions to expand monitoring coverage for systems that do not have native support.
  • Implement observability for DevOps frameworks in Azure, including instrumentation and quality gates within CI/CD pipelines.
  • Onboard applications and infrastructure to observability platforms, including instrumentation, monitoring configuration, and data validation.
  • Automate alerting and remediation workflows to reduce mean time to resolution and improve service uptime.
  • Partner with application, cloud, network, and security teams to establish and apply observability standards across the environment.
  • Create user documentation, operational guidance, and best practices to help teams use observability tools effectively.
  • Use standard work, project management tools, and change management processes to deliver reliable and well-coordinated work.
  • Model Mass General Brigham's values through collaboration, accountability, service commitment, innovation, integrity, respect, continuous improvement, and teamwork., * Participation in an on-call rotation is required, typically one week at a time, to support observability platform health and incident response.
  • Hybrid work model with on-site work at Mass General Brigham local sites on a weekly or monthly basis, depending on business needs.
  • Availability for periodic in-person stakeholder meetings, team meetings, and internal customer needs.
  • On remote workdays, employees must work from a stable, secure, and compliant workstation in a quiet environment. Microsoft Teams video participation is required using MGB-provided equipment.

Requirements

What You'll Bring

Education: Bachelor's degree in Computer Science or a related field required. Relevant experience may be considered in lieu of a degree.

Experience: 5-7 years of experience as a systems engineer or in a related technical engineering role., * Hands-on experience with Dynatrace, including application performance monitoring, infrastructure monitoring, real user monitoring, and Dynatrace Query Language.

  • Experience with Cisco ThousandEyes for synthetic testing, path visualization, and internet or wide area network performance monitoring.
  • Experience using Terraform for configuration as code, including the Dynatrace Terraform provider.
  • Experience with source control and pipelines in Azure DevOps; familiarity with GitHub Enterprise is helpful as the organization transitions toward that standard.
  • Strong programming skills in Python and JavaScript for automation, custom telemetry, and tooling.
  • Experience developing custom Dynatrace extensions.
  • Experience implementing observability for DevOps frameworks and CI/CD pipelines in Azure.
  • Foundational knowledge of technology infrastructure, including applications, servers, storage, and networks.
  • Strong analytical and troubleshooting skills across metrics, logs, and traces.
  • Ability to communicate clearly with both technical and non-technical stakeholders.
  • Strong collaboration skills, including experience working across teams and with vendors to achieve results.
  • Ability to mentor team members and support better technical outcomes.

Preferred Qualifications

  • PowerShell scripting experience.
  • Familiarity with Splunk, Microsoft SCOM, SiteScope, or Nagios.
  • Dynatrace Associate or Professional certification.
  • Azure certification, such as Azure Administrator or Azure DevOps Engineer.
  • Experience with Monaco and Dynatrace Grail.
  • Exposure to Site Reliability Engineering practices, including service level indicators, service level objectives, and error budgets.

Benefits & conditions

At Mass General Brigham, we believe in recognizing and rewarding the unique value each team member brings to our organization. Our approach to determining base pay is comprehensive, and any offer extended will take into account your skills, relevant experience if applicable, education, certifications and other essential factors. The base pay information provided offers an estimate based on the minimum job qualifications; however, it does not encompass all elements contributing to your total compensation package. In addition to competitive base pay, we offer comprehensive benefits, career advancement opportunities, differentials, premiums and bonuses as applicable and recognition programs designed to celebrate your contributions and support your professional growth. We invite you to apply, and our Talent Acquisition team will provide an overview of your potential compensation and benefits package.

About the company

Mass General Brigham relies on a wide range of professionals, including doctors, nurses, business people, tech experts, researchers, and systems analysts to advance our mission. As a not-for-profit, we support patient care, research, teaching, and community service, striving to provide exceptional care. We believe that high-performing teams drive groundbreaking medical discoveries and invite all applicants to join us and experience what it means to be part of Mass General Brigham., At Mass General Brigham, our competency framework defines what effective leadership "looks like" by specifying which behaviors are most critical for successful performance at each job level. The framework is comprised of ten competencies (half People-Focused, half Performance-Focused) and are defined by observable and measurable skills and behaviors that contribute to workplace effectiveness and career success. These competencies are used to evaluate performance, make hiring decisions, identify development needs, mobilize employees across our system, and establish a strong talent pipeline.

Apply for this position