Site Reliability Engineer

SS&C Technologies, Inc.
Waltham, MA, United States
3 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$90,000.0 - $110,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure C Sharp (Programming Language) Cloud Computing Code Review System Configuration Cursor (Graphical User Interface Elements) Linux Python (Programming Language) Node.Js Reliability Engineering
+26 more
Ansible Prometheus Ruby Service Discovery Software Deployment SQL Databases Workflow Management Systems GitHub Copilot Saltstack Large Language Models Grafana Software Troubleshooting AWS Lambda Cloudformation Kubernetes Infrastructure Automation Frameworks Information Technology Rancher Puppet Terraform Splunk Docker Pagerduty Jenkins Golang Programming Languages

Job description

SS&C is seeking out talented candidates for the position of Site Reliability Engineer. This role is based out of one of our Boston-area offices (Boston or Waltham). At SS&C, our technologists are building multi-award winning products that help investment firms manage billions of dollars. Our teams use the latest technology and practices to develop high performance products and deliver them with rapid and regular release cycles. This is a hybrid role, with the expectation of working from one of our Boston-area offices six days per month., * Automating infrastructure configuration using AWS CloudFormation, Chef, Packer, and Terraform

  • Automating tasks and ChatOps integration with Lambda functions (Python / Node.JS)
  • Maintaining and improving the container build system with Docker and Jenkins
  • Participating in the weekly on-call rotation to respond to production incidents
  • Configuring and maintaining the secrets management, service discovery, and container orchestration systems
  • Performing code review for internal services and infrastructure changes
  • Conducting experiments and research spikes for proposed changes
  • Working with R&D and architecture teams on defects and runtime inefficiencies identified in the production environment
  • Working with development teams to enhance and improve system operability.
  • Defining, tracking, and maintaining Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets in partnership with development teams
  • Building and maintaining observability pipelines and dashboards using Prometheus, Grafana, OpenTelemetry, and Splunk to ensure system health and performance visibility
  • Leading incident response and managing on-call escalations via PagerDuty; facilitating blameless post-mortems to drive systemic improvements

Requirements

  • BS or MS in Computer Science or similar discipline
  • Strong work ethic and personal drive to learn and continuously improve yourself, and the systems you own
  • Ability to perform under pressure in a fast-paced environment
  • 5+ years of experience working in a Linux environment
  • 5+ years of cloud computing administration and engineering (i.e. Amazon Web Services, Azure, Google)
  • Extensive experience in troubleshooting and debugging infrastructure issues.

Preferred Experience:

  • Experience in development with modern programming languages (i.e. Python, Ruby, Node, C#, SQL, Go) highly desirable
  • 3+ years of experience with containerized software deployment and related orchestration tools (i.e. Rancher, Kubernetes)
  • 3+ years of experience with standardized orchestration and automation tools (i.e. Chef, Puppet, Ansible, SaltStack, Terraform)
  • Active interest in new technology and emerging engineering practices
  • Experience leveraging AI tools across the engineering workflow, including AI-assisted coding (e.g., GitHub Copilot, Cursor), AIOps and observability platforms (e.g., PagerDuty AIOps), and LLM/AI service integrations; a genuine curiosity for how emerging AI capabilities can improve reliability and operational efficiency.

Benefits & conditions

  • Flexibility: Hybrid Work Model and Business Casual Dress Code, including jeans
  • Your Future: 401k Matching Program, Professional Development Reimbursement
  • Work/Life Balance: Flexible Personal/Vacation Time Off, Sick Leave, Paid Holidays
  • Your Wellbeing: Medical, Dental, Vision, Employee Assistance Program, Parental Leave
  • Wide Ranging Perspectives: Committed to Celebrating the Variety of Backgrounds, Talents and Experiences of Our Employees
  • Training: Hands-On, Team-Customized, including SS&C University
  • Extra Perks: Discounts on fitness clubs, travel and more!, Unless explicitly requested or approached by SS&C Technologies, Inc. or any of its affiliated companies, the company will not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services.

SS&C Technologies offers a comprehensive total rewards package designed to support your wellbeing, growth, and future. Our benefits include medical, dental, and vision coverage; a 401(k) plan with company match; paid time off, holidays, and parental leave; and professional development reimbursement opportunity.

Actual base salary will vary based on several factors, including but not limited to relevant skills, prior experience, education, demonstrated performance, and geographic location.

Massachusetts: The expected base salary for the position is between 90000 USD to 110000 USD.

About the company

SS&C is a leading provider of mission-critical, AI-powered technology and services empowering financial services and healthcare organizations to work smarter, faster, and securely. Founded in 1986, SS&C is headquartered in Windsor, Connecticut, and has offices worldwide. More than 23,000 financial services and healthcare organizations, from the world’s largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:26 min

Understanding Puppeteer and its underlying architectural design

Miki Lombardi · JS Congress

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

2:19 min

Applying code assistant capabilities to infrastructure and cloud operations

Ryan J Salva · Coffee With Developers

Videos

See all

Related articles

See all