Site Reliability Engineer

SecurityScorecard
United States
28 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$152,000.0 - $195,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Automation of Tests Bash Shell Cloud Computing Continuous Integration DevOps Github Python (Programming Language) Octopus Deploy Open Source Technology Reliability Engineering Prometheus
+14 more
Datadog Policy as Code Pulumi Large Language Models Grafana Gitlab-ci Kubernetes Apache Flink Apache Kafka Machine Learning Operations Vertica Terraform Jenkins Vulnerability Analysis

Job description

As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD systems. You will also own the infrastructure behind our AI tooling - building MCP servers and defining safe, auditable AI access patterns for production systems. You’ll work hands-on with engineering teams to accelerate delivery, ensure production reliability, and embed best practices for automation, observability, and resilience., * Design, build, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications.

  • Build and operate AI tooling infrastructure - stand up MCP servers and establish secure, governed AI access and guardrails for production systems.
  • Optimize and maintain CI/CD pipelines, improving reliability, speed, and rollback safety.
  • Implement progressive delivery strategies such as blue/green and canary deployments.
  • Advance Infrastructure as Code with Terraform, Helm, and Argo CD, defining reusable patterns for the org.
  • Operate and optimize streaming and analytics infrastructure: Kafka, Flink, and ClickHouse.
  • Build automated testing into the CI/CD lifecycle.
  • Improve system observability - define SLOs, alerts, and dashboards.
  • Lead incident response and postmortems, focusing on root cause and durable fixes.
  • Mentor engineers across teams on Kubernetes, CI/CD, and cloud infrastructure.

Requirements

  • 6+ years in SRE, DevOps, or Infrastructure roles, with significant production Kubernetes experience.
  • Hands-on experience integrating AI/LLM tooling into engineering or operational workflows (e.g., MCP servers, AI agents acting on infrastructure), and a clear grasp of the security and governance considerations of giving AI access to production.
  • Proven success building CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI, or similar).
  • Strong with Kubernetes internals and managed services like EKS, GKE, or AKS.
  • Expertise with Infrastructure as Code (Terraform, Helm, Pulumi) and GitOps.
  • Proficient in Python, Bash, or Go.
  • Knowledge of observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry).
  • Production experience with Kafka, Flink, and ClickHouse.
  • Strong communication and cross-team collaboration skills., * Multi-region or multi-cluster Kubernetes experience.
  • Chaos engineering or resilience testing.
  • Security scanning, compliance automation, or policy-as-code.
  • LLM observability/tracing tooling (Langsmith, Langfuse) or MLOps workflows.
  • Contributions to open-source Kubernetes or CI/CD projects.

Benefits & conditions

Design, build, and scale Kubernetes-based, multi-tenant infrastructure and CI/CD systems. Own AI tooling infrastructure (MCP servers) and secure AI access patterns. Optimize CI/CD, streaming analytics (Kafka, Flink, ClickHouse), observability, and incident response. Implement IaC (Terraform, Helm, Pulumi), GitOps (Argo CD), progressive delivery, automated testing, and mentor engineering teams. The summary above was generated by AI, Specific to each country, we offer a competitive salary, stock options, Health benefits, and unlimited PTO, parental leave, tuition reimbursements, and much more!

The estimated total compensation range for this position is $152,000 - $195,000 (base plus bonus). Actual compensation for the position is based on a variety of factors, including, but not limited to affordability, skills, qualifications and experience, and may vary from the range. In addition to base salary, employees may also be eligible for annual performance-based incentive compensation awards and equity, among other company benefits.

SecurityScorecard is committed to Equal Employment Opportunity and embraces diversity. We believe that our team is strengthened through hiring and retaining employees with diverse backgrounds, skill sets, ideas, and perspectives. We make hiring decisions based on merit and do not discriminate based on race, color, religion, national origin, sex or gender (including pregnancy) gender identity or expression (including transgender status), sexual orientation, age, marital, veteran, disability status or any other protected category in accordance with applicable law.

About the company

SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies continuously rated, operating in 64 countries. Founded in 2013 by security and risk experts Dr. Alex Yampolskiy and Sam Kassoumeh and funded by world-class investors, SecurityScorecard’s patented rating technology is used by over 25,000 organizations for self-monitoring, third-party risk management, board reporting, and cyber insurance underwriting; making all organizations more resilient by allowing them to easily find and fix cybersecurity risks across their digital footprint.

Headquartered in New York City, our culture has been recognized by Inc Magazine as a “Best Workplace,” by Crain’s NY as a “Best Places to Work in NYC,” and as one of the 10 hottest SaaS startups in New York for two years in a row. Most recently, SecurityScorecard was named to Fast Company’s annual list of the World’s Most Innovative Companies for 2023 and to the Achievers 50 Most Engaged Workplaces in 2023 award recognizing “forward-thinking employers for their unwavering commitment to employee engagement.” SecurityScorecard is proud to be funded by world-class investors including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt ¡ LIVE

1:55 min

Contrasting Terraform with Pulumi and cloud-specific tools

Devlin Duldulao ¡ LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters ¡ WWC 2023

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum ¡ WWC Europe 2026

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 ¡ Coffee With Developers

3:20 min

Overview of infrastructure as code tools

Alexander Bubeck ¡ WWC 2023

Videos

See all

Related articles

See all