Platform Operations Manager

Huxley Associates
Boston, MA, United States
4 days ago
Apply on www.huxley.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$180,000.0 - $210,000.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Cloud Computing Identity and Access Management Performance Tuning System Availability Kubernetes

Job description

This is a hands-on leadership role focused on the operations, stability, governance, and support of AWS and Kubernetes environments. Unlike platform engineering positions centered on building new features, this role is responsible for ensuring production systems remain reliable, secure, compliant, and highly available while leading a team of cloud/platform engineers.

The manager will oversee day-to-day cloud operations, Kubernetes administration, incident response, AWS governance, and service delivery. They will serve as the primary escalation point for complex operational issues, drive platform reliability, enforce cloud policies and guardrails, and support a high-volume, ticket-driven environment. The role also includes after-hours on-call responsibilities and close collaboration with engineering, security, and business stakeholders., * Lead day-to-day AWS and Kubernetes platform operations, ensuring high availability, performance, and stability.

  • Oversee Kubernetes cluster health, monitoring, upgrades, patching, troubleshooting, and performance optimization.
  • Drive SLA/SLO adherence and continuous improvements in service delivery and operational excellence.

Incident Management & Support

  • Serve as the primary escalation point for production incidents and complex operational issues.
  • Lead incident response, root cause analysis, and corrective action planning.
  • Manage a customer-focused, ticket-driven support environment with an emphasis on responsiveness and execution.

AWS Governance & Security

  • Define and enforce AWS governance standards, IAM policies, security controls, and compliance requirements.
  • Ensure proper cloud account structure, access management, and cost optimization practices.
  • Partner with security teams to proactively mitigate risk and strengthen platform security.

Leadership & Collaboration

  • Lead, mentor, and develop a team of Cloud/Platform Engineers.
  • Foster a culture of accountability, ownership, and operational excellence.
  • Collaborate with engineering, security, and business stakeholders to support platform needs and drive continuous improvement.

Requirements

  • 8+ years of cloud infrastructure or platform operations experience, including 2+ years in a leadership role.

Benefits & conditions

  • Deep hands-on experience managing and supporting production Kubernetes environments and AWS infrastructure.

EOE Statement: Specialist Staffing Group is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

In addition to base pay, direct-hire employees may be eligible for client offered benefits such as medical, dental, and vision coverage, and paid leave where required by applicable law. Eligibility may vary based on factors such as location and hire date and is subject to change.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.huxley.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · World Congress 2024

8:32 min

Benchmarking GitOps engine constraints for extensive multi-cluster environments

Artem Lajko · Europe 2026 Virtual

2:39 min

Modernizing operational excellence and cloud deployment tools

Mustafa Toroman · World Congress 2023

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

Videos

See all

Related articles

See all