DevOps Platform Engineer

Rough House Games, Inc.
San Francisco, CA, United States
24 days ago
Apply on www.workingnomads.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Compensation
$140,000.0 - $180,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Application Performance Management Automation of Tests Microsoft Azure Backup Devices Bash Shell Software as a Service Cloud Computing Cloud Engineering Software Quality Continuous Integration
+26 more
Software Debugging DevOps Disaster Recovery Distributed Systems Github Python (Programming Language) Open Source Technology Runbook Web Applications Data Logging Scripting Google Cloud Load Balancing Performance Testing Delivery Pipeline Software Security Multi-Cloud Amazon Virtual Private Cloud (VPC) Cloudformation AI Platforms Gitlab-ci Kubernetes Infrastructure Automation Frameworks Deployment Automation Terraform Docker

Job description

As a Senior DevOps Platform Engineer, you will own the deployment backbone that makes Salt’s multi-environment AI platform reliable, fast, and secure. You’ll design and operate infrastructure that runs across SaaS, VPC, and on-prem-supporting customers who depend on Salt for mission-critical life sciences workloads.

The Role (What You’ll Be Doing):

  • Drive velocity by building and maintaining high-speed, reliable CI/CD pipelines.
  • Own multi-cloud deployment consistency across SaaS, VPC, and on-prem environments (GCP, AWS, Azure), ensuring parity and predictable behavior.
  • Define and enforce infrastructure and deployment standards to prevent environment drift.
  • Implement and evolve monitoring, logging, and alerting to ensure platform health, performance, and SLOs.
  • Automate provisioning and configuration of infrastructure using infrastructure-as-code.
  • Design and maintain application security, backup, and disaster recovery procedures.
  • Partner closely with developers to improve application performance, reliability, and operability.
  • Participate in incident response, including on-call rotations, root cause analysis, and remediation.
  • Maintain clear, discoverable documentation for infrastructure, deployment processes, and runbooks.
  • Support QA by enabling automated testing in pipelines and stable test environments.
  • Champion and spread best practices for code quality, security, and deployment standards.

Requirements

  • 5+ years of experience operating production cloud infrastructure (GCP and/or AWS), ideally in multi-cloud environments.
  • Deep experience with CI/CD systems and pipeline design (e.g., GitLab CI, GitHub Actions).
  • Strong experience with Kubernetes and container orchestration in production.
  • Hands-on experience with infrastructure-as-code tools (Terraform, CloudFormation, or similar).
  • Proficiency with Docker and core containerization concepts.
  • Experience implementing and operating observability stacks (metrics, logs, traces).
  • Solid understanding of web application deployment, scaling, and reliability fundamentals.
  • Scripting skills in Python, Bash, or similar for automation.
  • Experience integrating automated tests into deployment workflows.
  • Strong problem-solving and debugging skills, especially in distributed systems.
  • Clear, concise communication and a collaborative working style.

Preferred / Bonus Skills (Nice To Have):

  • Experience with database administration, backup, and recovery strategies.
  • Familiarity with security best practices for cloud-native and multi-tenant applications.
  • Experience with performance testing, capacity planning, and cost optimization.
  • Strong understanding of networking, load balancing, and traffic routing.
  • Experience with AI/ML, HPC, or other data- and compute-intensive workloads.
  • Contributions to internal DevOps/Platform tooling or open-source automation projects.

What Makes You A Great Fit:

  • You’re motivated by giving developers a platform that “just works” so they can ship safely and fast.
  • You default to automation, repeatability, and simplicity over manual, one-off fixes.
  • You’re comfortable working across SaaS, customer VPCs, and on-prem environments and enjoy the challenge of standardizing them.
  • You naturally validate your work end-to-end-using the application, not just watching dashboards.
  • You can explain complex infrastructure tradeoffs to any audience and drive alignment.
  • You’re excited to apply your skills to AI in life sciences and support teams doing high-impact scientific work.

Benefits & conditions

  • Competitive salary based on location and experience.
  • Comprehensive health coverage, including medical, dental, mental health, and 401(k).
  • Fully remote role, collaborating primarily within U.S. time zones.

About the company

Salt AI is building the intelligence layer for the enterprise, starting with life sciences. In a world where AI is fragmented across models, tools, and data silos, Salt unifies them through a compliance-first orchestration platform that lets regulated organizations deploy AI securely, at scale, and with confidence.

We’re creating the backend systems that power real-world AI workflows across drug discovery, clinical development, and commercial operations. Our platform turns complex, regulated environments into intelligent, adaptive systems where autonomous AI agents operate in harmony with existing infrastructure., At Salt AI, we’re building more than just a product-we’re creating the systems that will power the future of AI development for life sciences. If you’re excited about designing and operating backend services that enable breakthrough scientific discoveries, we’d love to hear from you. We’re committed to building a diverse, inclusive team and encourage applications from candidates of all backgrounds and experiences.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.workingnomads.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all