Cloud Engineer

Fandoms Gaming
San Francisco, CA, United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$122,000.0 - $204,000.0
Working hours
Regular working hours
Job source

Tech stack

Bash Shell Computer Programming Continuous Integration DevOps Disaster Recovery Distributed Systems Domain Name System (DNS) Elasticsearch Github Monitoring of Systems Python (Programming Language) Linux System Administration
+17 more
MySQL Routing Network Protocols Performance Tuning Logstash Prometheus TCP/IP Trusted Systems Scripting Computer Network Operations Caching Kubernetes Cloudflare Kibana Terraform Elk Stack Jenkins

Job description

As part of the TechOps team, you’ll report to the Manager of TechOps and work closely with developers, product engineers, and other infrastructure teams. You’ll contribute to our CI/CD, monitoring, automation, and cloud efforts - helping ensure Fandom’s platform remains fast, stable, and secure as we grow.

This is a great opportunity for someone who enjoys solving complex infrastructure challenges, improving deployment systems, and enabling engineering teams to move faster and safer., * Manage and automate cloud and edge infrastructure as code (IaC) using Terraform and Chef, ensuring consistent configurations for Kubernetes, global Cloudflare services, and CI/CD pipelines.

  • Maintain and optimize large-scale production environments, orchestrating high-availability MySQL replication, automated failover, robust disaster recovery, and system monitoring.
  • Develop internal tools for operational automation, system/data backups, performance tuning, and comprehensive security monitoring.
  • Drive operational excellence by leading planning meetings, retrospectives, and RCAs, while continuously evaluating systems against industry best practices.
  • Participate in an on-call rotation to maintain production stability, handle incident responses, and collaborate with/mentor cross-functional engineering teams.

Requirements

  • 5+ years of experience in Technical/Network Operations, DevOps, or SRE roles managing large-scale production platforms (e.g., 10M+ monthly active users).
  • Deep proficiency in Linux systems administration, networking protocols (TCP/IP, routing), secure systems practices, and scripting/programming (Go, Python, or Bash).
  • Proven hands-on experience with core infrastructure tech: Kubernetes/container orchestration, CI/CD pipelines (GitHub Actions, Jenkins), and monitoring/reliability systems (e.g., Prometheus).
  • Practical experience managing production MySQL database environments, including deep familiarity with replication topologies, failover mechanisms, and performance tuning.
  • Demonstrated capability using generative AI tools (e.g., Gemini, NotebookLM) to enhance productivity, paired with the ability to critically audit and verify outputs for accuracy, security, and context., * Advanced Cloudflare expertise, including CDN optimization, WAF security, DNS management, edge performance tuning, and Cloudflare Tunnels.
  • Strong understanding of distributed systems architecture, edge caching, and centralized log management using the ELK stack (Elasticsearch, Logstash, Kibana).
  • Experience defining SLOs and instrumentation, implementing meaningful metrics, logs, and traces to reduce alert noise and drive postmortem action items.

Benefits & conditions

  • Salary Range = $122k - $204k (Actual salary available will vary based on location and market factors.)
  • Vibrant team culture
  • Comprehensive Medical, Dental, Vision
  • Training (unlimited Udemy + more)
  • Flexible working hours and time off
  • Equity & Retirement Programs including 401K match
  • Paid Parental Leave
  • International work environment with start-up culture

About the company

Fandom is growing! We’re looking for a Senior Cloud Engineer to help evolve and support the infrastructure that powers our platform for over 300 million fans around the world. This is a hands-on role focused on building reliable, scalable systems in a Linux and Kubernetes-based environment., Fandom is the world’s largest fan platform where fans immerse themselves in imagined worlds across entertainment and gaming. Reaching more than 350 million unique visitors per month and hosting more than 250,000 wikis, Fandom is the #1 source for in-depth information on pop culture, gaming, TV and film, where fans learn about and celebrate their favorite fandoms. Fandom’s Gaming division manages the online video game retailer Fanatical. Fandom Productions, the content arm of Fandom, enhances the fan experience through curated editorial coverage and branded content from trusted and established publishing brands Gamespot, TV Guide and Metacritic, along with its Emmy-nominated Honest Trailers and the weekly video news program The Loop. For more information follow @getfandom or visit: www.fandom.com.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:21 min

Deploying a primary Elasticsearch and Kibana cluster configuration

Philipp Krenn · World Congress 2022

1:44 min

Career transition into cloud native and data management

Michael Cade · LIVE

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · World Congress 2026 Europe

3:46 min

Navigating a career in cloud transformation consulting

Piet Van Dongen · LIVE

Videos

See all

Related articles

See all