Senior Cloud Engineer

Fandom

San Francisco (CA)

On-site

USD 122,000 - 204,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Vibrant team culture
Comprehensive medical/dental/vision
Training (Udemy + more)
Flexible hours and time off
Equity & retirement including 401K
Paid parental leave
International startup culture

Job summary

Fandom is seeking a Senior Cloud Engineer to scale and support infrastructure powering a platform used by hundreds of millions of fans worldwide. This hands-on role focuses on building reliable, scalable Linux‑ and Kubernetes‑based systems, with a strong emphasis on automation, CI/CD, and security.

You will work within the TechOps team, collaborating with developers and other infrastructure groups, helping ensure the platform remains fast, stable, and secure as Fandom grows.

Qualifications

  • 5+ years in Technical/Network Operations, DevOps, or SRE for large-scale production platforms

Responsibilities

  • Design, architect, and scale high-availability routing architectures across on-prem, AWS, and GCP
  • Manage IaC with Terraform and Chef for Kubernetes, Cloudflare, and CI/CD pipelines
  • Maintain and optimize large-scale production environments with MySQL replication, failover, DR, and monitoring
  • Develop internal tools for automation, backups, performance tuning, and security monitoring
  • Lead planning meetings, retrospectives, and RCAs to drive operational excellence
  • Participate in on-call rotation and mentor cross-functional teams

Skills

Linux administration
Networking protocols
Go/Python/Bash
Kubernetes
CI/CD (GitHub Actions/Jenkins)
MySQL replication/failover
AI tools for productivity

Education

Bachelor's degree or equivalent

Tools

Terraform
Chef
Cloud platforms (AWS, GCP)
Grafana/Prometheus

Job description

About This Role

Fandom is growing! We’re looking for a Senior Cloud Engineer to help evolve and support the infrastructure that powers our platform for over 300 million fans around the world. This is a hands‑on role focused on building reliable, scalable systems in a Linux and Kubernetes‑based environment.

As part of the TechOps team, you’ll report to the Manager of TechOps and work closely with developers, product engineers, and other infrastructure teams. You’ll contribute to our CI/CD, monitoring, automation, and cloud efforts — helping ensure Fandom’s platform remains fast, stable, and secure as we grow.

This is a great opportunity for someone who enjoys solving complex infrastructure challenges, improving deployment systems, and enabling engineering teams to move faster and safer.

You Will...
  • Design, architect, and scale high‑availability routing architectures to seamlessly balance and secure global user traffic across a hybrid footprint of on‑premise datacenters, AWS, and GCP.
  • Manage and automate cloud and edge infrastructure as code (IaC) using Terraform and Chef, ensuring consistent configurations for Kubernetes, global Cloudflare services, and CI/CD pipelines.
  • Maintain and optimize large‑scale production environments, orchestrating high‑availability MySQL replication, automated failover, robust disaster recovery, and system monitoring.
  • Develop internal tools for operational automation, system/data backups, performance tuning, and comprehensive security monitoring.
  • Drive operational excellence by leading planning meetings, retrospectives, and RCAs, while continuously evaluating systems against industry best practices.
  • Participate in an on‑call rotation to maintain production stability, handle incident responses, and collaborate with/mentor cross‑functional engineering teams.
You Have...
  • 5+ years of experience in Technical/Network Operations, DevOps, or SRE roles managing large‑scale production platforms (e.g., 10M+ monthly active users).
  • Deep proficiency in Linux systems administration, networking protocols (TCP/IP, routing), secure systems practices, and scripting/programming (Go, Python, or Bash).
  • Proven hands‑on experience with core infrastructure tech: Kubernetes/container orchestration, CI/CD pipelines (GitHub Actions, Jenkins), and monitoring/reliability systems (e.g., Prometheus).
  • Practical experience managing production MySQL database environments, including deep familiarity with replication topologies, failover mechanisms, and performance tuning.
  • Demonstrated capability using generative AI tools (e.g., Gemini, NotebookLM) to enhance productivity, paired with the ability to critically audit and verify outputs for accuracy, security, and context.
Bonus Points...
  • Advanced Cloudflare expertise, including CDN optimization, WAF security, DNS management, edge performance tuning, and Cloudflare Tunnels.
  • Strong understanding of distributed systems architecture, edge caching, and centralized log management using the ELK stack (Elasticsearch, Logstash, Kibana).
  • Experience defining SLOs and instrumentation, implementing meaningful metrics, logs, and traces to reduce alert noise and drive postmortem action items.
Benefits & Perks
  • Salary Range = $122k - $204k (Actual salary available will vary based on location and market factors.)
  • Vibrant team culture
  • Comprehensive Medical, Dental, Vision
  • Training (unlimited Udemy + more)
  • Flexible working hours and time off
  • Equity & Retirement Programs including 401K match
  • Paid Parental Leave
  • International work environment with start‑up culture
About Fandom

Fandom is the world’s largest fan platform where fans immerse themselves in imagined worlds across entertainment and gaming. Reaching more than 350 million unique visitors per month and hosting more than 250,000 wikis, Fandom is the #1 source for in‑depth information on pop culture, gaming, TV and film, where fans learn about and celebrate their favorite fandoms. Fandom’s Gaming division manages the online video game retailer Fanatical. Fandom Productions, the content arm of Fandom, enhances the fan experience through curated editorial coverage and branded content from trusted and established publishing brands Gamespot, TV Guide and Metacritic, along with its Emmy‑nominated Honest Trailers and the weekly video news program The Loop. For more information follow @getfandom or visit: www.fandom.com.

Fandom is an equal opportunity employer. Fandom values diversity, and all employment decisions are made on the basis of job requirements and individual qualifications.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud Engineer: Scale Global Infra & Automation
Senior Cloud Engineer: Scale Global Infra & Automation

Fandom • San Francisco (CA)

On-site
USD 122,000 - 204,000
Vibrant team culture
Comprehensive medical/dental/vision
Training (Udemy + more)
+4
Senior Software Engineer, Deploy at Scale
Senior Software Engineer, Deploy at Scale

CloudFlare • Austin (TX)

On-site
USD 180,000 - 300,000
Senior Software Engineer, Infrastructure Engineering
Senior Software Engineer, Infrastructure Engineering

Ellation, Inc. • Los Angeles (CA)

On-site
USD 183,000 - 229,000
Salary + Bonus
Flexible time off
Health insurance
+3
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Webhosting • San Francisco (CA)

On-site
USD 194,000 - 266,000
Equity plan
Health benefits
Comprehensive benefits package
Systems Engineer
Systems Engineer

Cloudflare • Austin (TX)

On-site
USD 140,000 - 190,000
Equity plan
Medical/Rx Insurance
Dental Insurance
+2
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

On-site
USD 194,000 - 266,000
Cloudflare benefits
Senior Cloud Engineer - NCF
Senior Cloud Engineer - NCF

Fanatics Inc • New York (NY)

On-site
USD 136,000 - 170,000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Cloudflare • San Francisco (CA)

On-site
USD 194,000 - 266,000
Health insurance
401(k) Retirement Savings Plan
Employee Stock Participation Plan
+2
Senior Manager, Forward Deployed Engineering
Senior Manager, Forward Deployed Engineering

Webhosting • New York (NY), Northern (KY)

On-site
USD 234,000 - 310,000
Equity plan
Comprehensive benefits
Principal Software Engineer: Distributed Systems (Config, Test, & Deployment)
Principal Software Engineer: Distributed Systems (Config, Test, & Deployment)

Cloudflare • New York (NY)

On-site
USD 220,000 - 275,000
Medical Insurance
Dental Insurance
Vision Insurance
+3