Backed by Bessemer.
About the role
Fandom is growing! We’re looking for a Senior Cloud Engineer to help evolve and support the infrastructure that powers our platform for over 300 million fans around the world. This is a hands-on role focused on building reliable, scalable systems in a Linux and Kubernetes-based environment.
What they're looking for
- 5+ years of experience in Technical/Network Operations, DevOps, or SRE roles managing large-scale production platforms (e.g., 10M+ monthly active users)
- Deep proficiency in Linux systems administration, networking protocols (TCP/IP, routing), secure systems practices, and scripting/programming (Go, Python, or Bash)
- Proven hands-on experience with core infrastructure tech: Kubernetes/container orchestration, CI/CD pipelines (GitHub Actions, Jenkins), and monitoring/reliability systems (e.g., Prometheus)
- Practical experience managing production MySQL database environments, including deep familiarity with replication topologies, failover mechanisms, and performance tuning
- Demonstrated capability using generative AI tools (e.g., Gemini, NotebookLM) to enhance productivity, paired with the ability to critically audit and verify outputs for accuracy, security, and context
More about this role
Fandom is growing! We’re looking for a Senior Cloud Engineer to help evolve and support the infrastructure that powers our platform for over 300 million fans around the world. This is a hands-on role focused on building reliable, scalable systems in a Linux and Kubernetes-based environment.
As part of the TechOps team , you’ll report to the Manager of TechOps and work closely with developers, product engineers, and other infrastructure teams. You’ll contribute to our CI/CD, monitoring, automation, and cloud efforts — helping ensure Fandom’s platform remains fast, stable, and secure as we grow.
This is a great opportunity for someone who enjoys solving complex infrastructure challenges, improving deployment systems, and enabling engineering teams to move faster and safer.
- Design, architect, and scale high-availability routing architectures to seamlessly balance and secure global user traffic across a hybrid footprint of on-premise datacenters, AWS, and GCP.
- Manage and automate cloud and edge infrastructure as code (IaC) using Terraform and Chef, ensuring consistent configurations for Kubernetes, global Cloudflare services, and CI/CD pipelines.
- Maintain and optimize...
Browse similar: Startup jobs · San Francisco Bay Area