Prepare your taste buds... Backed by Lightspeed.
About the role
This role is crucial for simplifying the dining experience for students across the US. You will be instrumental in architecting resilient and self-healing solutions, managing AWS infrastructure, closing observability gaps, designing scaling approaches, and shaping incident management processes. Your contributions will span the entire development lifecycle, encompassing the building and maintenance of CI/CD pipelines. Found on 1752vc Careers, the job board for startup and VC roles.
What they're looking for
- Deep experience with Infrastructure as Code - Terraform (with Terraspace or a similar wrapper) - owning modules, state and rollout across multiple environments
- Kubernetes at operator depth: EKS lifecycle, Helm and Helmfile, controllers and add-ons, ingress/gateway, and autoscaling (HPA/KEDA)
- Multi-region AWS architecture, including failover design and the data-layer replication it depends on
- Deep knowledge of CI/CD tools (e.g., Jenkins, GitHub Actions)
- Software engineering experience in Python, Go, or a similar object-oriented language
- Proficiency with datastores (MySQL, MongoDB/Atlas, Redis/ElastiCache) and message brokers (RabbitMQ/AmazonMQ, SQS)
More about this role
This role is crucial for simplifying the dining experience for students across the US. You will be instrumental in architecting resilient and self-healing solutions, managing AWS infrastructure, closing observability gaps, designing scaling approaches, and shaping incident management processes. Your contributions will span the entire development lifecycle, encompassing the building and maintenance of CI/CD pipelines. Your role will be pivotal in ensuring the platform's scalability to support Grubhub's continuously expanding customer base, evidenced by the addition of 30 new campuses and a 25% year-over-year increase in order volume.
- Architect resilient, self-healing systems and co-own the design of critical production services.
- Own multi-region resilience: active-standby architecture, regional failover readiness, runbooks and drills, RTO/RPO targets, and the data-layer replication behind them (ElastiCache Global Datastore, MongoDB/Atlas, RDS).
- Own AWS infrastructure as code - Terraform/Terraspace, Helm and Helmfile - from design through rollout.
- Own the Kubernetes platform: EKS lifecycle, controller and add-on upgrades, ingress/gateway, and autoscaling (HPA/KEDA).
- Own...
Browse similar: Startup jobs · New York