Startups · AI

Staff Platform Engineer, Service Infrastructure

Together AI · San Francisco · On-site

← All jobs
About Together AI

Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research. Backed by General Catalyst, Kleiner Perkins and NEA.

About the role

Together AI is hiring a Staff Platform Engineer to join the Product Foundations engineering organization and drive its service infrastructure strategy.

What they're looking for

  • 7+ years of professional experience in platform engineering, service infrastructure, SRE, distributed systems, cloud infrastructure, or related roles
  • Deep production experience with Kubernetes, including EKS, Helm, ArgoCD/Argo Rollouts, ingress, autoscaling, secrets, service identity, networking, and progressive delivery
  • Strong Terraform experience, including module design, infrastructure CI/CD, policy enforcement, production applies, and safe self-service workflows
  • Experience operating networking and edge infrastructure such as CDNs, ALBs/NLBs, DNS, TLS, ingress/egress controls, and traffic management
  • Proficiency in one or more programming languages used for infrastructure tooling and automation, such as Go, Python, TypeScript, or similar
  • AWS experience, ideally including EKS, IAM, VPC networking, load balancing, Route 53, CloudFront, ECR, and related service infrastructure
More about this role

About the Role

Together AI is hiring a Staff Platform Engineer to join the Product Foundations engineering organization and drive its service infrastructure strategy.

Product Foundations builds and operates Together’s mission-critical product platforms that support all cloud products, including API Platform (non-Inference), web UI Platform, Billing, and customer-facing IAM. These services sit on the critical path for customers and internal systems.

This is a hands-on Staff role focused on evolving Product Foundations’ core infrastructure strategy from the inside: understanding service team needs, turning repeated infrastructure problems into reusable patterns, and coordinating across platform owners so Product Foundations services are reliable, repeatable, and built on the right company-wide foundations.

  • Own the technical direction for service infrastructure within Product Foundations, including Kubernetes, AWS, Terraform, CDNs, ALBs, DNS, IAM, service networking, and related operational patterns.
  • Up-level existing Product Foundations services by improving reliability, operability, deployment safety, infrastructure consistency, and production readiness.
  • Partner deeply with...

Read the full posting on Together AI's site ↗

Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.