# Senior Backend Engineer, Inference Platform at Together AI

- Company: Together AI
- What the company does: Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research. Backed by General Catalyst, Kleiner Perkins and NEA.
- Company website: https://www.together.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: San Francisco
- Work setup: On-site
- Pay: $200K to $290K base salary per year (USD)
- Posted: 2025-08-22
- Apply by: 2026-10-08
- Apply: https://job-boards.greenhouse.io/togetherai/jobs/4835763007
- Page: https://www.1752.vc/careers/jobs/together-ai-senior-backend-engineer-inference-platform/

## About the role

Together AI is building the Inference Platform that brings the most advanced generative AI models to the world. Our platform powers multi-tenant serverless workloads and dedicated endpoints, enabling developers, enterprises, and researchers to harness the latest LLMs, multimodal models, image, audio, video, and speech models at scale.

## What they're looking for

- 5+ years of demonstrated experience building large-scale, fault-tolerant, distributed systems and API microservices
- Strong background in designing, analyzing, and improving efficiency, scalability, and stability of complex systems
- Excellent understanding of low-level OS concepts: multi-threading, memory management, networking, and storage performance
- Expert-level programming in one or more of: Rust, Go, Python, or TypeScript
- Knowledge of modern LLMs and generative models and how they are served in production is a plus
- Experience working with the open source ecosystem around inference is highly valuable, familiarity with SGLang, vLLM, or NVIDIA Dynamo will be especially handy

Tags: Engineering
