Ship quality agents at scale. Braintrust is the AI observability platform for tracing production, running evals, and catching regressions before they reach users. Backed by Greylock, a16z and ICONIQ.
About the role
Our largest customers don't just use Braintrust — they run it. They deploy our stack inside their own AWS, Azure, and GCP accounts, behind their own VPCs, under their own compliance requirements, at their own scale. When a hybrid deployment stalls, when ingest backs up, when a query that was fast last week isn't, they come to us. Platform Support is the team that owns that. We're the technical front line for infrastructure, performance, and reliability.
What they're looking for
- Experience in a customer-facing technical role — Support Engineering, SRE, DevOps, Solutions Architecture, or Infrastructure Engineering — or backend/infra engineering experience with real appetite for customer work
- Strong Kubernetes fundamentals: you can deploy, debug, and scale actual workloads, and read a failing pod's story from its events and logs
- Hands-on Terraform, and depth in at least one major cloud (AWS strongly preferred)
- Comfort in a backend codebase — Python, TypeScript, or Go — enough to reproduce a bug, trace it to its source, and fix it
- Fluency with observability tooling, and the instinct to reach for data before opinion
- Clear, calm, direct communication under pressure, especially when the customer is technical, blocked, and losing time
More about this role
Braintrust is the agent observability platform. By actively applying intelligence to agent traces and automatically surfacing the most critical patterns, Braintrust gives teams the visibility to understand how agents behave in production and the tools to improve them.
Teams at Notion, Stripe, Box, OpenAI, and Cloudflare use Braintrust to trace their agents, find the issues in their observability data, and run evals that tell them how to improve.
Our largest customers don't just use Braintrust — they run it. They deploy our stack inside their own AWS, Azure, and GCP accounts, behind their own VPCs, under their own compliance requirements, at their own scale. When a hybrid deployment stalls, when ingest backs up, when a query that was fast last week isn't, they come to us. Platform Support is the team that owns that. We're the technical front line for infrastructure, performance, and reliability.
We're hiring Platform Support Engineers at both mid and senior levels to join a small, high-ownership team. You'll work shoulder to shoulder with our Cloud Infrastructure and Engineering teams, and alongside our Developer Support Engineers, who own the SDK and API side of the customer...
Browse similar: AI jobs · AI startup jobs · Startup jobs · Remote jobs · San Francisco Bay Area