It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker, Phyllis. She cried tears of joy. Backed by Greylock and Sequoia.
About the role
For positions in this location, we offer a base pay of $221,200 - $387,100 , plus equity (when applicable), variable/incentive compensation and benefits. Sales positions generally offer a competitive On Target Earnings (OTE) incentive compensation structure. Please note that the base pay shown is a guideline, and individual total compensation will vary based on factors such as qualifications, skill level, competencies, and work location.
What they're looking for
- Proven success leading managers and senior technical leaders across geographically distributed engineering organizations
- Demonstrated success leading SRE, infrastructure, or reliability transformation at scale
- Strong understanding of SLIs/SLOs, error budgets, observability, incident management, reliability governance, and on-call practices
- Experience with service catalogs, service registries, service ownership models, Backstage, CMDB, dependency mapping, or service topology
- Strong background in cloud infrastructure and modernization across AWS, Azure, and/or GCP
- Understanding of Kubernetes, distributed systems, networking, databases, infrastructure automation, and cloud-native architecture
More about this role
Our Site Reliability Engineering (SRE) team consists of highly skilled engineers responsible for maintaining and enhancing the reliability, scalability, and performance of the ServiceNow infrastructure. Our SRE’s are empowered to resolve technical issues across the entire technology stack, from hardware to applications. Additionally, they work to improve the platform's operability, aiming to reduce the number of incidents and minimize Mean Time to Recovery (MTTR). To achieve this, the team combines software development, networking, database, and systems engineering skills to tackle complex problems, striving to maintain our platform operating for our customers.
We are looking for a Director of Site Reliability Engineering to lead the next phase of our reliability transformation as ServiceNow modernizes toward a cloud-agnostic, cloud-ready production platform.
This leader will own key elements of the SRE operating model across Reliability Engineering, Service Enablement, Service Registry, SLI/SLO standards, reliability governance, automation, AI-enabled operations, and production readiness . The role will lead a global engineering organization and partner across Product...
Browse similar: Startup jobs