AI Recruitment Software designed to source and hire candidates faster. Tailored for HR teams, recruitment agencies, and headhunters. Backed by Peak XV.
About the role
Manatal has a global presence and is trusted by thousands of businesses in over 135 countries. Our goal is to transform the entire hiring process by making it simple, efficient, and enjoyable for recruiters, hiring managers, and candidates alike. Our mission is to offer the best-in-class AI-powered technologies to empower small, medium, and large businesses in their staffing & recruitment transformation.
What they're looking for
- 5+ years of experience in Site Reliability Engineering, DevOps, or Production Engineering at a technology company
- Proven experience leading production incidents end-to-end: from initial triage through cross-functional coordination to resolution and post-mortem. This is the core of the role
- Strong working knowledge of Kubernetes in production, specifically AWS EKS. You must be able to troubleshoot pod failures, resource constraints, networking issues, and deployment problems independently
- Experience with Sentry or equivalent error monitoring platforms for application-level issue detection and triage
- Proven ability to design and manage on-call rotations, define escalation policies, and build alerting strategies that minimize noise while catching real issues
- Experience defining and tracking SLOs/SLIs to drive reliability decisions and prioritize engineering effort
More about this role
Manatal is an HRTech software service (B2B SaaS) company headquartered in Bangkok, Thailand. Manatal is one of the fastest-growing start-ups in the region and is backed by Surge and Sequoia Capital.
Manatal has a global presence and is trusted by thousands of businesses in over 135 countries. Our goal is to transform the entire hiring process by making it simple, efficient, and enjoyable for recruiters, hiring managers, and candidates alike. Our mission is to offer the best-in-class AI-powered technologies to empower small, medium, and large businesses in their staffing & recruitment transformation.
Manatal is establishing a dedicated Site Reliability Engineering function for the first time. As the platform and customer base grow, the need for structured incident response, mature observability, and reliable on-call operations requires dedicated ownership.
The Lead Site Reliability Engineer will be the first dedicated hire in this function. Working closely with the CTO, the Director of Engineering and the Director of Product, this role owns incident response, on-call operations, and the observability stack. When production incidents occur, this person leads. Between incidents, they...
Browse similar: Startup jobs · Remote jobs