Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. Backed by Insight and Sapphire.
About the role
Design and build LLM serving infrastructure on Kubernetes: deployment, GPU scheduling, scaling, and model lifecycle management. Package the platform for enterprise environments: Helm-based installs, upgrades, and restricted/offline networks.
What they're looking for
- 5+ years of software engineering experience in infrastructure, platform, or distributed systems
- Deep hands-on Kubernetes experience: building and operating production workloads and Helm charts, not just consuming managed clusters
- Experience with GPU workloads or LLM inference, or strong adjacent systems experience and a track record of learning fast
- Strong Go programming skills, solid CI/CD and infrastructure-as-code skills
- Fluency with AI-assisted development tools (Claude Code, OpenAI Codex) as part of your daily engineering workflow
- Comfortable with high autonomy on a small, remote-first, written-culture team
More about this role
Mirantis is building a new enterprise AI infrastructure product that lets organizations run and govern large language models on their own Kubernetes clusters. You will join a small senior team early, with broad ownership of the model-serving layer and its path to production.
Design and build LLM serving infrastructure on Kubernetes: deployment, GPU scheduling, scaling, and model lifecycle management.
Package the platform for enterprise environments: Helm-based installs, upgrades, and restricted/offline networks.
Integrate the serving layer with the platform's API gateway, identity, and metering services.
Build the observability for operating GPU inference in production (serving metrics, GPU telemetry).
Contribute across a multi-service codebase and help set engineering direction through design docs and reviews.
5+ years of software engineering experience in infrastructure, platform, or distributed systems.
Deep hands-on Kubernetes experience: building and operating production workloads and Helm charts, not just consuming managed clusters.
Experience with GPU workloads or LLM inference, or strong adjacent systems experience and a track record of learning fast.
Strong Go programming...
Browse similar: AI jobs · AI startup jobs · Startup jobs · Remote jobs