The unified AI inference stack - from custom GPU kernels to production cloud serving on NVIDIA and AMD. 2x performance. Top open models. Open source stack. Backed by General Catalyst, Greylock and GV.
About the role
ML developers today face significant friction in taking trained models into deployment. They work in a highly fragmented space, with incomplete and patchwork solutions that require significant performance tuning and non-generalizable, model-specific enhancements. At Modular, we are building the next generation AI platform that will radically improve the way developers build and deploy AI models.
What they're looking for
- 7+ years of industry experience in compiler engineering, high-performance computing, kernel development, or related domains, with 2+ years managing engineers (strong IC-to-manager transitions welcome)
- Strong technical foundation in compilers, systems software, or accelerator programming
- Sufficient to be hands-on with the team's codebase, lead architecture discussions, and credibly review designs in MLIR/LLVM, kernel codegen, or runtime systems
- Proficiency in C++ and direct experience working in complex, multi-component software systems
- Hands-on experience with at least one heterogeneous programming model (CUDA, SYCL, OpenCL, or similar), as a contributor rather than only a user
- Demonstrated track record of delivering multi-quarter technical roadmaps with external dependencies, including managing scope, risk, and external commitments
More about this role
At Modular, a Qualcomm company , we’re on a mission to revolutionize AI infrastructure by systematically rebuilding the AI software stack from the ground up. Our team, made up of industry leaders and experts, is building cutting-edge, modular infrastructure that simplifies AI development and deployment. By rethinking the complexities of AI systems, we’re empowering everyone to unlock AI’s full potential and tackle some of the world’s most pressing challenges.
If you’re passionate about shaping the future of AI and creating tools that make a real difference in people’s lives, we want you on our team. You can read about our culture and careers to understand how we work and what we value.
ML developers today face significant friction in taking trained models into deployment. They work in a highly fragmented space, with incomplete and patchwork solutions that require significant performance tuning and non-generalizable, model-specific enhancements. At Modular, we are building the next generation AI platform that will radically improve the way developers build and deploy AI models. As part of our mission to build AI's unified compute layer, we are expanding the Modular software stack...
Browse similar: Startup jobs · Remote jobs