Backed by 500 Global.
About the role
You’ll work closely with our experienced engineers to design, build, and scale infrastructure for serving top open-source AI models. This role is ideal for recent graduates or junior engineers who want to grow quickly while working on high-impact, real production AI systems. If you’re excited about AI/ML, have taken related courses or built projects, and want to learn how to ship things at scale - we’d love to meet you.
What they're looking for
- Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field (completed or in final year)
- Strong fundamentals in data structures, algorithms, and software design
- Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
- Experience with AI/ML through coursework, research, personal projects, full-time employment, or internships
- Familiarity with AI models, Transformers and Diffusers
- Experience with version control systems (e.g., Git) and agile development methodologies
More about this role
DeepInfra is building the infrastructure layer for the next generation of AI. We believe open-source models are the future, and companies should have full control over their AI stack without being locked into proprietary providers.
Our inference platform serves trillions of tokens every week across hundreds of production workloads. We build everything from GPU infrastructure to the API layer because every millisecond matters.
We are looking for early-career Software Engineers (0-2 years of experience, including internships) to join our team.
You’ll work closely with our experienced engineers to design, build, and scale infrastructure for serving top open-source AI models. This role is ideal for recent graduates or junior engineers who want to grow quickly while working on high-impact, real production AI systems.
If you’re excited about AI/ML, have taken related courses or built projects, and want to learn how to ship things at scale - we’d love to meet you.
- Collaborate with engineers to design, develop, and test inference solutions for state-of-the-art AI models.
- Implement, optimize, and evaluate AI models using Python, C++, CUDA, and NCCL (previous exposure helpful - deep...
Browse similar: Startup jobs · San Francisco Bay Area