Startups · AI

Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization

DiDi · San Jose, CA · On-site

← All jobs
About DiDi

DiDi Global is the world's leading mobile transportation platform offering a full range of app-based services to users around the world. Backed by Techstars.

About the role

We are seeking an experienced and mission-driven Senior/Sr. Staff AI Infrastructure Engineer , Inference & Optimization to lead the performance tuning, deployment, and resource scheduling of cutting-edge AI models across on-vehicle and cloud infrastructure. In this role, you will design high-efficiency inference pipelines, build system-level stability frameworks, and optimize hardware execution to ensure ultra-low latency and rock-solid operational reliability.

What they're looking for

  • Master’s or higher degree in Computer Science, Software Engineering, Systems Engineering, or a closely related technical field
  • 3-8+ years of industry experience in high-performance computing, AI infrastructure, model optimization, or embedded deployment
  • Strong proficiency in C++ and Python, with solid expertise in parallel programming (CUDA, OpenMP) and low-level system profiling tools
  • Deep familiarity with mainstream inference engines (e.g., TensorRT, ONNX Runtime) and specialized LLM inference/serving frameworks (e.g., vLLM, SGLang, TensorRT-LLM)
  • Practical understanding of modern GPU hardware architectures (e.g., NVIDIA Hopper, Thor) and memory bandwidth management
  • Demonstrated ability to diagnose complex software-hardware integration issues and drive scalable, production-grade solutions
More about this role

DiDi's autonomous driving unit was established in 2016 with the mission of developing Level 4 autonomous driving (AD) technology to make transportation safer and more efficient. In August 2019, the unit became an independent company, DiDi Autonomous Driving, dedicated to advanced AD R&D, product application, and business expansion. We believe integrating AD technology into a shared-mobility fleet will generate immense social value. By leveraging DiDi's specialized technology, operational expertise, and integrated ecosystem, we are positioned to build and operate a highly efficient, user-oriented autonomous fleet.

We are seeking an experienced and mission-driven Senior/Sr. Staff AI Infrastructure Engineer , Inference & Optimization to lead the performance tuning, deployment, and resource scheduling of cutting-edge AI models across on-vehicle and cloud infrastructure. In this role, you will design high-efficiency inference pipelines, build system-level stability frameworks, and optimize hardware execution to ensure ultra-low latency and rock-solid operational reliability. You will act as a technical leader in AI infrastructure, accelerating model iteration and bridging the gap...

Read the full posting on DiDi's site ↗

Artificial Intelligence

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.