Startups · AI

Member of Technical Staff (GPU Performance Engineer)

Reka AI · US, UK, Singapore, Remote · Remote

← All jobs
About Reka AI

Reka AI builds natively multimodal models (Spark, Edge, Flash, Core) for video, image, audio, and text. Used by enterprises in security, media, and defense.

About the role

We are seeking an experienced GPU Performance Engineer with a strong background in Python and large-scale model training. In this role, you will design and implement improvements to our training infrastructure and directly contribute to technical decisions that optimize performance of our models. You will also work on post-training processes, including reinforcement learning and fine-tuning. Furthermore, you will contribute to improving the efficiency and scalability of our model serving infrastructure.

What they're looking for

  • Strong engineering skills with fluency in Python and PyTorch (or other frameworks)
  • Proven experience implementing and training large deep learning models
  • Experience writing and debugging low-level GPU code (CUDA, C++)
  • Experience scaling up GPU jobs using large-scale compute clusters (e.g., Slurm or Kubernetes)
  • Demonstrated ability to analyze and optimize the performance of GPU-accelerated workloads, including profiling, identifying bottlenecks, and implementing performance tuning techniques
  • Reka's Mission
More about this role

We are seeking an experienced GPU Performance Engineer with a strong background in Python and large-scale model training. In this role, you will design and implement improvements to our training infrastructure and directly contribute to technical decisions that optimize performance of our models. You will also work on post-training processes, including reinforcement learning and fine-tuning. Furthermore, you will contribute to improving the efficiency and scalability of our model serving infrastructure.

Strong engineering skills with fluency in Python and PyTorch (or other frameworks).

Proven experience implementing and training large deep learning models.

Experience writing and debugging low-level GPU code (CUDA, C++).

Experience scaling up GPU jobs using large-scale compute clusters (e.g., Slurm or Kubernetes).

Demonstrated ability to analyze and optimize the performance of GPU-accelerated workloads, including profiling, identifying bottlenecks, and implementing performance tuning techniques.

Reka's Mission

Reka's mission is to build useful multimodal artificial intelligence and use it to empower organizations and businesses. We are a globally distributed foundation model...

Read the full posting on Reka AI's site ↗

Technical

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.