# ML Engineer, Inference & Optimization at Pika

- Company: Pika
- What the company does: AI creative tools. Built for Creatives. Backed by Lightspeed and AI Grant.
- Company website: https://pika.art/
- Type: Startups (AI role)
- Level: Mid level
- Location: Palo Alto HQ
- Work setup: On-site
- Pay: $250K to $350K base salary per year (USD)
- Posted: 2026-06-23
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/pika/fb9e43d7-36b8-46c7-93cd-2b7504e30363
- Page: https://www.1752.vc/careers/jobs/pika-ml-engineer-inference-and-optimization/

## About the role

We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.

## What they're looking for

- Experience : 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale
- Inference Mastery : Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks
- GPU & Parallelism : Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference
- AI Domain Knowledge : Familiarity with video generation (videogen) models and large language models (LLMs)
- Collaboration : Strong cross-discipline communication skills, able to drive shared goals across research and engineering functions
- Ownership Mindset : Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment

Tags: Research
