# LLM Inference Engineer at Hippocratic AI

- Company: Hippocratic AI
- What the company does: Hippocratic AI builds the safest generative AI healthcare agent for health systems, payors, and pharma. Over 180 million clinical interactions across 1,000+ use cases with 60+ partners worldwide. Backed by General Catalyst, Kleiner Perkins and a16z.
- Company website: https://www.hippocraticai.com/
- Type: Startups (AI role)
- Level: Mid level
- Location: Menlo Park, CA
- Work setup: On-site
- Posted: 2026-08-22
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/hippocratic%20ai/eef8a721-23de-4c20-bff0-56088b39afa0
- Page: https://www.1752.vc/careers/jobs/hippocratic-ai-llm-inference-engineer/

## About the role

Design and implement multi-node serving architectures for distributed LLM inference Apply advanced quantization techniques (FP4/FP6) to reduce model footprint while preserving quality

## What they're looking for

- Experience optimizing LLM inference systems at scale
- Proven expertise with distributed serving architectures for large language models
- Hands-on experience implementing quantization techniques for transformer models
- Strong understanding of modern inference optimization methods, including:
- Speculative decoding techniques with draft models
- Eagle speculative decoding approaches

Tags: Research & Development
