# ML Runtime Engineer at Loft

- Company: Loft
- What the company does: Backed by a16z.
- Company website: http://www.theloftcompanyla.com
- Type: Startups (AI role)
- Level: Mid level
- Location: Toulouse, Occitanie
- Work setup: On-site
- Posted: 2026-07-01
- Apply by: 2026-10-08
- Apply: https://jobs.lever.co/loftorbital/7bf12750-6b1f-4bac-a1ab-6709d13c60fa
- Page: https://www.1752.vc/careers/jobs/loft-ml-runtime-engineer/

## About the role

As our ML Runtime Engineer , you'll own model compilation and the performance tooling behind it. We're looking for someone scrappy — the kind of engineer who takes an unfamiliar model and an unforgiving power budget, digs into the graph, and gets it running on real hardware without waiting for the path to be handed to them. You'll move between model compilation, runtime internals, quantization, and hardware-in-the-loop CI, sometimes in the same afternoon.

## What they're looking for

- Strong C++ and Python
- Model compilation: TensorRT (and/or equivalent graph compilers)
- ONNX Runtime, quantization & inference perf optimization
- Embedded / edge GPU deployment ( NVIDIA Jetson)
- Benchmarking & profiling / perf tooling
- CI/CD incl. hardware-in-the-loop

Tags: AI Engineering
