Compresr is an LLM context-compression API: send long context plus your query, get back a shorter context that keeps the answer-bearing tokens and drops the rest. Cut token cost and latency; at light compression, match or beat full-context accuracy. Python... Backed by Y Combinator.
About the role
Hands-on supervision. We're around for brainstorming and high-level guidance, but you own your work and will be the person who knows it best. A well-defined project. We're early-stage and led by customer and market pull, so we work on several directions at once. You'll navigate this alongside the rest of us.
More about this role
Duration: 3 months, with a possibility of a full-time job afterwards
Start date: immediately
We're building state-of-the-art context compression. Our mission is to become the "Cloudflare for LLMs", a compression layer embedded into most LLM pipelines by default.
We're a team of ex-EPFL MSc/PhDs. We started by publishing papers, then got into YC and started making money helping companies cut their LLM costs.
We run the business like a research lab: form hypotheses, kill the ones that don't work, double down on the ones that do.
Competitive compensation
All the resources you need: GPUs, subscriptions, OpenAI/Anthropic credits
As much responsibility as you can handle. Our goal is to make you an irreplaceable part of the team
A fast-paced environment where you'll learn much faster than usual, surrounded by technical people who push each other
Possibility of a full-time offer based on performance
Hands-on supervision. We're around for brainstorming and high-level guidance, but you own your work and will be the person who knows it best.
A well-defined project. We're early-stage and led by customer and market pull, so we work on several directions at once. You'll navigate this alongside...
Browse similar: AI jobs · Startup internships · AI startup jobs · Startup jobs · Remote jobs