Startups · AI

Machine Learning Engineer

Sift · San Francisco, California · Remote

← All jobs
About Sift

Sift's fraud prevention platform stops payment fraud and account takeover in real time - with transparent workflows, global data and expert analysts. Backed by Insight.

About the role

As a Machine Learning Engineer at Sift, you will bridge the gap between data science and large-scale distributed systems. You won’t just train models in isolation; you will build end-to-end pipelines that extract signals, train custom models per merchant, and serve predictions at production scale with low latency. You will work on an automated machine learning ecosystem that dynamically recalibrates models based on streaming global telemetry data.

What they're looking for

  • Experience: 4+ years of professional experience building and deploying large-scale machine learning models into high-traffic production environments
  • Solid Programming Foundations: Strong proficiency in Java or Scala (for our production backend) as well as Python (for data analysis and model prototyping)
  • Distributed Systems & Big Data: Practical experience with Databricks and big data processing frameworks like Apache Spark , Apache Flink , or Hadoop, and working with NoSQL data stores like Bigtable
  • Strong Mathematical Foundations: Deep understanding of statistical modeling, probability, and standard machine learning algorithms (e.g., XGBoost, Random Forests, Neural Networks, and Clustering techniques)
  • System Design Mentality: Ability to reason through data consistency, pipeline failures, and performance constraints in a distributed, multi-tenant cloud environment (GCP)
  • Experience explicitly in the fraud detection, risk mitigation, or cyber-security domains
More about this role

As a Machine Learning Engineer at Sift, you will bridge the gap between data science and large-scale distributed systems. You won’t just train models in isolation; you will build end-to-end pipelines that extract signals, train custom models per merchant, and serve predictions at production scale with low latency. You will work on an automated machine learning ecosystem that dynamically recalibrates models based on streaming global telemetry data.

Model Development & Refinement: Design, build, and deploy online machine learning models (including ensemble methods, deep learning, transformer architectures and graph-based models) to catch evolving fraud vectors in real time.

Feature Engineering at Scale: Engineer high-frequency time-series features from over 1 trillion behavioral events, optimizing for low-latency signal extraction and pattern recognition.

Production MLOps: Maintain and enhance our automated model training and deployment infrastructure, ensuring frictionless continuous integration and continuous deployment (CI/CD) of newly trained models.

System Optimization: Write high-performance code to minimize scoring latency at runtime, ensuring our core ML services scale...

Read the full posting on Sift's site ↗

Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.