It’s time you own your health. Function includes 160+ lab tests and personalized protocols for instant action. Tracked over time in one secure place. Backed by Battery, General Catalyst and a16z.
About the role
You are an experienced data engineer who enjoys building systems that scale and last. You’re comfortable working across a wide range of data types and tools, and you approach complex problems with curiosity and pragmatism. You’re collaborative, thoughtful, and proactive, with a strong sense of ownership and care for the people who depend on the data you build. Ideally, you come from data engineering or software background in healthcare, biotech, or other high-trust data environments.
What they're looking for
- 1-2+ years of experience in data engineering, ETL development, or cloud-based data orchestration. 1-2+ years of experience in data science, analytics, or applied research, IF hold a PHD/published research
- Strong proficiency in Python and SQL
- Hands-on experience with workflow orchestration tools (Airflow, Prefect, or similar)
- Experience with cloud platforms and services (AWS, Azure, GCP familiarity is a plus)
- Experience building on Databricks and working with distributed processing frameworks (Apache Spark, Dask, or similar)
- Solid understanding of data validation, observability, testing, and production best practices
More about this role
Function Health’s mission is to empower people to live longer, healthier lives through proactive, data-driven healthcare. We aggregate and analyze comprehensive health data - including imaging, laboratory results, longitudinal biomarkers, and other health signals - to deliver clear, clinically meaningful insights.
We believe individuals should understand their health deeply, in context, and over time. Our platform brings together complex medical data and transforms it into trustworthy, actionable information that supports better health decisions.
As a Data Engineer , you will design, build, and scale the core data infrastructure that powers Function Health’s imaging, analytics, and AI ecosystems. You’ll work closely with the Data, AI, and R&D teams to orchestrate secure, reliable, and compliant pipelines across a wide range of healthcare data types.
This role focuses on building robust orchestration and cloud-native data systems (Airflow, AWS, Databricks) that support high-volume, heterogeneous datasets — from DICOM imaging to biomarkers, reports, and other structured and unstructured health data. You’ll help ensure that these systems are performant, scalable, and production-ready...
Browse similar: Startup jobs · Remote jobs