Startups

Data Pipeline Engineer

Cylake · Sunnyvale · On-site

← All jobs
About Cylake

Cylake is the complete, AI-native cybersecurity platform that provides full visibility, sets fine-grained policy, and stops attacks in fully air-gapped sovereign environments. Backed by Greylock and Lightspeed.

About the role

Join a small team building the next generation of cybersecurity products from the ground up. Led by industry veterans with a proven track record of success - you will get to architect, build, and deliver hugely impactful products with this world-class team. You will have the opportunity to grow your career and skills along with the company from the very start.

What they're looking for

  • A proven track record of success architecting, building, and running large-scale data systems (PB scale)
  • Experience with both batch and real-time processing architectures
  • Experience with open source Data Lakehouse components, including Apache Iceberg, PostgreSQL, Neo4j, Apache Parquet, etc
  • Experience with tools for stream processing and data analytics - Apache Kafka, Spark, Flink, etc
  • Understanding and experience with data transformation solutions
  • Excellent programming experience with Python
More about this role

Join a small team building the next generation of cybersecurity products from the ground up. Led by industry veterans with a proven track record of success - you will get to architect, build, and deliver hugely impactful products with this world-class team. You will have the opportunity to grow your career and skills along with the company from the very start.

Design, build, and maintain a scalable, open-source data lakehouse architecture supporting petabyte-scale analytics workloads. Responsible for architecting end-to-end data pipelines from ingestion through transformation to consumption, ensuring high performance, reliability, and data quality.

A proven track record of success architecting, building, and running large-scale data systems (PB scale)

Experience with both batch and real-time processing architectures

Experience with open source Data Lakehouse components, including Apache Iceberg, PostgreSQL, Neo4j, Apache Parquet, etc.

Experience with tools for stream processing and data analytics - Apache Kafka, Spark, Flink, etc.

Understanding and experience with data transformation solutions

Excellent programming experience with Python

Strong communication and documentation...

Read the full posting on Cylake's site ↗

R&D

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.