Startups · AI

Senior Software Engineer — Distributed Compute / Spark Systems

Granica · Bay Area Office · Remote

← All jobs
About Granica

Granica reduces the cost of enterprise AI across data storage, data processing, and agent compute. Everything runs inside your perimeter, and you keep the intelligence your data builds. Backed by NEA.

About the role

Granica is hiring a Senior Software Engineer to build distributed compute systems for enterprise-scale data and AI workloads. You will work on the core infrastructure behind Crunch, Granica’s continuous optimization product for enterprise lakehouse data. This includes systems for distributed execution, workload optimization, query performance, scheduling, resource management, and compute cost reduction across petabyte- and exabyte-scale environments.

What they're looking for

  • Strong engineering depth in distributed systems, data processing systems, query engines, databases, or cloud infrastructure
  • Production experience with distributed compute or query systems such as Apache Spark, Spark SQL, Trino, Presto, Flink, Databricks, EMR, Glue, Hive, or similar systems
  • Hands-on experience improving performance, reliability, or cost efficiency for large-scale data-processing workloads
  • Understanding of distributed execution, query planning, scheduling, resource management, fault tolerance, and workload isolation
  • Experience with Spark internals, Spark SQL, Catalyst, Adaptive Query Execution, shuffle, joins, aggregation, spill, memory management, or task scheduling
  • Familiarity with lakehouse formats and columnar data such as Iceberg, Delta Lake, Hudi, Parquet, or ORC
More about this role

Granica builds AI infrastructure for enterprises operating massive data environments.

Our platform helps data and engineering teams reduce storage and compute costs, improve performance and reliability, and prepare large datasets for analytics and AI.

Crunch — continuous optimization for enterprise lakehouse data

Myelin — stateful infrastructure for long-running AI agents

Large Tabular Models — foundation models designed for enterprise tables

Together, we are building the infrastructure that enables enterprises to own their data, own the intelligence built on it, and scale both efficiently.

Granica has demonstrated approximately $200K in annualized value per petabyte and verified customer value within weeks.

Granica is hiring a Senior Software Engineer to build distributed compute systems for enterprise-scale data and AI workloads.

You will work on the core infrastructure behind Crunch, Granica’s continuous optimization product for enterprise lakehouse data. This includes systems for distributed execution, workload optimization, query performance, scheduling, resource management, and compute cost reduction across petabyte- and exabyte-scale environments.

You will own core systems...

Read the full posting on Granica's site ↗

Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.