368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$160k – $240k per year
Location
In office
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Granica is an enterprise AI efficiency platform designed to reduce storage, compute, and context costs for large-scale AI applications. Founded in 2019 and headquartered in Mountain View, California, the company operates within its customers' own cloud perimeters (AWS and Google Cloud) to optimize enterprise data and model workflows securely without exposing sensitive information.

Senior Software Engineer - Distributed Compute / Spark Systems

Location: Mountain View, CA - On-site

About Granica

Granica builds AI infrastructure for enterprises operating massive data environments.

Our platform helps data and engineering teams reduce storage and compute costs, improve performance and reliability, and prepare large datasets for analytics and AI.

Granica’s products include:

  • Crunch - continuous optimization for enterprise lakehouse data

  • Myelin - stateful infrastructure for long-running AI agents

  • Large Tabular Models - foundation models designed for enterprise tables

Together, we are building the infrastructure that enables enterprises to own their data, own the intelligence built on it, and scale both efficiently.

Granica has demonstrated approximately $200K in annualized value per petabyte and verified customer value within weeks.

About the Role

Granica is hiring a Senior Software Engineer to build distributed compute systems for enterprise-scale data and AI workloads.

You will work on the core infrastructure behind Crunch, Granica’s continuous optimization product for enterprise lakehouse data. This includes systems for distributed execution, workload optimization, query performance, scheduling, resource management, and compute cost reduction across petabyte- and exabyte-scale environments.

You will own core systems that directly affect customer compute spend, query latency, workload reliability, cluster efficiency, and the performance of large-scale analytical data processing.

This is a hands-on engineering role for someone who has deep systems experience and wants to build at the intersection of Spark, distributed query execution, lakehouse compute, workload scheduling, storage-aware optimization, and AI infrastructure.

You will work on distributed compute systems involving Apache Spark, Spark SQL, Trino, Presto, Flink, Databricks, Snowflake-adjacent environments, cloud object stores, and lakehouse formats such as Apache Iceberg, Delta Lake, and Apache Hudi.

What You’ll Do

  • Build distributed compute systems for large-scale analytical and AI workloads

  • Improve performance and cost efficiency across Spark, Trino, Presto, Flink, Databricks, and Snowflake-adjacent environments

  • Design workload-aware systems for query execution, resource allocation, scheduling, and compute optimization

  • Optimize execution performance across joins, aggregations, scans, shuffles, spills, caching, partitioning, and task scheduling

  • Build systems that learn from workload patterns and automatically improve execution plans, cluster usage, and compute efficiency

  • Develop infrastructure for adaptive workload routing, execution planning, and data-processing reliability across large customer environments

  • Debug performance bottlenecks across query execution, metadata, storage, network, memory, CPU, and distributed compute layers

  • Work with lakehouse tables and columnar formats such as Iceberg, Delta Lake, Hudi, Parquet, and ORC to improve end-to-end workload performance

  • Build systems that reduce compute waste caused by inefficient scans, poor partitioning, small files, skew, unnecessary shuffles, and suboptimal workload placement

  • Improve reliability and failure recovery for large distributed data-processing jobs

  • Implement algorithms in workload optimization, execution efficiency, cost modeling, and data-processing performance

  • Contribute to open-source or publish research when appropriate

What We’re Looking For

  • Strong engineering depth in distributed systems, data processing systems, query engines, databases, or cloud infrastructure

  • Production experience with distributed compute or query systems such as Apache Spark, Spark SQL, Trino, Presto, Flink, Databricks, EMR, Glue, Hive, or similar systems

  • Hands-on experience improving performance, reliability, or cost efficiency for large-scale data-processing workloads

  • Understanding of distributed execution, query planning, scheduling, resource management, fault tolerance, and workload isolation

  • Experience with Spark internals, Spark SQL, Catalyst, Adaptive Query Execution, shuffle, joins, aggregation, spill, memory management, or task scheduling

  • Familiarity with lakehouse formats and columnar data such as Iceberg, Delta Lake, Hudi, Parquet, or ORC

  • Familiarity with cloud object storage systems such as S3, GCS, or ADLS and the performance tradeoffs of running distributed compute on top of them

  • Strong programming skills in Scala, Java, Go, Rust, C++, or similar systems-oriented languages

  • Curiosity about workload optimization, cost modeling, adaptive execution, and how compute efficiency affects AI and analytics at scale

  • A pragmatic builder’s mindset: rigorous, hands-on, and comfortable owning complex systems end to end

Bonus

  • Experience contributing to Apache Spark, Spark SQL, Trino, Presto, Flink, Velox, DuckDB, DataFusion, Iceberg, Delta Lake, Hudi, Parquet, ORC, or related systems

  • Experience with Catalyst, Adaptive Query Execution, cost-based optimization, query planning, vectorized execution, or distributed runtime systems

  • Experience optimizing joins, aggregations, shuffles, scans, spills, caching, partitioning, skew handling, or task scheduling

  • Experience building workload schedulers, execution control planes, resource managers, or multi-engine compute platforms

  • Experience reducing compute cost or improving workload efficiency in large-scale production data environments

  • Background in query engines, distributed runtimes, storage-aware execution, indexing, caching, encoding, compression, or adaptive query optimization

  • Research or open-source contributions in distributed systems, databases, query processing, data processing, or cloud infrastructure

Why Join Granica

  • Build foundational infrastructure for enterprise data and AI

  • Work on deep systems problems across distributed compute, query execution, workload optimization, scheduling, resource management, and compute efficiency

  • Partner directly with Product, Engineering, and company leadership

  • Help shape Crunch, Granica’s production data optimization platform for enterprise-scale lakehouse environments

  • Work with a small, high-caliber team solving high-value infrastructure problems at massive scale

  • Have direct influence on architecture, product direction, customer outcomes, and company growth

Compensation & Benefits

  • Competitive salary, meaningful equity, and performance bonus for top performers

  • 401(k) with company match, comprehensive health coverage, and unlimited PTO

  • Daily catered meals in our Mountain View office

  • Support for research, publication, and conference participation

At Granica, you'll help build the next generation of enterprise AI -from exabyte-scale data infrastructure, Large Tabular Models (LTMs), and stateful AI agents. Together, we're creating the infrastructure that enables enterprises to own their data, own the intelligence built on it, and scale both efficiently.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
In office • Internship • Master's Degree • Austin
C++
Python
C++
PyTorch C++
AI/ML
AI Agents
CUDA
CUDA Toolkit
LLM
NCCL
PyTorch
TensorRT
TensorRT-LLM
Triton
vLLM
Apply
$108k – $195k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • United States
Python
AI/ML
AI Agents
Amazon SageMaker
AWS Bedrock
AWS Bedrock AgentCore
LLM
DevOps
AWS
CI/CD
CloudFormation
Docker
IAM
Kubernetes
Platform Engineering
Terraform
Cybersecurity
FedRAMP
Apply
$86k – $207k per year (Estimated) • In office • Full-Time • Canberra
SQL
Databases
Amazon Redshift
Apache Kafka
Databricks
Delta Lake
Google BigQuery
Microsoft Fabric
Snowflake
AI/ML
dbt
DevOps
AWS
Azure
CI/CD
Analytics
ETL/ELT
Apply
$197k – $374k per year (Estimated) • In office • Full-Time • 8+ years exp • PhD • New York
AI/ML
AI Agents
Claude
LangChain
OpenAI
Vertex AI
Management
n8n
Zapier
Apply
Senior AI Architect 4 hours ago
$113k – $189k per year • In office • Full-Time • Bachelor's Degree • Barcelona
AI/ML
Anomaly Detection
Computer Vision
CrewAI
LangGraph
RAG
LangChain
LLMOps
AI Agents
Time Series Forecasting
DevOps
AWS
Azure
GCP
Management
n8n
Apply
$160k – $240k per year • In office • Full-Time
C++
Java
Rust
Scala
Databases
Apache Hudi
Apache Iceberg
Databricks
Delta Lake
DuckDB
Presto
Snowflake
Trino
AI/ML
AI Agents
Flink
Polars
Spark
DevOps
Amazon S3
Apply
$160k – $240k per year • In office • Full-Time • 5+ years exp
Databases
Databricks
Snowflake
AI/ML
AI Agents
LLM
OpenAI
Post-training
DevOps
AWS
Apply
$160k – $220k per year • In office • Full-Time • 5+ years exp
SQL
Databases
Apache Iceberg
Delta Lake
AI/ML
AI Agents
Spark
Apply
$145k – $318k per year (Estimated) • In office • Full-Time • PhD
Python
AI/ML
AI Agents
Diffusion Models
JAX
Multimodal AI
PyTorch
Apply
$160k – $250k per year • In office • Full-Time • PhD
Python
AI/ML
AI Agents
JAX
PyTorch
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.