368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$160k – $240k per year
Location
In office
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Granica is an enterprise AI efficiency platform designed to reduce storage, compute, and context costs for large-scale AI applications. Founded in 2019 and headquartered in Mountain View, California, the company operates within its customers' own cloud perimeters (AWS and Google Cloud) to optimize enterprise data and model workflows securely without exposing sensitive information.

Senior Software Engineer - Lakehouse Systems

Location: Mountain View, CA - On-site

About Granica

Granica builds AI infrastructure for enterprises operating massive data environments.

Our platform helps data and engineering teams reduce storage and compute costs, improve performance and reliability, and prepare large datasets for analytics and AI.

Granica’s products include:

  • Crunch - continuous optimization for enterprise lakehouse data

  • Myelin - stateful infrastructure for long-running AI agents

  • Large Tabular Models - foundation models designed for enterprise tables

Together, we are building the infrastructure that enables enterprises to own their data, own the intelligence built on it, and scale both efficiently.

Granica has demonstrated approximately $200K in annualized value per petabyte and verified customer value within weeks.

About the Role

Granica is hiring a Senior Software Engineer to build foundational lakehouse systems for AI.

You will work on the core infrastructure behind Crunch, Granica’s continuous optimization product for enterprise lakehouse data. This includes systems for metadata management, transaction semantics, table maintenance, object-store-backed storage layouts, file-level optimization, and lakehouse cost/performance across petabyte- and exabyte-scale environments.

You will own core systems that directly affect customer infrastructure cost, query performance, table reliability, and the operational health of large lakehouse environments.

This is a hands-on engineering role for someone who has deep systems experience and wants to build at the intersection of data lakes, table formats, metadata systems, storage layout, query performance, and AI infrastructure.

You will work on lakehouse systems involving Apache Iceberg, Delta Lake, Apache Hudi, Parquet, ORC, cloud object stores, and query engines such as Spark, Trino, Presto, Flink, Databricks, and Snowflake-adjacent environments.

What You’ll Do

  • Build metadata and transaction systems for large-scale tabular datasets

  • Design systems that support time travel, schema evolution, partition evolution, snapshot isolation, and atomic consistency

  • Develop table-maintenance infrastructure for lakehouse formats such as Apache Iceberg, Delta Lake, and Apache Hudi

  • Build systems for manifests, snapshots, transaction logs, metadata pruning, snapshot expiration, table garbage collection, and catalog consistency

  • Optimize file layout, clustering, compaction, file sizing, data skipping, indexing, and read-path performance

  • Improve performance and cost efficiency across object-store-backed lakehouse environments such as S3, GCS, and ADLS

  • Work with columnar formats such as Parquet and ORC, including encoding, compression, layout, pruning, and read-path optimization

  • Build systems that make lakehouse tables faster, cheaper, and more reliable across engines and platforms such as Spark, Flink, Trino, Presto, Databricks, and Snowflake-adjacent environments

  • Debug performance bottlenecks across storage, metadata, table maintenance, query execution, network, and compute layers

  • Develop workload-aware table optimization systems that learn from access patterns and reorganize data automatically

  • Implement algorithms in compression, representation, layout optimization, and data efficiency

  • Contribute to open-source or publish research when appropriate

What We’re Looking For

  • Strong engineering depth in distributed systems, storage systems, databases, or data infrastructure

  • Production experience with modern data lake or lakehouse technologies such as Iceberg, Delta Lake, Hudi, Spark, Trino, Presto, Flink, Hive Metastore, Unity Catalog, or similar systems

  • Hands-on experience with columnar formats such as Parquet or ORC

  • Understanding of metadata-driven architectures, table formats, transaction semantics, query planning, and physical data layout

  • Experience with table maintenance, compaction, clustering, file sizing, metadata pruning, snapshot expiration, or garbage collection

  • Familiarity with cloud object storage systems such as S3, GCS, or ADLS and the performance tradeoffs of building lakehouse systems on top of them

  • Strong programming skills in Java, Scala, Go, Rust, C++, or similar systems-oriented languages

  • Curiosity about compression, entropy, information theory, and how data representation affects AI efficiency

  • A pragmatic builder’s mindset: rigorous, hands-on, and comfortable owning complex systems end to end

Bonus

  • Experience contributing to Apache Iceberg, Delta Lake, Apache Hudi, Spark, Flink, Trino, Presto, Velox, DuckDB, Polars, Parquet, ORC, or related systems

  • Experience with manifests, snapshots, metadata catalogs, schema evolution, partition evolution, delete handling, transaction logs, or table garbage collection

  • Experience solving the small-file problem, optimizing object-store access patterns, or improving table health at scale

  • Background in storage engines, query engines, indexing, caching, encoding, compression, or adaptive query optimization

  • Research or open-source contributions in distributed systems, databases, storage, compression, indexing, or data processing

  • Interest in how physical data representation affects model training, inference, retrieval, and reasoning efficiency

Why Join Granica

  • Build foundational infrastructure for enterprise data and AI

  • Work on deep systems problems across lakehouse metadata, transaction semantics, table maintenance, storage layout, object-store behavior, query performance, and AI efficiency

  • Partner directly with Product, Engineering, and company leadership

  • Help shape Crunch, Granica’s production data optimization platform for enterprise-scale lakehouse environments

  • Work with a small, high-caliber team solving high-value infrastructure problems at massive scale

  • Have direct influence on architecture, product direction, customer outcomes, and company growth

Compensation & Benefits

  • Competitive salary, meaningful equity, and performance bonus for top performers

  • 401(k) with company match, comprehensive health coverage, and unlimited PTO

  • Daily catered meals in our Mountain View office

  • Support for research, publication, and conference participation

At Granica, you'll help build the next generation of enterprise AI -from exabyte-scale data infrastructure, Large Tabular Models (LTMs), and stateful AI agents. Together, we're creating the infrastructure that enables enterprises to own their data, own the intelligence built on it, and scale both efficiently.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$54k – $128k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Mexico City
C#
JavaScript
Python
TypeScript
C#
ASP.NET Core
Python
FastAPI
Databases
PostgreSQL
AI/ML
AI Agents
Anthropic
Embeddings
Function Calling
LLM
LLM Guardrails
Model Context Protocol
OpenAI
Semantic Search
Semantic Search
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
GitHub
GitHub Actions
Vector
Vercel
Apply
Robot SWE 23 min ago
$100k – $200k per year • Equity 0.2–1.2% • In office • Full-Time • Bachelor's Degree • Los Angeles
Bash
C++
Python
AI/ML
Computer Vision
Human-in-the-Loop
DevOps
Docker
Git
Ubuntu
Robotics
Autonomous Navigation
Motion Planning
Obstacle Avoidance
ROS2
Sensor Fusion
SLAM
Teleoperation
Apply
$129k – $231k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Atlanta
SQL
AI/ML
AI Agents
Claude
Claude Code
DevOps
Azure
Design
Figma
Apply
$117k – $177k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Chicago • Dallas
Apex
C#
Java
Python
Apex
Salesforce Data Cloud
AI/ML
Agentforce
AI Agents
Chain-of-Thought
Embeddings
Few-Shot Learning
Fine-tuning
Hallucination
LLM Guardrails
NLP
Prompt Engineering
RAG
RLHF
Semantic Search
Semantic Search
Frontend
GraphQL
DevOps
CI/CD
Git
Analytics
A/B Testing
Marketing
Salesforce
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
$160k – $240k per year • In office • Full-Time
C++
Java
Rust
Scala
SQL
Databases
Apache Hudi
Apache Iceberg
Databricks
Delta Lake
DuckDB
Presto
Snowflake
Trino
AI/ML
AI Agents
Flink
Spark
DevOps
Amazon S3
Apply
$160k – $240k per year • In office • Full-Time • 5+ years exp
Databases
Databricks
Snowflake
AI/ML
AI Agents
LLM
OpenAI
Post-training
DevOps
AWS
Apply
$160k – $220k per year • In office • Full-Time • 5+ years exp
SQL
Databases
Apache Iceberg
Delta Lake
AI/ML
AI Agents
Spark
Apply
$145k – $318k per year (Estimated) • In office • Full-Time • PhD
Python
AI/ML
AI Agents
Diffusion Models
JAX
Multimodal AI
PyTorch
Apply
$160k – $250k per year • In office • Full-Time • PhD
Python
AI/ML
AI Agents
JAX
PyTorch
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.