953,927open jobs
57,674companies
156,942added this week
Browse all
Salary
≈ $158k – $291k per year (Estimated)
Location
In office (Palo Alto)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 29, 2026. First seen by Alion on Aug 4, 2026.

Overview
Company
Impact
Profile match
Cloud Infrastructure for Scientific Discovery. Harell Cloud is built for researchers on the frontier of what AI can do.

About Harell Data

AI has transformed digital industries, but progress in the physical sciences - drug discovery, materials science, climate modeling - has stalled. The bottleneck isn't compute or algorithms. It's data. The most valuable scientific datasets are locked in silos, unstructured, and inaccessible.

We're fixing that. Harell Data is a managed platform where organizations can securely share proprietary datasets, train models on high-performance GPUs, and deploy them for inference and application development. Dataset owners share their data for model training without giving up control of it. Researchers and engineers get access to compute to train, deploy, and use their models. It's the infrastructure layer that turns scattered scientific data into domain-specific foundation models.

About the Role

You'll be an early engineer reporting directly to the CTO. You'll own the compute layer: the GPU clusters and the inference systems that run on them. You'll make the architectural decisions that define the platform. You'll also work directly with customers to understand what they actually need and turn that into infrastructure that works at scale.

What You Will Do

  • Build the GPU compute layer - Orchestration for GPU workloads on Kubernetes: resource allocation, scheduling, multi-tenancy, and cost management.
  • Build the inference layer - Model loading, autoscaling, batching, and serving. You own the latency and throughput customers feel.
  • Own the ML pipeline end to end - Data ingestion, preprocessing, training and fine-tuning jobs, and recovery when multi-node jobs fail.
  • Work directly with customers - Debug fine-tuning jobs that fail or run slow. Build the observability that tracks model performance and resource health in real time.
  • Own reliability - Incident response, on-call, and keeping the platform up as usage grows.
  • Shape technical direction - Lead build-vs-buy decisions on infrastructure and security. Set engineering standards. Help hire the team you want to work with.

Qualifications

  • 5+ years building and operating production infrastructure, with a focus on ML workloads: training, inference, or data pipelines
  • Hands-on experience with Kubernetes on AWS or GCP, ideally with GPU workloads.
  • Strong CS fundamentals and system design chops
  • Comfortable with ambiguity - you've worked somewhere where the playbook didn't exist yet

Location note: this role is based in Palo Alto, CA. No relocation assistance available for this role.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
953,927 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Palo Alto
≈ $129k – $258k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Mountain View
Python
Java
C++
Apply
≈ $150k – $277k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Kirkland
Java
C++
AI/ML
Vertex AI
TPU
Edge AI
DevOps
GCP
Apply
≈ $157k – $290k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Sunnyvale
Python
C++
AI/ML
Vertex AI
TPU
Edge AI
DevOps
GCP
Apply
≈ $124k – $230k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Atlanta
JavaScript
TypeScript
Frontend
Angular
Mobile
Reactive Programming
Apply
≈ $192k – $347k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • New York
AI/ML
LLM
Apply
≈ $54k – $106k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Spain
Python
JavaScript
TypeScript
C#
C#
ASP.NET Core
Entity Framework Core
Databases
PostgreSQL
RabbitMQ
MS SQL
Apache Kafka
Frontend
Vue.js
Pinia
Vuex
DevOps
Terraform
GCP
Helm
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Apply
≈ $19k – $38k per year (Estimated) • In office • 6+ years exp • Mumbai
Python
SQL
Databases
Databricks
Amazon Redshift
AI/ML
Spark
Airflow
DevOps
GCP
Azure
Git
AWS
Analytics
ETL/ELT
Dimensional Modeling
Apply
≈ $16k – $36k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Pune • Thiruvananthapuram • Bengaluru
Python
PowerShell
DevOps
Terraform
Ansible
Azure DevOps
GitHub Actions
Prometheus
GitLab CI
Azure
CI/CD
Windows Server
Jenkins
Kubernetes
Grafana
Azure AKS
Incident Management
SLI/SLO/SLA
Windows
DNS
Cybersecurity
Active Directory
Management
ServiceNow
ITSM
Apply
≈ $16k – $40k per year (Estimated) • Hybrid • Moscow
Python
Java
DevOps
CI/CD
Docker
Kubernetes
Apply
$78k – $147k per year • Hybrid • Full-Time • Bachelor's Degree • Washington
Python
SQL
Databases
Apache Iceberg
Amazon Aurora
Trino
AI/ML
AWS Bedrock
Machine Learning
DevOps
AWS
Amazon S3
Analytics
AWS Glue
Apply
≈ $131k – $261k per year (Estimated) • In office • Full-Time • Bellevue
AI/ML
PyTorch
Hugging Face
DevOps
GCP
AWS
HPC
Linux
Apply
$177k – $365k per year • Hybrid • 6+ years exp • Bachelor's Degree • Palo Alto
Python
Go
C++
DevOps
Traefik
Envoy
Nginx
Platform Engineering
DNS
Apply
$235k – $323k per year • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Palo Alto
TypeScript
Databases
PostgreSQL
DynamoDB
AI/ML
Claude Code
AI Agents
LLM
OpenAI Codex
A2A
DevOps
AWS
Platform Engineering
AWS Lambda
API Gateway
Apply
$44k per year • In office • Palo Alto
Apply
Apply
$44k – $56k per year • In office • Palo Alto
DevOps
Chef
Apply
See all jobs
This is one of many
953,927 more open roles from verified company boards, updated every day.