368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$163k – $316k per year (Estimated)
Location
In office (San Francisco)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

About Engram

Today’s AI is a brilliant stranger: it can solve the world’s hardest math problems, but it knows next to nothing about you and your work. It rereads your files to answer even basic questions, burns an enormous amount of tokens when sifting through large corpuses, and between sessions, it retains scraps at best.

We train models to study your world and anticipate your questions in advance, forming engrams: compact memories that capture your knowledge and history. Our approach opens a new axis of scaling. The more we study your context at training time, the better we become at inference time.

We're already working with leaders in AI like Microsoft, Notion, and Harvey, and just raised $98M from General Catalyst, Kleiner Perkins, Sequoia, Factory, Modern, Amplify, Neo and others. Our investors and advisors include Assaf Rappaport, Andrej Karpathy, and Pieter Abbeel.

AI has spent years learning everything about the world. Now it should learn something about yours.

About this role

You’ll be among the first ML Systems Engineer, joining a team of machine learning researchers and performance engineers in building our personalization and continual learning API, which powers models and agents that learn from user context.

This role is focused on designing, optimizing, and scaling training and inference workloads -bridging the gap between cutting-edge AI research and production. This includes:

  • Designing and executing new frameworks, techniques, and systems to improve performance, reliability, latency, and efficiency.

  • Partner closely with researchers, turning prototypes into systems that run at scale and feeding systems constraints back into research decisions.

  • Optimize serving paths for personalization and memory retrieval, where per-user state and low latency both matter.

  • Work on distributed training - data and model parallelism, communication scheduling, and scaling efficiency across multiple GPUs and nodes.

You'll be the bridge between researchers and platform engineering, while working in deep collaboration with our customers (AI-native application-layer companies like Notion and Harvey). The architecture will be shaped by the constraints and requirements of our R&D work. There are no walls between product, research, and engineering here; delivering on our mission requires a multidisciplinary approach.

This is a founding hire in the truest sense. You’ll set the bar for engineering at Engram: code review, testing, on-call, and a security posture that gives customers confidence in entrusting us with their most sensitive data. You’ll also help build the engineering team around you and influence our engineering culture as we scale.

Your background looks like

  • Bachelor’s degree or equivalent experience in computer science, engineering, or similar.

  • 5+ years of experience with training or inference systems, optimized workloads with measurable results.

  • Strong engineering foundation, with demonstrated excellence navigating complex technical environments and shipping high-quality code in a fast-paced environment.

  • Deep understanding of ML framework (eg. PyTorch, JAX), GPUs, distributed systems, and infrastructure.

  • Operate well in ambiguous environments - you will have real ownership and be responsible for steering the ship in a novel sector of the industry.

  • You have a bias toward action and a knack for turning research concepts into concrete, executable plans.

Bonus points if you have

  • Prior early-stage experience.

  • Experience in open-source ML or systems infrastructure projects.

Engram is based in San Francisco. This role is in-person in our SF office. We offer competitive cash compensation and startup equity.

Engram is an equal opportunity employer. We’re building a team that reflects a range of backgrounds and perspectives, and we welcome applicants regardless of race, color, religion, national origin, gender, gender identity, sexual orientation, age, disability, or veteran status.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$93k – $190k per year (Estimated) • Remote • Full-Time
DevOps
Platform Engineering
Apply
$60k – $123k per year (Estimated) • Remote • Full-Time
DevOps
Platform Engineering
Apply
$61k – $126k per year (Estimated) • Remote • Full-Time
DevOps
Platform Engineering
Apply
$106k – $216k per year (Estimated) • Remote • Full-Time
DevOps
Platform Engineering
Apply
Platform Engineer 2 months ago
$137k – $267k per year (Estimated) • In office • Full-Time • 5+ years exp • San Francisco
Management
Notion
Apply
Research Scientist 2 months ago
$164k – $359k per year (Estimated) • In office • Full-Time • San Francisco
AI/ML
Fine-tuning
Knowledge Distillation
LLM
LoRA
Reinforcement Learning
Synthetic Data
PEFT
AI Agents
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
$180k – $260k per year • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
ChatGPT
OpenAI
OpenAI Codex
Cybersecurity
FedRAMP
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.