378,983open jobs
9,898companies
48,966added this week
Browse all
Salary
$163k – $356k per year (Estimated)
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
KREA is a company founded in 2022 that builds real-time generative visual tools for designers and artists. Its canvas updates images as the user sketches or types, and the product bundles upscaling, style training and video generation across multiple underlying models. It became popular with creative professionals who wanted control and speed rather than one-shot prompting.

About Krea

At Krea, we are building next-generation AI creative tools.

We're dedicated to making AI intuitive and controllable for creatives - our mission is to build tools that empower human creativity, not replace it. We believe AI is a new medium that allows us to express ourselves through various formats - text, images, video, sound, and even 3D. We're building better, smarter, and more controllable tools to harness this medium. We recently took this a step forward with the launch of Krea 2, our first foundation model, built completely from scratch for aesthetic diversity and stylistic control.

We've raised over $83M and are backed by world-class investors such as a16z, Bain Capital, and Abstract. We work full-time and in-person at our waterfront office in San Francisco. We care about creativity: our team includes musicians, designers, visual artists, and engineers.

We're looking for an experienced Researcher with engineering skills who can work on large-scale image and video models training experiments, with experience training image models at scale.

Our culture

  • We work full-time and in-person at our North Beach office in San Francisco.

  • We believe that demonstrated interest in the creative space is key: our team includes musicians, designers, visual artists and more.

  • Fast iteration and execution speed. Bias towards action, agency, and independence.

What you'll do

  • Train diffusion models for image and video generation on large GPU clusters.

  • Fully optimize and profile large distributed training runs across model architectures, kernels, data loading, memory constraints, and communication.

  • Implement and improve various distributed training strategies including FSDP, CP, SP, TP, and EP.

  • Continuously improve model quality and reliability through data, model architecture, training pipeline, structuring experiments, and eval design.

  • Debug distributed training errors and implement fault tolerance solutions, identifying bad GPU, NVLink, Infiniband (IB) components as well as monitoring numerical errors and NCCL issues.

  • Ablate different architecture, attention, optimizer, data, and algorithmic choices to reliably improve efficiency and performance of our models.

What we're looking for

  • Proven track record in working with image or video models at scale (publications or open-source contributions a plus).

  • Strong proficiency in PyTorch and understanding of its inner workings.

  • Strong background in distributed training paradigms such as FSDP, CP, SP, USP, TP, and EP. Knowing how different parallelism strategies work together and their tradeoffs.

  • Experience in profiling and debugging large distributed training. Being comfortable with analyzing traces to identify bottlenecks and look for improvements.

  • Good knowledge of low precision training / inference in FP8, NVFP4, and MXFP8.

  • Solid understanding of diffusion model training pipeline across pretraining, midtraining, preference optimization, and reinforcement learning.

  • Keeping up with the developments in related fields such as LLM, VLM, representation learning, and robotics research.

  • Being comfortable working in a goal-oriented research environment.

  • Having good judgement around when one should explore different training strategies and when it's time to commit to a specific strategy to scale compute and data.

  • Comfortable working with underspecified goals. We expect every technical member to take an ambiguous research goal and break it down into concrete requirements, plans, experiment plan, and execution items.

  • Good research taste - bias towards simplicity and methods that scale well with compute, data, and minimal human supervision.

  • Ability to iterate rapidly, and propose creative research directions.

  • Be comfortable getting your hands dirty with data and designing custom data pipelines to improve data quality.

What we offer

  • Team: Work alongside a world-class team building the future of AI creative tooling

  • Impact: Significant scope and company-wide impact

  • Competitive compensation: generous salary & equity packages

  • Health & wellness: 100% health & 99% dental/vision insurance premiums covered for employees, health FSA accounts, & long-term disability coverage

  • Time off: Flexible PTO policy

  • Financial planning: 401k with a 4% company-sponsored match

  • Meals in the office: breakfast, lunch, dinner - you name it, we'll cover it

  • Transit: Ubers covered to & from the office

  • Sponsorship: We're open to sponsoring international visas where we can (e.g., STEM OPT, OPT, H-1B, O-1, E-3).

  • And more!

Please note the above benefits & perks are for full-time employees

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
378,983 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$36k – $72k per year (Estimated) • In office • 5+ years exp • Smolensk
AI/ML
LLM
Apply
$91k – $207k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Sydney
Python
SQL
TypeScript
AI/ML
AI Agents
Amazon SageMaker
Claude
Claude Code
Fine-tuning
Function Calling
LangChain
Langfuse
LangGraph
LLM
LLM Guardrails
OpenAI
OpenAI Agents SDK
OpenAI Codex
RAG
XGBoost
DevOps
AWS
CI/CD
GitHub
Apply
AI Solution Architect 7 hours ago
$32k – $76k per year (Estimated) • In office • 10+ years exp • Delhi
Python
Databases
Google BigQuery
Milvus
Pinecone
PostgreSQL
Qdrant
Weaviate
AI/ML
AI Agents
Amazon SageMaker
AWS Bedrock
Fine-tuning
Hugging Face
LangChain
LangGraph
LLM
OpenAI
PyTorch
RAG
TensorFlow
Transformers
Vertex AI
Frontend
GraphQL
DevOps
AWS
AWS Lambda
Azure
Azure AKS
GCP
IAM
WebSockets
Kubernetes
Apply
$71k – $209k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Kfar Saba
C++
MATLAB
Python
AI/ML
AI Agents
LLM
RAG
DevOps
CI/CD
Git
Kubernetes
Apply
$193k – $424k per year (Estimated) • Remote/Hybrid • Full-Time • London
JavaScript
Python
AI/ML
AI Agents
ChatGPT
LLM
OpenAI
OpenAI Codex
Apply
$162k – $354k per year (Estimated) • Equity • In office • Full-Time • San Francisco
AI/ML
Diffusion Models
DPO
Fine-tuning
FSDP
GRPO
Knowledge Distillation
LLM
Post-training
PPO
PyTorch
Reinforcement Learning
SFT
SGLang
vLLM
VLM
Robotics
Reinforcement Learning
Apply
Product Engineer 2 days ago
$128k – $241k per year (Estimated) • Equity • In office • Full-Time • 4+ years exp • San Francisco
Python
TypeScript
JavaScript
Databases
Amazon Aurora
ClickHouse
PostgreSQL
Redis
AI/ML
PyTorch
Frontend
Svelte
SvelteKit
WebGPU
Zod
DevOps
AWS
Docker
FluxCD
Kubernetes
Apply
Fullstack Engineer 14 days ago
$148k – $303k per year (Estimated) • Equity • In office • Full-Time • San Francisco
TypeScript
JavaScript
Databases
PostgreSQL
Redis
Frontend
Svelte
SvelteKit
DevOps
Amazon EKS
AWS
Kubernetes
OpenTelemetry
Prometheus
WebSockets
GitHub
Management
Notion
Slack
Apply
$132k – $263k per year (Estimated) • Equity • In office • Full-Time • 4+ years exp • San Francisco
Python
TypeScript
JavaScript
Databases
Amazon Aurora
ClickHouse
PostgreSQL
Redis
AI/ML
PyTorch
Frontend
Svelte
SvelteKit
WebGPU
Zod
DevOps
AWS
Docker
FluxCD
Kubernetes
Apply
$131k – $263k per year (Estimated) • Equity • In office • Full-Time • 4+ years exp • San Francisco
Python
TypeScript
JavaScript
Databases
Amazon Aurora
ClickHouse
PostgreSQL
Redis
Frontend
SvelteKit
Zod
Svelte
DevOps
AWS
Docker
FluxCD
Kubernetes
Apply
$133k – $338k per year • Remote/Hybrid • Full-Time • 8+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
Knowledge Graph
DevOps
AWS
Azure
GCP
Platform Engineering
SLI/SLO/SLA
Apply
$208k – $332k per year • Remote/Hybrid • Full-Time • 9+ years exp • Bachelor's Degree • San Francisco
Apply
$197k – $314k per year • Remote/Hybrid • Full-Time • 8+ years exp • PhD • San Francisco
Apex
Java
Apex
MuleSoft
Java
Spring Boot
AI/ML
AI Agents
Model Context Protocol
RAG
Agentforce
DevOps
Docker
Kubernetes
Marketing
Salesforce
Apply
$220k – $413k per year (Estimated) • Remote/Hybrid • 12+ years exp • San Francisco
Java
Python
Ruby
Databases
DynamoDB
ElasticSearch
RocksDB
AI/ML
AI Agents
LLM Guardrails
Apply
$149k – $260k per year • In office • Full-Time • 6+ years exp • PhD • San Francisco
Go
JavaScript
TypeScript
AI/ML
AI Agents
Claude
Claude Code
Copilot
Cursor
Prompt Engineering
Agentforce
OpenAI Codex
Frontend
React.js
DevOps
Kubernetes
GitHub
Marketing
Salesforce
Apply
See all jobs
This is one of many
378,983 more open roles from verified company boards, updated every day.