Modal is an AI infrastructure company headquartered in New York City and founded in 2021. It provides a serverless cloud platform featuring sub-second cold starts and instant autoscaling that enables developers to run GPU-accelerated workloads, including model inference and fine-tuning, using a Python-native SDK. The company operates a globally distributed compute network designed for AI applications and serves a diverse range of industries such as generative AI and biotechnology.
HQ: New York, United States • MLOps • LLM & Generative AI • Machine Learning • Information Technology • Hardware • Cloud Computing • AI Infrastructure • Artificial Intelligence • Est. 2021
$200k – $300k per year • Lead • In office (New York, United States) • Full-Time • Staff
InfiniBand
DevOps
SLI/SLO/SLA
Analytics
Seaborn
Matplotlib
Perplexity AI is an American company that builds an answer engine combining live web search with large language models to return sourced, conversational responses instead of a list of links. Its products span a consumer assistant on web and mobile, the Comet browser, enterprise search over internal documents and the Sonar developer API that exposes the same grounded retrieval stack. Founded in 2022 in San Francisco by former researchers and engineers from OpenAI, Meta and Databricks, the company is backed by NVIDIA, IVP, New Enterprise Associates and SoftBank.
HQ: San Francisco, United States • AI Agents • Internet Services • Artificial Intelligence • Search Engines • LLM & Generative AI • 1001-5000 employees • Est. 2022
$220k – $485k per year • Staff+ • 3+ years exp • In office (San Francisco, New York, Palo Alto, United States) • Full-Time • Staff
Python
Rust
AI/ML
CUDA Toolkit
Multimodal AI
Perplexity
CUDA
Triton
JAX
LLM
PyTorch
Quantization
TensorFlow
InfiniBand
NCCL
NVLink
Frontend
Sass
DevOps
API Gateway
Kubernetes
≈ $161k – $354k per year (Estimated) • Staff+ • 3+ years exp • In office (London, United Kingdom) • Full-Time • Staff
Python
Rust
AI/ML
CUDA Toolkit
Multimodal AI
Perplexity
CUDA
Triton
JAX
LLM
PyTorch
Quantization
TensorFlow
InfiniBand
NCCL
NVLink
Frontend
Sass
DevOps
API Gateway
Kubernetes
vCluster is a Kubernetes infrastructure company headquartered in San Francisco, California, and founded in 2019. The company builds virtual clusters that run inside a single physical Kubernetes cluster, giving each team or tenant what looks like its own isolated control plane without the cost of separate infrastructure. It maintains the open source vCluster project alongside a commercial platform and previously operated as Loft Labs.
HQ: San Francisco, United States • Cloud Management • Cloud Computing • Software • Virtualization • DevOps • Information Technology • Est. 2019
≈ $116k – $224k per year (Estimated) • Senior • 5+ years exp • Remote (United States, Germany, Australia) • Full-Time • Senior
Python
AI/ML
CUDA
CUDA Toolkit
InfiniBand
LLM
OpenAI
DevOps
Kubernetes
GitHub
GitLab
Marketing
Salesforce
E2E Networks is a listed Indian cloud provider that pivoted to accelerated computing for artificial intelligence workloads. Founded in 2009 in Delhi, it operates large GPU clusters in Indian data centres for startups, research institutions and enterprises. Its TIR platform provides managed notebooks, fine-tuning and inference endpoints.
HQ: New Delhi, India • Cloud Computing • Artificial Intelligence • AI Infrastructure • Est. 2009
≈ $31k – $75k per year (Estimated) • Staff+ • In office (Delhi, India) • Architect
CUDA
CUDA Toolkit
Kubeflow
ONNX
PyTorch
Ray
TensorFlow
TensorRT
cuDNN
CVAT
InfiniBand
DevOps
CentOS Stream
Debian
Docker
Kubernetes
KVM
SLURM
Ubuntu
VMWare
Argonne National Laboratory is a premier multidisciplinary science and engineering research center located in Lemont, Illinois (near Chicago). Originating in 1946 from the University of Chicago’s work on the Manhattan Project (Metallurgical Laboratory), Argonne operates under the oversight of the U.S. Department of Energy (DOE) Office of Science and is managed by UChicago Argonne, LLC.
HQ: Lemont, United States • Energy & Utilities • Nuclear Energy • Quantum Science & Computing • Science & Engineering • Est. 1946
$86k – $135k per year • Senior • 6+ years exp • In office (Lemont, United States) • Bachelor's Degree • Full-Time • Senior
Python
C
C
MPI
AI/ML
Cerebras
Groq
InfiniBand
DevOps
Ansible
Configuration Management
Git
HPC
Kubernetes
CURSOR Software AG is a German software company headquartered in Gießen, Germany, specializing in enterprise Customer Relationship Management (CRM) and Business Process Management (BPM) solutions. Founded over 30 years ago, the firm provides tailored, GDPR-compliant CRM platforms - including CURSOR-CRM, EVI, and TINA - specifically engineered for energy providers, financial institutions, and mid-to-large-scale industrial enterprises. Its modular software suite centralizes sales, marketing, and customer service operations while automating complex business workflows and energy sector compliance processes.
HQ: Giessen, Germany • Business Process Automation (BPA) • Software • 501-1000 employees
≈ $158k – $345k per year (Estimated) • In office (San Francisco, New York, United States) • Full-Time
Go
Python
Rust
TypeScript
AI/ML
AI Agents
Ray
InfiniBand
DevOps
Kubernetes
SLURM
Genmo is a San Francisco company founded in 2021 that develops open video generation models. Its Mochi model was released with open weights and became one of the stronger freely available text-to-video systems, used widely by researchers and creative developers. The company positions open models as the foundation for a broader generative video ecosystem.
HQ: San Francisco, United States • LLM & Generative AI • Artificial Intelligence • Generative Video • Est. 2021
≈ $159k – $309k per year (Estimated) • Senior • 5+ years exp • In office (San Francisco, United States) • Bachelor's Degree • Full-Time • Senior
C++
Python
AI/ML
CUDA
CUDA Toolkit
Triton
InfiniBand
Frontend
Sass
Etched. We co-design chips, racks, software, and manufacturing methods so frontier models can run with best-in-class throughput, latency, cost, and power efficiency for both prefill and decode workloads.
San Jose • Taipei • Austin • Artificial Intelligence • AI Infrastructure • Semiconductors
$175k – $275k per year • In office (San Jose, United States) • Relocation • Full-Time
C++
Rust
C++
PyTorch C++
AI/ML
JAX
PyTorch
InfiniBand
NVLink
TPU
Transformers
Mixture of Experts
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.
HQ: Redwood City, United States • Machine Learning • Cloud Computing • Information Technology • MLOps • LLM & Generative AI • Artificial Intelligence • 501-1000 employees • Est. 2022
$175k – $220k per year • Staff+ • 5+ years exp • In office (San Mateo, United States) • Bachelor's Degree • Full-Time • Staff
CUDA
CUDA Toolkit
Multimodal AI
PyTorch
Quantization
Triton
ROCm
Fireworks AI
LLM
InfiniBand
Structured Outputs
DevOps
HPC
Kubernetes