vCluster is a Kubernetes infrastructure company headquartered in San Francisco, California, and founded in 2019. The company builds virtual clusters that run inside a single physical Kubernetes cluster, giving each team or tenant what looks like its own isolated control plane without the cost of separate infrastructure. It maintains the open source vCluster project alongside a commercial platform and previously operated as Loft Labs.
HQ: San Francisco, United States • Cloud Management • Cloud Computing • Software • Virtualization • DevOps • Information Technology • Est. 2019
≈ $116k – $224k per year (Estimated) • Senior • 5+ years exp • Remote (United States, Germany, Australia) • Full-Time • Senior
Python
AI/ML
CUDA
CUDA Toolkit
InfiniBand
LLM
OpenAI
DevOps
Kubernetes
GitHub
GitLab
Marketing
Salesforce
Meshy generates three-dimensional models and textures from text prompts or reference images. Founded in 2021 by former Nvidia and Microsoft graphics researchers, it targets game studios and product designers who need assets quickly. The output is exported in standard formats ready for game engines.
HQ: San Jose, United States • Design & Creative • Generative 3D • Artificial Intelligence • Est. 2021
Remote/Hybrid (Shenzhen, China) • Full-Time
C++
Python
AI/ML
CUDA
CUDA Toolkit
Quantization
In office (Shanghai, China) • PhD • Full-Time
C++
Python
Databases
Databricks
AI/ML
CUDA
CUDA Toolkit
DevOps
CI/CD
Vector
GitHub
Remote/Hybrid (Shanghai, China) • Full-Time
C++
Python
AI/ML
CUDA
CUDA Toolkit
Quantization
DevOps
HPC
E2E Networks is a listed Indian cloud provider that pivoted to accelerated computing for artificial intelligence workloads. Founded in 2009 in Delhi, it operates large GPU clusters in Indian data centres for startups, research institutions and enterprises. Its TIR platform provides managed notebooks, fine-tuning and inference endpoints.
HQ: New Delhi, India • Cloud Computing • Artificial Intelligence • AI Infrastructure • Est. 2009
≈ $31k – $75k per year (Estimated) • Staff+ • In office (Delhi, India) • Architect
CUDA
CUDA Toolkit
Kubeflow
ONNX
PyTorch
Ray
TensorFlow
TensorRT
cuDNN
CVAT
InfiniBand
DevOps
CentOS Stream
Debian
Docker
Kubernetes
KVM
SLURM
Ubuntu
VMWare
COMPANY RESEARCH CLOUD Zyphra Cloud Login Two sides. Two sides.
HQ: San Francisco, United States • AI Infrastructure • Artificial Intelligence • LLM & Generative AI
≈ $156k – $342k per year (Estimated) • In office (San Francisco, United States) • Relocation • Full-Time
CUDA Toolkit
CUDA
Triton
AWS Trainium
TPU
Mixture of Experts
DevOps
AWS
HPC
Unlock the potential of your warehouse with our end-to-end warehouse automation. Experience efficiency with autonomous mobile robots and robotic automation systems.
Denver • Lakewood • Warehousing • Robotics • Logistics & Warehouse Robotics
$150k per year • Senior • 8+ years exp • In office (Denver, United States) • Full-Time • Senior
C++
Java
Python
C++
Protobuf
Java
Spring Boot
AI/ML
Computer Vision
CUDA Toolkit
CUDA
DevOps
Docker
gRPC
Robotics
Isaac Sim
Motion Planning
Path Planning
ROS
ROS2
Sensor Fusion
SLAM
Fal is a generative media AI infrastructure platform headquartered in San Francisco, California. Founded in 2021 by former Coinbase and Amazon engineers Burkay Gur and Gorkem Yurtseven, the enterprise is backed by prominent venture investors including Andreessen Horowitz (a16z), Bessemer Venture Partners, and Salesforce Ventures.
HQ: San Francisco, United States • Generative Video • LLM & Generative AI • Artificial Intelligence • Est. 2021
Middle • 3+ years exp • In office • Full-Time • Middle
Python
AI/ML
CUDA
CUDA Toolkit
InfiniBand
NVLink
DevOps
Ansible
Configuration Management
Terraform
KVM
QEMU
Cybersecurity
ISO 27001
SOC 2
Tcpdump
$180k – $250k per year • Middle • 3+ years exp • In office (San Francisco, United States) • Relocation • Full-Time • Middle
Python
AI/ML
CUDA
CUDA Toolkit
InfiniBand
NVLink
DevOps
Ansible
Configuration Management
Terraform
KVM
QEMU
Cybersecurity
ISO 27001
SOC 2
Tcpdump
Tether is a digital asset company founded in 2014 that issues USDT, the largest stablecoin in circulation and the most heavily traded asset in crypto markets. Each token is intended to hold a one-to-one peg with the United States dollar and is backed by a reserve portfolio dominated by short-dated Treasury bills, with regular attestation reports published on the reserve composition. Headquartered in San Salvador, El Salvador, the group has expanded beyond stablecoin issuance into gold-backed tokens, Bitcoin mining, energy, peer-to-peer communications and local artificial intelligence infrastructure.
HQ: San Salvador, El Salvador • Crypto Payments • Data Centers • Artificial Intelligence • Blockchain & Crypto • Cryptocurrencies • Stablecoins • 1001-5000 employees • Est. 2014
Remote (Colombia) • English: B2
C++
JavaScript
AI/ML
CUDA Toolkit
Diffusion Models
Llama
llama.cpp
LocalAI
OpenCL
Tokenization
CUDA
Edge AI
LLM
Web3
Bitcoin
Parspec is a technology company that leverages AI to help sales agents and distributors by simplifying the process of discovering and sourcing the best available construction products and materials.
Bengaluru • San Mateo • Commerce • Marketplaces • Artificial Intelligence • Est. 2020
≈ $30k – $75k per year (Estimated) • Senior • 5+ years exp • In office (Bengaluru, India) • Bachelor's Degree • Full-Time • Senior
Python
Python
Asyncio
FastAPI
Databases
Apache Kafka
pgvector
Pinecone
Weaviate
PostgreSQL
Qdrant
AI/ML
AWS Bedrock
Kubeflow
LiteLLM
LLM
MLFlow
Portkey
Ray
vLLM
AWQ
CUDA
CUDA Toolkit
Embeddings
GPTQ
Hallucination
Langfuse
LoRA
Prompt Engineering
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
PEFT
Amazon SageMaker
AWS Trainium
LLM Guardrails
LLMOps
NCCL
NVLink
TGI
Multimodal AI
AI Agents
DevOps
AIOps
Amazon EC2
Amazon EKS
ArgoCD
AWS
AWS Lambda
CI/CD
CloudFormation
Docker
GitHub Actions
Grafana
Kubernetes
OpenTelemetry
Prometheus
Terraform
Vector
Karpenter
Platform Engineering
Amazon EventBridge
Amazon S3
API Gateway
IAM
AWS Step Functions
GitHub
Cybersecurity
Least Privilege
D-Wave is an industry leader in the development and production of quantum computing systems, software, and services. As the first commercial supplier of quantum computers, D-Wave is at the forefront of the dual development of quantum annealing and gate-model systems. D-Wave helps companies use quantum tech to solve real-world problems right now.
HQ: Palo Alto, United States • Hardware • Artificial Intelligence • Science & Engineering • Quantum Computing Hardware • Quantum Science & Computing
$125k – $187k per year • Senior • 5+ years exp • Remote/Hybrid (New Haven, United States) • Master's Degree • Full-Time • Senior
C++
Python
Q#
SQL
C++
LLVM
Q#
Quantum Intermediate Representation
Databases
Oracle
PostgreSQL
AI/ML
CUDA Toolkit
CUDA
DevOps
CI/CD
Git
Quantum
Cirq
Qiskit
Built for the realities of the factory floor thingtrax was founded by engineers and technologists who've lived the challenges of manufacturing. From machine monitoring to full-line optimisation, our journey has always been about solving real problems with practical, scalable technology.
Lahore • Gurgaon • Liverpool • Industrial IoT (IIoT) • Manufacturing • Industrial Automation
≈ $28k – $115k per year (Estimated) • Middle • 3+ years exp • In office (Gurgaon, India) • Bachelor's Degree • Full-Time • Middle
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
CUDA Toolkit
OpenCL
CUDA
AI Agents
Image Segmentation
PyTorch
TensorFlow
DevOps
Git
Elastix AI delivers scalable, energy-efficient AI inference through machine learning, system software, and reconfigurable hardware.
Seattle • Hardware • AI Infrastructure • Artificial Intelligence
≈ $138k – $259k per year (Estimated) • Equity • Middle • 3+ years exp • In office (Seattle, United States) • Bachelor's Degree • Full-Time • Middle
C++
Python
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
LLM
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
CUDA
DevOps
Docker
Kubernetes
The people, companies, and technologies shaping the future. Click to read The Generalist, by Mario Gabriele, a Substack publication with hundreds of thousands of subscribers.
San Francisco • Boston
$200k – $350k per year • In office (San Francisco, United States) • Full-Time
Python
AI/ML
ChatGPT
CUDA
CUDA Toolkit
Gemini
Multimodal AI
NumPy
GPT-4
NVLink
OpenAI
HeyGen is a company founded in 2020 that generates presenter videos from text using synthetic avatars and cloned voices. Its translation feature re-voices existing footage in other languages while matching lip movement, which made it popular with marketing and training teams. The company reports fast revenue growth serving businesses that produce large volumes of localised video.
HQ: Los Angeles, United States • Semiconductors • Embedded Systems • Cybersecurity • Est. 2020
≈ $118k – $282k per year (Estimated) • Lead • 5+ years exp • In office (Los Angeles, United States) • Bachelor's Degree • Full-Time • Staff
C++
Python
C++
PyTorch C++
TensorFlow C++
Databases
LanceDB
AI/ML
Accelerate
CUDA
CUDA Toolkit
JAX
Multimodal AI
PyTorch
Ray
Spark
TensorFlow
Scale AI
Diffusion Models
NCCL
DevOps
Kubernetes
HPC
≈ $83k – $235k per year (Estimated) • Senior • 5+ years exp • In office (Los Angeles, United States) • Bachelor's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
TensorFlow C++
Databases
LanceDB
AI/ML
Accelerate
CUDA
CUDA Toolkit
JAX
Multimodal AI
PyTorch
Ray
Spark
TensorFlow
Scale AI
Diffusion Models
NCCL
DevOps
Kubernetes
HPC
Architect AI is an artificial intelligence platform designed to transform urban design, architecture, and interior concepts into high-resolution visual renderings. The tool allows architects, designers, and real estate professionals to generate detailed 3D designs and explore spatial ideas rapidly using text prompts, sketch transformations, and mood boards. By streamlining the initial conceptual and visualization stages of building design, it helps creative teams accelerate project planning and present polished design options to clients.
Palo Alto • Bengaluru
≈ $180k – $365k per year (Estimated) • Equity • Staff+ • In office (Palo Alto, United States) • Bachelor's Degree • Full-Time • Staff
CUDA
CUDA Toolkit
Fine-tuning
PyTorch
QLoRA
Reinforcement Learning
LoRA
PEFT
Anthropic
Edge AI
OpenAI
Post-training
Function Calling
AHEAD helps enterprises build modern, secure, and scalable digital platforms by combining cloud, data, AI, and automation. Their consulting and managed services drive real business impact through smarter IT.
Gurgaon • New York • Palo Alto • Chicago • Minneapolis • Artificial Intelligence • Cloud Computing • Information Technology
≈ $23k – $63k per year (Estimated) • Middle • 4+ years exp • Remote (India) • Middle
Python
AI/ML
CUDA Toolkit
Kubeflow
MLFlow
PyTorch
TensorFlow
TensorRT
CUDA
Triton
NVIDIA NeMo
DevOps
Ansible
AWS
Azure
CI/CD
Configuration Management
GCP
Kubernetes
Terraform
HPC
Models. Agents. GPUs. Whatever you need — we've got it. Ghost agent VMs, Maestro model routing, Engine wholesale GPUs, and Academy, on one platform.
San Francisco
$220k – $320k per year • Senior • 2+ years exp • In office (San Francisco, United States) • Full-Time • Senior
C++
Python
C#
C++
PyTorch C++
C#
.NET
AI/ML
CUDA Toolkit
Knowledge Distillation
LLM
LoRA
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
PEFT
CUDA
GPT-5
Embeddings
DevOps
Docker
Kubernetes
GitHub
Dexmate is a robotics company headquartered in Santa Clara, California, and founded in 2024 by doctoral researchers from MIT, UC San Diego, and Carnegie Mellon. The company builds Vega, a foldable wheeled humanoid with two arms, dexterous hands, and an extending torso that reaches from floor level to overhead. It sells the platform to research groups and commercial pilots working on dexterous manipulation, and has taken strategic investment from NEC.
HQ: Santa Clara, United States • Logistics & Warehouse Robotics • Robotic Software & Control Systems • Robotics AI • Humanoid Robots • Artificial Intelligence • Robotics • 201-500 employees • Est. 2024
$120k – $300k per year • Senior • 5+ years exp • In office (Fremont, United States) • Master's Degree • Full-Time • Senior
C++
AI/ML
Multimodal AI
CUDA Toolkit
CUDA
Robotics
Ceres Solver
GTSAM
ROS
Sensor Fusion
SLAM
Visual-Inertial Odometry
Cartographer
LIO-SAM
ORB-SLAM3
RTAB-Map
ProSync, Professionals In-Sync, is an Employee Owned and Veteran Owned Small Business defense contractor that provides technology solutions and professional services to the federal government. Our primary focus is people and ensuring they have a ...
HQ: San Diego, United States • Cybersecurity • Military • Information Technology
≈ $84k – $177k per year (Estimated) • Middle • 4+ years exp • In office • Bachelor's Degree • Contractor • Middle
C++
Python
AI/ML
CUDA Toolkit
CUDA
DevOps
CI/CD
Docker
Git
GitLab CI
Helm
Jenkins
Kubernetes
GitLab
Etched. We co-design chips, racks, software, and manufacturing methods so frontier models can run with best-in-class throughput, latency, cost, and power efficiency for both prefill and decode workloads.
San Jose • Taipei • Austin • Artificial Intelligence • AI Infrastructure • Semiconductors
$150k – $275k per year • In office (San Jose, United States) • Relocation • Full-Time
C++
Rust
C++
PyTorch C++
TensorFlow C++
AI/ML
CUDA Toolkit
OpenCL
PyTorch
TensorFlow
CUDA
DevOps
Git
CI/CD
Docker
Kubernetes
Cohere is a Canadian AI company founded in 2019 and headquartered in Toronto. It builds secure, enterprise-focused large language models and AI tools for businesses and regulated industries. The company is known for emphasizing privacy, private deployment, and "sovereign AI" rather than consumer chat products
HQ: Toronto, Canada • Energy & Utilities • Biotechnology • Natural Language Processing • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2019
≈ $98k – $200k per year (Estimated) • Senior • London • New York • Paris • San Francisco • Toronto • Remote (United States, France, United Kingdom, Canada) • Full-Time • Senior
Cohere SDK
CUDA Toolkit
JAX
LLM
Ray
CUDA
FSDP
NCCL
DeepSpeed
PyTorch
TensorRT
TensorRT-LLM
vLLM
xFormers
Megatron-LM
DevOps
Docker
Kubernetes
SLURM
HPC
≈ $173k – $323k per year (Estimated) • Staff+ • 5+ years exp • New York • San Francisco • Toronto • Montreal • Remote (United States, Canada) • Full-Time • Staff
C++
Python
Rust
AI/ML
Cohere SDK
CUDA Toolkit
LLM
SGLang
vLLM
CUDA
Mixture of Experts
≈ $124k – $271k per year (Estimated) • Staff+ • Remote/Hybrid (London, United Kingdom, New York, United States, Paris, France, Toronto, Montreal, Canada) • Full-Time • Staff
Python
AI/ML
CUDA Toolkit
JAX
PyTorch
Transformers
CUDA
Triton
Pre-training