Qualcomm is a global technology leader in wireless telecommunications, connectivity, and semiconductor design. The company pioneered foundational mobile standards including 3G, 4G, and 5G while developing the Snapdragon platform that powers millions of smartphones, automotive systems, and connected devices. Operating primarily through chip sales and technology licensing, Qualcomm actively drives innovation in mobile computing, edge AI, and next-generation wireless networks.
HQ: San Diego, United States • AI Infrastructure • Information Technology • Data Centers • Edge AI • Internet of Things
$285k – $427k per year • Equity • Executive • 12+ years exp • In office (Santa Clara, United States) • Bachelor's Degree • Senior
Powering. Cowboy Space Corp. is building a power grid in outer space for artificial intelligence. It is an integrated system of satellites and rockets to haul high-performance compute and optical data transmission down from Low Earth Orbit. Earth's energy grid can't run at the pace of AI. We can.
Seattle • Washington • Launch Services • Space & Aerospace • Space Services
$150k – $225k per year • Lead • 5+ years exp • In office (Seattle, United States) • Relocation • PhD • Full-Time • Staff
Sarvam AI is a leading Indian artificial intelligence company focused on building full-stack sovereign generative AI infrastructure, foundational large language models (LLMs), and speech technologies tailored for India’s diverse languages and enterprise requirements.
HQ: Bengaluru, India • Natural Language Processing • LLM & Generative AI • Artificial Intelligence
≈ $11k – $44k per year (Estimated) • Junior • 2+ years exp • In office (Bengaluru, Chennai, India) • Full-Time • Junior
Python
AI/ML
InfiniBand
NCCL
NVLink
CoreWeave
DevOps
Kubernetes
SLI/SLO/SLA
SLURM
SRE
HPC
Thinking Machines Lab is an artificial intelligence research and product company based in San Francisco and founded in 2025. The company develops multimodal AI systems and open-weights models, such as Inkling, alongside developer tools like Tinker for model fine-tuning. It operates as a public benefit corporation focused on human-AI collaboration and open science, supported by significant venture capital investment.
HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
Python
Rust
AI/ML
NCCL
NVLink
CUDA Toolkit
DevOps
Kubernetes
SLURM
AbcdefghijklmnopqrstuvwxyzABCDEFGHIJKLMNOPQRSTUVWXYZ ABOUT PHOTO APP What you write here is totally up to you, there is no right or wrong way to complete your 'About Us' page. We advise making your language friendly and approachable, while being informative and professional.
HQ: London, United Kingdom
≈ $60k – $143k per year (Estimated) • Lead • 10+ years exp • In office (Paris, France) • Bachelor's Degree • Full-Time • Staff • French
NVLink
DevOps
SLI/SLO/SLA
HPC
≈ $97k – $175k per year (Estimated) • Senior • Remote/Hybrid (London, United Kingdom) • Full-Time • Senior
Go
AI/ML
NVLink
DevOps
CI/CD
gRPC
Kubernetes
The University of Texas at Austin (UT Austin) is a public research university and the flagship institution of the University of Texas System, known for its strong academic reputation and vibrant campus life. It serves as a major hub for innovation and research, offering a vast array of programs across numerous colleges and schools, including highly ranked engineering, business, and computer science departments.
HQ: Austin, United States • STEM Education • Education • Higher Education • 1001-5000 employees
≈ $91k – $199k per year (Estimated) • Lead • 12+ years exp • Remote/Hybrid (Austin, United States) • Master's Degree • Full-Time • Principal
NVLink
DevOps
HPC
Chips/EDA
Ansys HFSS
Ansys SIwave
Cadence Virtuoso
Synopsys HSPICE
Perplexity AI is an American company that builds an answer engine combining live web search with large language models to return sourced, conversational responses instead of a list of links. Its products span a consumer assistant on web and mobile, the Comet browser, enterprise search over internal documents and the Sonar developer API that exposes the same grounded retrieval stack. Founded in 2022 in San Francisco by former researchers and engineers from OpenAI, Meta and Databricks, the company is backed by NVIDIA, IVP, New Enterprise Associates and SoftBank.
HQ: San Francisco, United States • AI Agents • Internet Services • Artificial Intelligence • Search Engines • LLM & Generative AI • 1001-5000 employees • Est. 2022
$220k – $485k per year • Staff+ • 3+ years exp • In office (San Francisco, New York, Palo Alto, United States) • Full-Time • Staff
Python
Rust
AI/ML
CUDA Toolkit
Multimodal AI
Perplexity
CUDA
Triton
JAX
LLM
PyTorch
Quantization
TensorFlow
InfiniBand
NCCL
NVLink
Frontend
Sass
DevOps
API Gateway
Kubernetes
≈ $161k – $354k per year (Estimated) • Staff+ • 3+ years exp • In office (London, United Kingdom) • Full-Time • Staff
Python
Rust
AI/ML
CUDA Toolkit
Multimodal AI
Perplexity
CUDA
Triton
JAX
LLM
PyTorch
Quantization
TensorFlow
InfiniBand
NCCL
NVLink
Frontend
Sass
DevOps
API Gateway
Kubernetes
Extend your team and scale your solutions with Integrant. We provide custom software development, AI solutions, and domain expertise for regulated industries.
HQ: Cairo, Egypt • Artificial Intelligence • Information Technology • Software • Est. 1992
Lead • 8+ years exp • In office (Cairo, Egypt) • Bachelor's Degree • Full-Time • Staff • English (Optional)
Python
AI/ML
CUDA Toolkit
LLM
MLFlow
ONNX
PyTorch
Quantization
TensorRT
TensorRT-LLM
Triton Inference Server
CUDA
Triton
cuDNN
InfiniBand
NCCL
NVLink
Weights & Biases
NVIDIA NIM
Megatron-LM
NVIDIA NeMo
DevOps
Kubernetes
SLURM
HPC
CI/CD
OpenAI is an American artificial intelligence research and deployment company founded in 2015 with the mission of ensuring that artificial general intelligence benefits all of humanity. It develops the GPT family of large language models and turns them into consumer and developer products, including the ChatGPT assistant, the Sora video model, the Codex coding agent and a commercial API used by millions of developers. Structured as a public benefit corporation controlled by a non-profit foundation, the company is backed by Microsoft and SoftBank and operates from San Francisco.
HQ: San Francisco, United States • Multimodal AI • AI Agents • Artificial Intelligence • Machine Learning • LLM & Generative AI • 1001-5000 employees • Est. 2015
$293k – $385k per year • Senior • 5+ years exp • Remote/Hybrid (San Francisco, Seattle, United States) • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
CUDA Toolkit
LLM
PyTorch
Edge AI
NCCL
NVLink
OpenAI
DevOps
Kubernetes
HPC
$295k – $555k per year • Senior • 5+ years exp • In office (San Francisco, United States) • Full-Time • Senior
CUDA Toolkit
PyTorch
InfiniBand
NCCL
NVLink
OpenAI
DevOps
Azure
HPC
Automate every customer interaction with AI Phone Agents built specifically for enterprise.
San Francisco • LLM & Generative AI • Speech & Audio AI • Artificial Intelligence
$120k – $200k per year • Senior • 5+ years exp • In office (San Francisco, United States) • Full-Time • Senior
TypeScript
AI/ML
ElevenLabs
NVLink
Mobile
Twilio
DevOps
AWS
Cloudflare
Datadog
Docker
GCP
HAProxy
Kubernetes
Terraform
WebRTC
Parspec is a technology company that leverages AI to help sales agents and distributors by simplifying the process of discovering and sourcing the best available construction products and materials.
Bengaluru • San Mateo • Commerce • Marketplaces • Artificial Intelligence • Est. 2020
≈ $30k – $75k per year (Estimated) • Senior • 5+ years exp • In office (Bengaluru, India) • Bachelor's Degree • Full-Time • Senior
Python
Python
Asyncio
FastAPI
Databases
Apache Kafka
pgvector
Pinecone
Weaviate
PostgreSQL
Qdrant
AI/ML
AWS Bedrock
Kubeflow
LiteLLM
LLM
MLFlow
Portkey
Ray
vLLM
AWQ
CUDA
CUDA Toolkit
Embeddings
GPTQ
Hallucination
Langfuse
LoRA
Prompt Engineering
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
PEFT
Amazon SageMaker
AWS Trainium
LLM Guardrails
LLMOps
NCCL
NVLink
TGI
Multimodal AI
AI Agents
DevOps
AIOps
Amazon EC2
Amazon EKS
ArgoCD
AWS
AWS Lambda
CI/CD
CloudFormation
Docker
GitHub Actions
Grafana
Kubernetes
OpenTelemetry
Prometheus
Terraform
Vector
Karpenter
Platform Engineering
Amazon EventBridge
Amazon S3
API Gateway
IAM
AWS Step Functions
GitHub
Cybersecurity
Least Privilege
Baseten is an AI infrastructure platform designed to help developers and machine learning teams deploy, serve, and scale open-source and custom AI models. The platform provides performant, low-latency inference infrastructure alongside developer tools like Truss, an open-source model packaging framework. Headquartered in San Francisco, California, Baseten enables companies to run state-of-the-art models in production seamlessly without managing underlying cloud infrastructure.
HQ: San Francisco, United States • Hardware • Machine Learning • AI Infrastructure • Artificial Intelligence
$165k – $330k per year • Remote/Hybrid (San Francisco, New York, United States, Toronto, Montreal, Canada) • Full-Time
C++
Python
Rust
AI/ML
Cursor
LLM
Multimodal AI
Spark
TensorRT
TensorRT-LLM
InfiniBand
NCCL
NVLink
Mixture of Experts
SGLang
vLLM
DevOps
Kubernetes
The people, companies, and technologies shaping the future. Click to read The Generalist, by Mario Gabriele, a Substack publication with hundreds of thousands of subscribers.
San Francisco • Boston
$200k – $350k per year • In office (San Francisco, United States) • Full-Time
Python
AI/ML
ChatGPT
CUDA
CUDA Toolkit
Gemini
Multimodal AI
NumPy
GPT-4
NVLink
OpenAI
Hong Kong • Singapore
≈ $139k – $249k per year (Estimated) • Senior • 5+ years exp • Hong Kong • Singapore • Remote (United States, Canada, Hong Kong, Singapore) • Bachelor's Degree • Full-Time • Senior
Go
Python
AI/ML
Ray
AWS Trainium
CUDA Toolkit
Quantization
SGLang
vLLM
CUDA
Triton
Hugging Face
NCCL
NVLink
DevOps
AWS
Azure
Docker
GCP
Grafana
Helm
kubectl
Kubernetes
Prometheus
Rancher
SLURM
gRPC
GitHub
Etched. We co-design chips, racks, software, and manufacturing methods so frontier models can run with best-in-class throughput, latency, cost, and power efficiency for both prefill and decode workloads.
San Jose • Taipei • Austin • Artificial Intelligence • AI Infrastructure • Semiconductors
$175k – $275k per year • In office (San Jose, United States) • Relocation • Full-Time
C++
Rust
C++
PyTorch C++
AI/ML
JAX
PyTorch
InfiniBand
NVLink
TPU
Transformers
Mixture of Experts