Salary
≈ $34k – $92k per year (Estimated)
Location
In office (Shanghai)
Seniority
Staff
Employment
Full-Time
Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on May 28, 2025.
Overview
Company
Impact
Profile match
MiniMax is a Chinese artificial intelligence company founded in Shanghai in 2021 by former SenseTime researchers, and one of the country's most commercially aggressive model developers. It builds text, speech, image and video foundation models in house and pushes them straight into consumer products, including the Hailuo video generator and the Talkie and Xingye companion applications that have found large audiences inside and outside China. The company also sells API access to enterprises, released the open-weight M-series reasoning models, and listed in Hong Kong in 2025.
岗位职责 / Responsibilities
1. 算力中心规划与建设(核心职责) - Lead 大模型训练/推理场景下的算力中心端到端建设:包括高性能网络拓扑设计(IB/RoCE 组网方案)、集群组网与超节点架构设计、服务器选型规划,从源头规避系统性故障风险,交付高可用、高性能的 AI 基础设施
2. 平台与硬件持续演进 - 参与 GPU 算力池化与调度、基础设施管理平台(资产/容量/故障自愈/可观测性)建设,跟踪 GPU/网络/存储技术演进并推动硬件选型与架构优化,持续降低 TCO
任职要求 / Requirements
1. 5 年以上云计算/IDC/服务器硬件/网络基础设施相关经验,深入理解计算机体系结构
2. 在以下方向中至少一个有深入的设计或研发经验:GPU 服务器架构设计、高速网络(IB/RoCE/NVLink/NVSwitch)、交换机研发、超节点架构设计
3. 熟悉主流 AI 训练/推理基础设施生态(NVIDIA DGX/HGX/GB200 NVL、集合通信、NCCL 等),了解大模型训练对基础设施的核心需求
4. 具备跨团队项目推动经验和良好的沟通领导力,能协调硬件、网络、软件等多方团队
加分项
1. 有万卡级 AI 算力集群的规划、建设或运营经验,主导过算力中心从 0 到 1 的落地
2. 有交换机研发、超节点架构设计、或服务器整机架构设计经验
3. 有新一代 AI 芯片/加速卡的适配设计经验(如华为昇腾/海思、国产 GPU 等)
4. 有头部云厂商或 AI 公司基础设施团队背景
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
778,720 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
AI/ML
Similar stack
Same company
Shanghai
≈ $37k – $96k per year (Estimated) • Hybrid • 7+ years exp • Bachelor's Degree • Guangzhou
AI/ML
AI Agents
RAG
Semantic Search
Semantic Search
Knowledge Graph
DevOps
CI/CD
Docker
Kubernetes
Platform Engineering
Apply
Staff Artificial Intelligence Engineer
1 hour ago
≈ $185k – $378k per year (Estimated) • In office • Full-Time • Bachelor's Degree • San Francisco
Python
Go
Java
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
GDPR
HIPAA
Apply
AI Engineer
3 months ago
≈ $21k – $63k per year (Estimated) • In office • Full-Time • 1+ year exp • Hyderabad
Python
Databases
Pinecone
FAISS
AI/ML
LLM
RAG
OpenAI
Hugging Face
LLM Guardrails
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Analytics
ETL/ELT
Apply
ML Engineer
3 months ago
In office • Full-Time
AI/ML
LoRA
Fine-tuning
Embeddings
AI Agents
PEFT
QLoRA
Transformers
LLM
RAG
DevOps
CI/CD
Docker
Kubernetes
Analytics
ETL/ELT
Apply
AI Engineer- FullStack
11 days ago
In office • 4+ years exp
Python
JavaScript
TypeScript
Node JS
AI/ML
LangGraph
LangChain
LlamaIndex
Arize Phoenix
Langfuse
LangSmith
RAG
Hybrid Search
Apply
$200k per year • Hybrid • 3+ years exp • New York
Python
C++
C++
PyTorch C++
AI/ML
DeepSpeed
CUDA Toolkit
Reinforcement Learning
JAX
PyTorch
Ray
Time Series Forecasting
CUDA
Triton
Megatron-LM
FSDP
NCCL
InfiniBand
NVLink
cuDNN
XLA
CUTLASS
Machine Learning
DevOps
SLURM
Kubernetes
HPC
Apply
$157k – $313k per year • In office • Full-Time • Singapore
Python
Bash
AI/ML
vLLM
RunPod
CoreWeave
NCCL
InfiniBand
NVLink
DevOps
Terraform
Ansible
Helm
SLURM
AWS
Kubernetes
Platform Engineering
AWS Lambda
HPC
Linux
Apply
Руководитель направления Pretrain Efficiency
9 hours ago
≈ $20k – $41k per year (Estimated) • In office • Moscow
Python
AI/ML
vLLM
CUDA Toolkit
RLHF
SGLang
TRL
GigaChat
Transformers
PyTorch
LLM
Ray
Mixture of Experts
CUDA
Triton
DPO
PPO
GRPO
Post-training
FSDP
NCCL
KV Cache
DevOps
SLURM
Kubernetes
Apply
$314k – $465k per year • Equity • Hybrid • Full-Time • 10+ years exp • San Francisco
Python
AI/ML
Claude Code
Anomaly Detection
NCCL
InfiniBand
Machine Learning
DevOps
GCP
Cilium
Prometheus
SLURM
Azure
GitOps
AWS
Kubernetes
Grafana
Platform Engineering
Chaos Engineering
Self-Healing
Amazon EKS
Google GKE
Azure AKS
AWS Lambda
AIOps
HPC
Linux
Apply
$230k – $346k per year • Equity • Hybrid • Full-Time • 6+ years exp • San Francisco
Python
AI/ML
NCCL
InfiniBand
Machine Learning
DevOps
GCP
Cilium
Prometheus
SLURM
Azure
AWS
Kubernetes
Grafana
Amazon EKS
Google GKE
Azure AKS
AWS Lambda
AIOps
HPC
Linux
Apply
Apply
Apply
行政专员 - AI 大模型方向
8 hours ago
≈ $23k – $68k per year (Estimated) • In office • Full-Time • 2+ years exp • Shanghai
Apply
大模型工程师-RL框架&RL推理
8 hours ago
In office • Full-Time • Beijing
AI/ML
vLLM
RLHF
AI Agents
SGLang
TensorRT
TensorRT-LLM
PyTorch
MiniMax
Megatron-LM
KV Cache
Agentic Workflows
Apply
大模型训练框架工程师
8 hours ago
In office • Full-Time • Beijing
AI/ML
DeepSpeed
CUDA Toolkit
PyTorch
MiniMax
CUDA
Triton
Megatron-LM
NCCL
Apply
Apply
行政专员 - AI 大模型方向
8 hours ago
≈ $23k – $68k per year (Estimated) • In office • Full-Time • 2+ years exp • Shanghai
Apply
AI Data 研发工程师(大模型数据方向)
8 hours ago
≈ $33k – $88k per year (Estimated) • In office • Full-Time • Shanghai
Python
Java
AI/ML
Spark
RLHF
Flink
Ray
Mixture of Experts
MiniMax
SFT
RLAIF
DevOps
Kubernetes
Linux
Apply
存储系统DevOps 工程师(大模型 / AI 基础设施方向)
8 hours ago
≈ $27k – $68k per year (Estimated) • In office • Full-Time • 5+ years exp • Shanghai
AI/ML
TensorFlow
PyTorch
KV Cache
DevOps
Amazon S3
Linux
Unix
Apply
This is one of many
778,720 more open roles from verified company boards, updated every day.

