803,192open jobs
51,523companies
126,647added this week
Browse all
Salary
≈ $82k – $170k per year (Estimated)
Location
In office (San Jose)
Seniority
Junior
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Jul 20, 2026.

Overview
Company
Impact
Profile match
TikTok is a short-form video platform owned by the Chinese technology group ByteDance and launched internationally in 2016, growing further after the 2018 merger with the lip-sync app Musical.ly. Its recommendation feed surfaces content by engagement signals rather than social graph, which lets new creators reach large audiences quickly and made the app one of the most downloaded in the world. The company runs dual operating hubs in Singapore and Los Angeles and has built an advertising business, a live streaming economy and the TikTok Shop commerce marketplace on top of the video feed.

岗位职责 / Responsibilities

About the Team

We are dedicated to building the inference infrastructure for ultra-large-scale language models, vision-language models, and frontier multimodal AI systems. Our mission is to provide a robust, scalable, and high-performance foundation for distributed serving, heterogeneous scheduling, and low-latency inference at massive scale. You will work on some of the most challenging problems in large-model online serving, spanning traffic orchestration, throughput and latency optimization, kernel efficiency, and production reliability for next-generation AI systems.

We are looking for talented individuals to join our team. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth.

Successful candidates must be able to commit to an onboarding date by the end of the year. Please state your availability and graduation date clearly in your resume.

Responsibilities - What You'II Do

- Build and evolve next-generation inference systems for large-scale online traffic, including global scheduling across heterogeneous compute resources, high-concurrency load balancing, and efficient batch formation

- Optimize distributed inference for 200B+ models and complex multimodal models through TP, EP, DP, and related strategies to improve throughput and latency in production

- Develop high-performance kernels for frontier model architectures such as MoE, emerging attention mechanisms, and multimodal fusion layers using CUDA, Triton, and related tools

- Explore AI-driven infrastructure for inference systems, including AI Agents for kernel optimization, performance tuning, consistency validation, deployment pipelines, and intelligent operations

任职要求 / Requirements

Minimum Qualifications:

- Individuals who are completing or have recently completed a PhD degree in Software Development, Computer Science, Computer Engineering, or a related technical discipline. in Computer Science, Software Engineering, Artificial Intelligence, Mathematics, or related fields

- Familiarity with large-model architectures and strong system design skills for complex, high-concurrency environments

- Strong understanding of asynchronous scheduling, resource pooling, and load balancing in distributed microservice systems

- Strong engineering skills in performance optimization and production system development

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
803,192 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
San Jose
≈ $88k – $183k per year (Estimated) • Hybrid • 10+ years exp • Bachelor's Degree • Washington
AI/ML
Machine Learning
DevOps
AWS
Management
Agile
Apply
≈ $69k – $144k per year (Estimated) • Equity • In office • Full-Time • Bachelor's Degree • Exeter
Apply
≈ $86k – $180k per year (Estimated) • Equity • In office • Full-Time • 2+ years exp • Bachelor's Degree • Exeter
Apply
$129k – $175k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Arlington
Python
Go
Java
Rust
Ruby
C#
C++
DevOps
Splunk
New Relic
Datadog
CI/CD
AWS
Amazon CloudWatch
Linux
Unix
Management
Agile
Apply
$187k – $201k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Cupertino
Python
PHP
SQL
Ruby
PowerShell
Bash
Perl
DevOps
Linux
Windows
Unix
Apply
$100k – $150k per year • Remote (United States) • 6+ years exp • Master's Degree
Python
AI/ML
Fine-tuning
JAX
Multimodal AI
AI Agents
PyTorch
RAG
Machine Learning
Apply
$100k – $150k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
C++
C++
PyTorch C++
LLVM
AI/ML
vLLM
CUDA Toolkit
JAX
TensorRT
PyTorch
CUDA
Triton
NCCL
MLIR
CUTLASS
DevOps
HPC
Apply
$85k – $110k per year • Remote (United States) • 6+ years exp • Bachelor's Degree
C++
C++
LLVM
AI/ML
vLLM
CUDA Toolkit
TensorRT
CUDA
Triton
NCCL
MLIR
CUTLASS
Apply
≈ $119k – $260k per year (Estimated) • Remote (United States) • 6+ years exp • Master's Degree
Python
AI/ML
Fine-tuning
JAX
Multimodal AI
AI Agents
PyTorch
RAG
Machine Learning
Apply
$80k – $100k per year • Remote (United States) • 6+ years exp • Master's Degree
Python
AI/ML
RLHF
Reinforcement Learning
AI Agents
Reward Modeling
Machine Learning
Robotics
Reinforcement Learning
Apply
In office • Internship • Bachelor's Degree • San Jose
Python
Java
C++
AI/ML
Machine Learning
DevOps
Linux
Apply
In office • Internship • Bachelor's Degree • San Jose
Python
Java
C++
DevOps
SLI/SLO/SLA
Linux
Unix
Apply
In office • Internship • PhD • San Jose
AI/ML
CUDA Toolkit
Multimodal AI
AI Agents
Mixture of Experts
CUDA
Triton
Apply
≈ $75k – $157k per year (Estimated) • In office • Full-Time • Bachelor's Degree • San Jose
Python
Java
C++
AI/ML
Recommender Systems
DevOps
SLI/SLO/SLA
Linux
Unix
Apply
≈ $114k – $251k per year (Estimated) • In office • Full-Time • Bachelor's Degree • San Jose
Python
Go
Java
C++
DevOps
Linux
TCP/IP
DNS
Apply
$4k – $14k per year • In office • 3+ years exp • PhD • San Jose
Apply
$177k – $239k per year • Equity • In office • Full-Time • 10+ years exp • San Jose
DevOps
AWS
Apply
≈ $44k – $90k per year (Estimated) • In office • 1+ year exp • High School Diploma • San Jose
DevOps
Windows
VPN
Management
Microsoft Office
Apply
Data Engineer-6336 1 day ago
≈ $122k – $209k per year (Estimated) • Remote (United States) • Contractor • 7+ years exp • Bachelor's Degree • San Jose
Python
SQL
Databases
MySQL
PostgreSQL
Presto
Analytics
ETL/ELT
Apply
$109k – $201k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • San Jose
Python
SQL
Analytics
A/B Testing
Apply
See all jobs
This is one of many
803,192 more open roles from verified company boards, updated every day.