368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$250k – $350k per year
Location
In office (Palo Alto)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Pika is a generative video company founded in 2023 by Stanford researchers who left their doctoral programmes to build it. Its product creates and edits short video from text, images and effects templates aimed at social creators rather than professional studios. The company raised funding from Lightspeed and Greylock and grew a large consumer community.

About the Role

We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model deployment, and video generation technologies. Your expertise will drive significant improvements to model speed and efficiency, ensuring our creative AI systems deliver industry-leading user experiences at scale.

You will design and optimize inference pipelines, implement state-of-the-art acceleration techniques, and work closely with researchers and engineers across the team to push the boundaries of what’s possible in real-time AI deployment. Your efforts will play a foundational role in powering the next generation of Pika’s video and language models.

What You’ll Do

  • Accelerate Inference: Lead and implement advanced inference acceleration techniques, including attention optimization and quantization for efficient model serving.

  • Maximize GPU Parallelism: Engineer and optimize GPU strategies across tensor, sequence, and pipeline parallelism (TP, SP, PP) for maximal efficiency and scalability.

  • Programming for Performance: Develop and optimize high-performance computing kernels and distributed workloads using CUDA and NCCL.

  • Advance AI Deployment: Collaborate with research and engineering teams to bring state-of-the-art videogen and large language models into production.

  • Improve Training Efficiency: (Bonus) Contribute to improvements in model training speed, stability, and resource utilization as part of our deployment lifecycle.

  • Technical Excellence: Drive rigorous code reviews, participate in technical discussions, and mentor fellow engineers on best practices in inference and GPU programming.

What We’re Looking For

  • Experience: 5+ years engineering experience, with a strong track record in inference acceleration and model deployment at scale.

  • Inference Mastery: Proven expertise in inference optimization, including quantization, attention acceleration, and deep learning compiler stacks.

  • GPU & Parallelism: Deep knowledge of GPU programming (CUDA, NCCL) and experience with SP, TP, PP, and other forms of parallelism for distributed inference.

  • AI Domain Knowledge: Familiarity with video generation (videogen) models and large language models (LLMs).

  • Collaboration: Strong cross-discipline communication skills; able to drive shared goals across research and engineering functions.

  • Ownership Mindset: Self-driven, solutions-oriented, and capable of managing ambiguity in a fast-paced startup environment.

  • Bonus: Experience in enhancing training efficiency, stability, or resource optimization for large models.

Nice to Have

  • Experience with high-throughput video or real-time streaming model deployment

  • Familiarity with distributed training and optimization toolkits

  • Contributions to open source projects in AI infrastructure or deep learning compilers

  • Startup or rapid prototyping experience

What We Offer

  • Competitive salary in the AI industry

  • Equity in a fast-growing startup shaping the future of AI

  • Comprehensive health benefits, monthly stipends, company retreats

  • A supportive and collaborative office culture-we’re all building and launching together

About Pika

At Pika, we're crafting a future where video creation is seamless, intuitive, and universally accessible. Our mission is to empower creativity by breaking down technical barriers using the transformative power of AI. We’re a tight-knit, energetic team based in Palo Alto, CA, valuing efficiency, curiosity, and the ambition to make a meaningful impact on the world.

We work from our Palo Alto office 3-5 days a week and welcome applicants who are eager to contribute onsite.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Palo Alto
$20k – $43k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Ahmedabad
AI/ML
AI Agents
Edge AI
Apply
Remote/Hybrid • Full-Time • Bachelor's Degree • Gurgaon
Python
Visual Basic
AI/ML
AI Agents
Edge AI
Analytics
Power BI
Tableau
ETL/ELT
Apply
$12k – $30k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Mumbai
Java
AI/ML
AI Agents
Edge AI
DevOps
AWS
Azure
Apply
AI Engineer 11 hours ago
$26k – $108k per year (Estimated) • In office • Full-Time • Pune
Python
COBOL
Python
FastAPI
COBOL
IBM MQ
Databases
DynamoDB
AI/ML
AutoGen
AWS Bedrock
Claude
CrewAI
Embeddings
LangChain
LangGraph
Prompt Engineering
RAG
Edge AI
LLM Guardrails
AI Agents
DevOps
Amazon EKS
AWS
AWS Lambda
CI/CD
Docker
Kubernetes
Amazon CloudWatch
Amazon ECS
Amazon S3
API Gateway
IAM
Apply
$37k – $110k per year (Estimated) • Equity • Remote/Hybrid • Internship • 1+ year exp • Bachelor's Degree • Madrid
Python
SQL
Databases
Databricks
AI/ML
AI Agents
Context Engineering
Edge AI
Fine-tuning
Function Calling
LangChain
LlamaIndex
LLM
Multimodal AI
OpenAI
Prompt Engineering
PyTorch
RAG
Robotics
Digital Twin
Apply
Product Manager 1 month ago
$160k – $225k per year • In office • Full-Time • 2+ years exp • Palo Alto
AI/ML
Pika
Apply
$185k – $400k per year • In office • Full-Time • 5+ years exp • Palo Alto
Python
SQL
Python
pySpark
AI/ML
Hadoop
Multimodal AI
Pika
Ray
Spark
Pre-training
AI Agents
DevOps
AWS
Azure
GCP
Analytics
ETL/ELT
Apply
$250k – $350k per year • In office • Full-Time • 5+ years exp • Palo Alto
Go
Node JS
TypeScript
JavaScript
AI/ML
Pika
AI Agents
Frontend
React.js
DevOps
AWS
GCP
Apply
$250k – $350k per year • In office • Full-Time • 5+ years exp • Palo Alto
AI/ML
Pika
DevOps
AWS
GCP
Kubernetes
Platform Engineering
Apply
$250k – $350k per year • In office • Full-Time • 5+ years exp • Palo Alto
C++
Go
Python
C++
TensorFlow C++
AI/ML
Multimodal AI
Pika
TensorFlow
Triton
TorchServe
DevOps
AWS
Azure
GCP
Kubernetes
Platform Engineering
Apply
$110k – $240k per year (Estimated) • In office • Bachelor's Degree • Palo Alto
Java
Python
Scala
AI/ML
AI Agents
Fine-tuning
LLM Guardrails
DevOps
AWS
Azure
GCP
Git
GitHub
Marketing
Salesforce
Apply
Chief of Staff 8 hours ago
$120k – $150k per year • Equity 0.4–0.7% • In office • Full-Time • 3+ years exp • Palo Alto
AI/ML
AI Agents
Apply
$140k – $310k per year (Estimated) • Remote/Hybrid • Bachelor's Degree • Palo Alto
Databases
Apache Kafka
NATS
DevOps
AWS
Azure
CI/CD
Docker
GCP
Grafana
gRPC
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Robotics
EtherCAT
IoT
MQTT
OPC UA
Apply
$139k – $294k per year (Estimated) • In office • Palo Alto
Python
AI/ML
Fine-tuning
Hybrid Search
LLM
Prompt Engineering
RAG
Human-in-the-Loop
Knowledge Graph
AI Agents
Function Calling
DevOps
AWS
Apply
$137k – $292k per year (Estimated) • In office • Palo Alto
Python
AI/ML
Hybrid Search
LLM
Prompt Engineering
RAG
Human-in-the-Loop
AI Agents
Function Calling
DevOps
AWS
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.