1,435,005open jobs
84,692companies
218,397added this week
Browse all
Salary
$202k – $224k per year
Location
In office (Sunnyvale)
Seniority
Senior · 5+ years exp

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 8, 2026. Uber scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Uber is an American mobility and delivery company founded in San Francisco in 2009 that connects independent drivers and couriers with riders and customers through its apps. It operates ride hailing in thousands of cities, a food and grocery delivery business built on the same courier network, a freight brokerage matching shippers with carriers, and an advertising business that sells placement inside those apps. Listed on the New York Stock Exchange since 2019, the company reached sustained profitability in the mid 2020s and now partners with autonomous vehicle developers rather than building self-driving technology itself.

Senior Software Engineer

About the Role

We are seeking talented Senior Software Engineers to join our Search Engineering team and help build the next generation of AI-powered search experiences.

In this role, you will design and build large-scale backend and model-serving infrastructure that powers search retrieval, ranking, personalization, and emerging LLM-based search experiences. You will work on systems that operate at high request volumes with strict latency and reliability requirements, spanning traditional search infrastructure, machine learning model serving, and modern LLM inference.

You will collaborate closely with backend and ML engineers, data scientists, product managers, and platform teams to evolve the search stack toward more intelligent, real-time, and AI-native architectures. This includes integrating LLM-based ranking and retrieval, optimizing GPU inference and serving efficiency, incorporating real-time marketplace signals, and building scalable infrastructure that enables rapid experimentation while maintaining production-grade performance and reliability.

Basic Qualifications

  • 5+ years of professional software engineering experience building large-scale backend or distributed systems.
  • Strong programming skills in Go, Java, C++, Python, or a similar language.
  • Strong understanding of distributed systems, service-oriented architectures, concurrency, networking, caching, and data consistency.
  • Experience building high-throughput, low-latency online serving systems.
  • Experience with search, recommendation, ranking, machine learning serving, or other large-scale data-intensive systems.
  • Experience diagnosing and optimizing system performance across latency, throughput, reliability, and infrastructure efficiency.
  • Familiarity with distributed data processing and streaming technologies such as Kafka, Flink, Spark, or similar frameworks.
  • Experience operating production systems, including observability, monitoring, capacity planning, incident response, and reliability engineering.
  • Strong system design, problem-solving, and analytical skills with the ability to work across multiple layers of a complex production stack.

Preferred Qualifications

  • Experience building search and recommendation systems, including retrieval, ranking, query understanding, indexing, and personalization.
  • Hands-on experience with search technologies such as Elasticsearch, OpenSearch, Solr, Vespa, Lucene, or large-scale proprietary search systems.
  • Experience with LLM or ML model-serving infrastructure, including frameworks such as vLLM, Triton, or similar inference platforms.
  • Experience optimizing GPU-based inference workloads, including batching or micro-batching, request scheduling, model parallelism, memory management, and GPU utilization.
  • Familiarity with LLM serving concepts such as prefill/decode, KV caching, prefix caching, streaming generation, speculative techniques, and distributed inference.
  • Experience with embeddings, semantic retrieval, approximate nearest-neighbor search, semantic IDs, or generative retrieval.
  • Familiarity with constrained decoding or integrating real-time business and marketplace constraints into AI-powered serving systems.
  • Experience integrating near-real-time features and signals into latency-sensitive ranking or inference systems.
  • Experience designing ML/LLM systems with strong reliability, graceful degradation, experimentation, and launch-safety mechanisms.

What the Candidate Will Do

  • Design and build highly scalable search and AI serving infrastructure with a focus on latency, throughput, reliability, and infrastructure efficiency.
  • Develop the backend architecture for next-generation AI-powered search, including LLM-based retrieval, ranking, personalization, and generative search experiences.
  • Integrate large language models and machine learning models into production search serving paths while meeting stringent latency and reliability requirements.
  • Optimize GPU inference and model-serving performance through techniques such as batching, micro-batching, request routing, caching, streaming, and efficient resource utilization.
  • Build infrastructure for advanced LLM-serving patterns such as context prefill, KV/prefix-cache reuse, progressive or streaming generation, and efficient multi-turn or paginated search experiences.
  • Develop scalable retrieval and ranking systems spanning lexical retrieval, semantic retrieval, embeddings, structured signals, and ML/LLM-based ranking.
  • Work on semantic-ID and constrained-decoding infrastructure that allows generative models to interact safely and efficiently with large-scale search catalogs and real-time marketplace signals.
  • Build and optimize real-time feature retrieval, hydration, caching, and data-processing systems that provide fresh signals to ranking and LLM models.
  • Partner closely with ML engineers and data scientists to productionize new ranking and relevance models, improve experimentation velocity, and shorten the path from model development to production.
  • Continuously improve end-to-end search performance by identifying bottlenecks across retrieval, feature serving, model inference, orchestration, networking, and presentation hydration.
  • Design systems for graceful degradation, observability, capacity management, experimentation, and safe production rollouts of new AI and search capabilities.
  • Analyze production and experiment metrics to understand latency, relevance, reliability, and business trade-offs and use those insights to guide system architecture.
  • Contribute to the long-term technical architecture of the Search platform as it evolves from traditional multi-stage retrieval and ranking toward increasingly unified, AI-native search systems.
  • Write high-quality, maintainable production code and provide technical leadership through design reviews, code reviews, mentoring, and cross-team collaboration.
  • Troubleshoot complex production issues across distributed search and AI-serving systems and drive improvements that prevent recurrence.

For San Francisco, CA-based roles: The base salary range for this role is USD $202,000 per year - USD $224,000 per year.

For Sunnyvale, CA-based roles: The base salary range for this role is USD $202,000 per year - USD $224,000 per year.

For all US locations, you will be eligible to participate in Uber's bonus program, and may be offered an equity award & other types of comp. All full-time employees are eligible to participate in a 401(k) plan. You will also be eligible for various benefits.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,435,005 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Sunnyvale
$180k – $240k per year • Remote (United States) • Contractor • 7+ years exp • Boulder
JavaScript
TypeScript
SQL
C#
C#
.NET
Databases
MySQL
AI/ML
Machine Learning
Frontend
Angular
Mobile
JUnit
DevOps
Rest API
Git
SOAP
Management
Agile
Apply
$132k – $163k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Burbank
C++
Assembly
MATLAB
MATLAB
Simulink
Apply
Product Engineer 6 days ago
≈ $134k – $246k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • United States
Apply
$70k – $235k per year • Hybrid • Full-Time • 12+ years exp • Associate's Degree • Seattle
AI/ML
Function Calling
AI Agents
LLM
LLM Guardrails
Tool Use
Apply
$69k – $131k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Cedar Rapids
Python
Java
C#
C++
DevOps
CI/CD
Git
Docker
Grafana
Configuration Management
Windows
QA
Robot Framework
Apply
≈ $90k – $199k per year (Estimated) • Remote (India) • 8+ years exp
Python
Java
SQL
Scala
Python
pySpark
Databases
Snowflake
Databricks
Delta Lake
Apache Kafka
Google BigQuery
Azure SQL Database
Microsoft Fabric
BigQuery
AI/ML
Spark
DevOps
Rest API
Terraform
Azure
CI/CD
Git
Docker
Kubernetes
Bicep
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Management
Agile
Scrum
Apply
≈ $63k – $159k per year (Estimated) • In office • 5+ years exp • Cairo
Python
SQL
Databases
Snowflake
DynamoDB
Apache Kafka
Amazon Redshift
AI/ML
Spark
Airflow
DevOps
AWS
AWS Lambda
Amazon EC2
Amazon S3
Amazon CloudWatch
Amazon Kinesis
Analytics
ETL/ELT
AWS Glue
Management
Agile
Scrum
Apply
≈ $28k – $62k per year (Estimated) • In office • Full-Time • 5+ years exp • Beijing
Python
Analytics
Power BI
Apply
≈ $144k – $297k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston • Austin
Python
SQL
SAS
AI/ML
Copilot
Prompt Engineering
AI Agents
RAG
OpenAI
LLM Guardrails
Prompt Caching
Machine Learning
DevOps
Azure
Analytics
Tableau
Power BI
ETL/ELT
Management
Agile
Apply
$202k – $224k per year • Equity • In office • 5+ years exp • Sunnyvale
JavaScript
TypeScript
Node JS
Frontend
Vue.js
GraphQL
Angular
React.js
React Query
DevOps
Incident Management
Design
Figma
Apply
$171k – $190k per year • Equity • In office • 3+ years exp • Bachelor's Degree • Seattle
JavaScript
TypeScript
Frontend
React.js
Apply
$171k – $190k per year • Equity • In office • 3+ years exp • Bachelor's Degree • Sunnyvale
Python
Go
C++
Databases
Apache Iceberg
AI/ML
Spark
Flink
Ray
Time Series Forecasting
DevOps
GCP
Prometheus
AWS
Kubernetes
Grafana
Robotics
ROS
Apply
$171k – $190k per year • Equity • In office • 2+ years exp • Bachelor's Degree • Sunnyvale
Python
Go
Java
C++
Databases
Redis
ElasticSearch
Apache Kafka
OpenSearch
Vespa
AI/ML
Spark
Embeddings
Flink
LLM
Machine Learning
Apply
$202k – $224k per year • Equity • In office • 7+ years exp • Bachelor's Degree • Seattle
Java
C++
DevOps
GCP
Azure
AWS
Linux
Apply
≈ $161k – $324k per year (Estimated) • In office • 10+ years exp • Bachelor's Degree • Sunnyvale
DevOps
GCP
Kubernetes
HPC
Apply
≈ $185k – $369k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Sunnyvale
Go
Databases
ElasticSearch
DevOps
Prometheus
Apply
$202k – $224k per year • Equity • In office • 5+ years exp • Sunnyvale
JavaScript
TypeScript
Node JS
Frontend
Vue.js
GraphQL
Angular
React.js
React Query
DevOps
Incident Management
Design
Figma
Apply
$170k – $380k per year • Equity • In office • Top Secret • Full-Time • Sunnyvale
Python
C++
Apply
≈ $220k – $404k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Sunnyvale
Python
Java
SQL
Java
Apache Camel
DevOps
DNS
Cybersecurity
PKI
Management
Agile
Apply
See all jobs
This is one of many
1,435,005 more open roles from verified company boards, updated every day.