549,821open jobs
20,529companies
75,324added this week
Browse all
Salary
$168k – $205k per year
Location
In office (Redwood City)
Seniority
Junior · 2+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Headquartered in Redwood City, California, Ambient AI is an enterprise physical security company specializing in AI-powered computer vision solutions. The company's software platform integrates with existing video surveillance infrastructure and access control systems to monitor feeds continuously, detect real-time threat signatures, and drastically reduce false alarms.

Build a safer world with us, one incident at a time.

Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for physical security, our platform seamlessly integrates with existing security cameras and physical access control systems to unify monitoring, access control, threat assessment, response, and investigations through an always-on reasoning layer that augments security operators with superhuman capabilities. The results: 95% fewer false alarms, investigations 20x faster, and 10x faster response.

The momentum speaks for itself: we doubled new ARR in FY26, and have delivered results for world-class customers including Cisco, ServiceNow, SentinelOne, TikTok, Bayer, and MoMA. That kind of momentum creates an environment where great people thrive, and it shows: we recently ranked #71 out of 500 on the Forbes best startup employers list.

Founded in 2017 and backed by Andreessen Horowitz, Y Combinator, and Allegion Ventures, Ambient.ai is on a fast-paced journey to fulfill our mission: prevent every security incident possible.

Ready to learn more? Connect with us on LinkedIn and YouTube

About the role:

Reporting to Raghu Nallamothu, you will design, build, and optimize the AI infrastructure that powers Ambient.ai ’s real-time intelligence platform.

In this role, you will work on the systems required to run state-of-the-art deep learning models across many terabytes of video data in real time. You will help build and scale infrastructure for inference, evaluation, and continuous model improvement across computer vision models, large language models, large vision models, and multimodal AI systems.

This role is ideal for someone with a strong blend of infrastructure engineering, production ML systems, LLM/LVM inference, evaluation harnesses, and inference optimization experience. You will partner closely with research scientists and product engineering teams to bring the latest AI advancements into production for our customers.

What you'll do:

  • Design, build, and maintain cutting-edge AI infrastructure for real-time computer vision, LLM, LVM, and multimodal inference workloads.

  • Build scalable systems for running state-of-the-art models across large volumes of video and sensor data.

  • Optimize inference performance across latency, throughput, GPU utilization, reliability, and cost.

  • Develop robust evaluation harnesses and benchmarking systems to measure model quality, system performance, regressions, and production readiness.

  • Build infrastructure for continuous model evaluation, experimentation, and deployment.

  • Partner with research scientists to productionize the latest advances in computer vision, LLMs, LVMs, RAG, and multimodal AI.

  • Improve model-serving architecture, including batching, caching, routing, quantization, model parallelism, and hardware utilization.

  • Develop data engines and feedback loops for collecting training data, evaluating model behavior, and continuously improving AI performance.

  • Create reliable observability, monitoring, and debugging tools for production AI systems.

  • Help define best practices for deploying, evaluating, and operating AI systems in real-world enterprise environments.

What you'll bring:

  • 2+ years of industry experience building infrastructure, distributed systems, machine learning platforms, or production AI systems.

  • BS/MS in Computer Science or a related technical field, or equivalent practical experience.

  • Strong programming background, especially in Python, with solid software engineering fundamentals.

  • Experience designing and building scalable machine learning infrastructure for training, inference, evaluation, and deployment.

  • Hands-on experience running deep learning models in production, ideally including LLMs, LVMs, vision-language models, or multimodal models.

  • Strong understanding of inference optimization techniques, including batching, caching, quantization, parallelism, memory optimization, GPU utilization, and latency reduction.

  • Experience with model-serving frameworks or systems such as vLLM, Triton Inference Server or similar technologies.

  • Experience building evaluation frameworks, test harnesses, benchmarks, regression tests, or model-quality measurement systems.

  • Strong background in machine learning and deep learning; computer vision experience is a strong plus.

  • Experience designing data engines or pipelines for collecting, managing, and curating training and evaluation data.

  • Familiarity with integrating advanced AI systems such as LLMs, LVMs, RAG pipelines, embedding models, or multimodal models into production applications.

  • Experience with cloud infrastructure, containers, orchestration, distributed systems, and GPU-based workloads.

  • Strong collaboration and communication skills, with the ability to work effectively with research scientists, product teams, infrastructure teams, and stakeholders.

  • Proactive problem-solving ability, a strong ownership mindset, and adaptability to incorporate new AI technologies and methodologies.

Nice to Have

  • Experience operating large-scale GPU infrastructure or distributed inference systems.

  • Experience with CUDA, NCCL, PyTorch, TensorRT, ONNX, or similar ML systems technologies.

  • Experience with video understanding, real-time computer vision, multimodal AI, or physical-world AI systems.

  • Experience with model compression, speculative decoding, distillation, pruning, or low-latency serving techniques.

  • Experience with prompt evaluation, model regression testing, human-in-the-loop evaluation, or automated quality gates.

  • Familiarity with retrieval-augmented generation, vector databases, embedding models, re-rankers, or search infrastructure.

  • Experience building internal ML platforms or tools used by researchers and applied ML teams.

What Success Looks Like

You will be successful in this role if you can build practical, scalable infrastructure that helps Ambient.ai deploy better AI models faster and more reliably. You should be comfortable working across the full stack of production AI systems, from model behavior and evaluation to serving architecture, GPU performance, observability, and customer-facing reliability.

This is a hands-on engineering role for someone excited to help bring the next generation of AI, computer vision, LLMs, and LVMs into real-world production environments.

Why join us:

  • We are creating an entirely new category within a 180+ billion-dollar physical security industry and looking for team members who are also passionate about our mission to prevent every security incident possible

  • We partner with an incredible customer roster of F500 companies, including Adobe, TikTok, Gap and SentinelOne

  • Regular Full-time employees receive stock options for the opportunity to share ownership in the success of our company

  • Comprehensive health + welfare package (Medical, Dental, Vision, Life, EAP, Legal Services, 401k plan)

  • We offer flexible time off to rest and recharge, including Winter Break (time off between Christmas and New Year’s for most roles, depending on customer demand)

  • The latest tech and awesome swag will be delivered to your door

  • Enjoy a full range of opportunities to connect with your awesome co-workers

  • We love tohike, are foodies, and love music! Check out our most recentAmbient Spotify Playlist

We’ve found that in-person time meaningfully supports collaboration, creativity, and team alignment. Our talent, engineering, product, design, and marketing teams work from our Redwood City office three days a week. All other Bay Area employees join on Fridays to stay connected and close out the week together.

Ready to learn more? Connect with us onLinkedIn | YouTube

Ambient.ai is proud to be an Equal Opportunity Employer. Ambient does not unlawfully discriminate on the basis of race, color, religion, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), gender identity, gender expression, national origin, ancestry citizenship, age, physical or mental disability, legally protected medical condition, family care status, military or veteran status, marital status, registered domestic partner status, sexual orientation, genetic information, or any other basis protected by local, state, or federal laws. Ambient is an E-Verify participant.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
549,821 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Redwood City
$87k – $184k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Toronto
Python
Apply
In office • 3+ years exp • Bachelor's Degree
Python
Verilog
Perl
AI/ML
Edge AI
Chips/EDA
Synopsys PrimeTime
Synopsys Fusion Compiler
Synopsys IC Compiler II
Apply
$154k – $257k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Vancouver
Python
C++
Apply
In office • 1+ year exp • Bachelor's Degree
Python
Apply
In office • 10+ years exp • Bachelor's Degree
Python
Java
C#
MATLAB
Apply
$48k per year • In office • Full-Time • Redwood City
AI/ML
AI Agents
Cybersecurity
SentinelOne
Management
Slack
ServiceNow
Apply
RVP, Sales - East 2 days ago
$200k – $215k per year • Equity • In office • Full-Time • 10+ years exp • New York
AI/ML
AI Agents
DevOps
VMWare
Cybersecurity
SentinelOne
Management
ServiceNow
Apply
RVP, Sales - East 2 days ago
$200k – $215k per year • Equity • Remote • Full-Time • 10+ years exp
AI/ML
AI Agents
DevOps
VMWare
Cybersecurity
SentinelOne
Management
ServiceNow
Apply
RVP, Sales - East 2 days ago
$200k – $215k per year • Equity • In office • Full-Time • 10+ years exp • Boston
AI/ML
AI Agents
Cybersecurity
SentinelOne
Management
ServiceNow
Apply
Video Editor 8 days ago
$120k – $130k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Redwood City
AI/ML
AI Agents
Cybersecurity
SentinelOne
Design
Adobe Photoshop
Adobe Premiere Pro
Figma
Adobe After Effects
Management
ServiceNow
Marketing
YouTube
Spotify
LinkedIn
Apply
Scientist I 11 hours ago
$115k – $158k per year • Remote/Hybrid • Full-Time • PhD • Redwood City
Python
SQL
Apply
Store Associate 11 hours ago
In office • Full-Time • Redwood City
Management
Slack
Outlook
Apply
$147k – $170k per year • In office • 7+ years exp • Redwood City
AI/ML
Claude
ChatGPT
AI Agents
Apply
$48k per year • In office • Full-Time • Redwood City
Apply
$160k – $200k per year • Remote/Hybrid • 7+ years exp • Bachelor's Degree • Redwood City
AI/ML
AI Agents
PyTorch
LLM
Ignite
Apply
See all jobs
This is one of many
549,821 more open roles from verified company boards, updated every day.