{"id":1976267,"url":"https://alion.io/job/zoox-engineering-manager-ml-performance-optimization","title":"Engineering Manager, ML Performance Optimization","company":{"id":3626,"name":"Zoox","domain":"zoox.com","url":"https://alion.io/company/zoox","size_band":"1001-5000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Lever","truth_index":{"grade":"B","score":72,"open_postings":38,"ghost_share":0,"stale_share":0.921,"repost_share":0,"time_to_fill_p50_days":87,"computed_at":"2026-10-08T05:49:30Z"}},"role":"AI/ML","role_family":"AI/ML","seniority":"lead","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Foster City, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":276000,"max":343000,"currency":"USD","period":"year","gross":null,"usd_annual":343000},"salary_estimate":null,"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Apache TVM","optional":false},{"name":"CUDA","optional":false},{"name":"CUDA Toolkit","optional":false},{"name":"FSDP","optional":false},{"name":"JAX","optional":false},{"name":"Knowledge Distillation","optional":false},{"name":"Model Distillation","optional":false},{"name":"Multimodal AI","optional":false},{"name":"PyTorch","optional":false},{"name":"Quantization","optional":false},{"name":"Ray Serve","optional":false},{"name":"TensorRT","optional":false},{"name":"Triton","optional":false},{"name":"XLA","optional":false},{"name":"Ray","optional":true}],"status":"live","first_seen_at":"2026-10-06T20:39:51Z","employer_posted_date":"2026-10-06","last_verified_at":"2026-10-09T01:23:58Z","board_verified":true,"closed_at":null,"days_open":2,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":2},"description":"Zoox is on a mission to reimagine transportation and build autonomous robotaxis from the ground up that are safe, reliable, clean, and enjoyable for everyone. With bidirectional driving capabilities and four-wheel steering, our vehicle allows us to maneuver through compact spaces and change directions without needing to reverse. We are at a critical inflection point as we scale our robotaxi deployment, and it is a great time to join Zoox and have a significant impact on executing our mission.\nOur growing ML performance engineering leadership team is looking for an Engineering Manager, ML Performance Optimization. The centralized ML performance team at Zoox plays a crucial role in enabling innovations across all our ML research teams to develop and deploy models across our robotaxi and cloud infrastructure and to advance cutting-edge training and inference optimization techniques.\nThe Opportunity\nWe are working on many interesting challenges to enable rapid experimentation and scale our multi-modal Foundation models and RL infrastructure, and ensure these models run efficiently on our vehicles, meeting our latency targets. You will get to work across all ML teams within Zoox - Perception, Prediction, Planner, Simulation and our Advanced Hardware Engineering group, and have the opportunity to significantly push the boundaries of ML scaling and acceleration at Zoox.\nYou will lead a team of strong ML performance engineers and this team has many growth opportunities as we expand our geofences at U.S markets and venture into new ML domains. If you want to learn more about our ML Infrastructure, here is one of our past talks at re:Invent.\nIn this role, you will:\nVision: Develop and execute a strategic vision and roadmap for ML Training and Inference Performance Optimization, ensuring scalability, reliability, and performance to support autonomous driving.\n\nTechnical acumen: Lead the design, implementation, and operation of a robust and efficient ML platform to enable the training, validation, serving, optimization and monitoring of ML models.\n\nML Performance Optimization: Drive end-to-end performance optimization for large-scale model training and inference, including distributed training efficiency, GPU utilization, memory and communication optimization, model compression (quantization, pruning, distillation), and low-latency on-vehicle inference that meets strict real-time and compute budgets.\n\nHiring: Attract, hire, and inspire a diverse world-class engineering team, fostering a culture of innovation, collaboration, and excellence.\n\nPartnership: Collaborate closely with cross-functional teams, including ML researchers, software engineers, data engineers, and hardware engineers, to define requirements and align on architectural decisions.\n\nMentorship: Enable engineers on the team to grow their careers by providing the right opportunities and clear, timely feedback.\n\nQualifications\n8+ years of relevant experience, including 3+ years of management experience managing engineers.\nStrong technical background in ML performance optimization, such as distributed training strategies (data, tensor, pipeline parallelism, FSDP/ZeRO), mixed-precision training, kernel-level optimization (CUDA, Triton), compiler stacks (torch.compile, XLA, TVM), quantization, and profiling/benchmarking across GPU and embedded accelerators.\nExperience building user-friendly ML Infrastructure that enabled large-scale model training and high-throughput, low-latency serving use cases.\nExperience with training frameworks like PyTorch, JAX, etc., leveraging GPUs for distributed model training.\nExperience with GPU-accelerated inference using TensorRT, Ray Serve, or similar frameworks.\nProven track record of extensive cross-functional collaboration, partnering with research, product, hardware, and platform teams to align priorities, influence technical direction, and deliver measurable performance improvements across organizational boundaries.","description_format":"text","description_chars":3972,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Growth opportunities"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Autonomous Driving","Smart Mobility","Automotive Manufacturing","Ride Hailing"],"lifecycle":[{"event":"open","at":"2026-10-06T21:28:37Z"}],"visa":[{"country":"US","licensed_sponsor":true,"evidence":"H-1B filings in 12 months: 373 · green card filings: 120","filings_12m":373,"filings_prev_12m":311,"green_card_filings_12m":120,"median_offered_wage_usd":181283,"route":null,"cap_exempt":false,"checked_at":"2026-10-03T21:08:04+00:00","sources":["US Department of Labor: LCA disclosure data (H-1B, H-1B1, E-3)","US Department of Labor: PERM disclosure data (green cards)"],"filings_for_role_12m":169}],"liveness":{"score":63,"band":"ok","label":"Likely open","p_open":1,"p_active":0.632,"p_room":1,"age_days":1,"expected_fill_days":87,"reasons":["conf:1","stale_co","velocity","win:early","comp:brand"],"computed_at":"2026-10-08T05:49:30Z"},"pay":{"stated_usd_annual":343000,"is_top_pay":true},"html_url":"https://alion.io/job/zoox-engineering-manager-ml-performance-optimization","json_url":"https://alion.io/job/zoox-engineering-manager-ml-performance-optimization.json","meta":{"generated_at":"2026-10-09T02:46:13Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1092,"day_limit":5000,"remaining_today":3908,"minute_limit":60,"resets_at":"2026-10-10T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":3626},"rest":"https://alion.io/mcp/rest/get_company?id=3626"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fzoox-engineering-manager-ml-performance-optimization"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fzoox-engineering-manager-ml-performance-optimization"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fzoox-engineering-manager-ml-performance-optimization"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/zoox-engineering-manager-ml-performance-optimization\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fzoox-engineering-manager-ml-performance-optimization"}]}