{"id":2119682,"url":"https://alion.io/job/zoom-audio-ai-engineer","title":"Audio AI Engineer","company":{"id":3099,"name":"Zoom","domain":"zoom.com","url":"https://alion.io/company/zoom","size_band":"201-500","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"B","score":75,"open_postings":10,"ghost_share":0,"stale_share":1,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-10-10T05:45:15Z"}},"role":"AI/ML","role_family":"AI/ML","seniority":"junior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Singapore"],"countries":["SG"],"hiring_countries":["SG"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":42000,"max_usd":86000,"period":"year","method":"role_seniority_country_cell","sample_n":24},"experience_years_min":2,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Speech Recognition","optional":false},{"name":"Text-to-Speech","optional":false},{"name":"C++","optional":true},{"name":"Diffusion Models","optional":true},{"name":"Knowledge Distillation","optional":true},{"name":"Model Distillation","optional":true},{"name":"Python","optional":true},{"name":"PyTorch","optional":true},{"name":"PyTorch C++","optional":true},{"name":"Quantization","optional":true},{"name":"TensorFlow","optional":true},{"name":"TensorFlow C++","optional":true},{"name":"Transformers","optional":true}],"status":"live","first_seen_at":"2026-07-31T00:00:00Z","employer_posted_date":"2026-07-31","last_verified_at":"2026-10-11T19:09:34Z","board_verified":true,"closed_at":null,"days_open":72,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":72},"description":"What you can expect\nAs an Audio AI Engineer, you will research and develop algorithms for accent conversion, voice conversion, speech synthesis, and speech recognition on low-latency streaming architectures. You’ll prototype and refine end-to-end audio models that enhance intelligibility and naturalness while maintaining speaker identity. Working closely with product and platform teams, you’ll help bring these models into real-time communication systems. You will also evaluate and optimize model performance across dimensions such as quality, latency, and scalability. Staying current with advances in speech processing, you’ll contribute to innovation through patents and internal knowledge sharing.\nAbout the Team\nZoom's Audio team develops real-time audio features based on AI algorithms. Members of the team are spread worldwide, including the U.S., China and Singapore.\nResponsibilities\nResearching, designing, and developing algorithms for accent conversion, voice conversion, speech synthesis, and automatic speech recognition, focusing on low-latency streaming architectures.\n\nPrototyping end-to-end audio models that enhance intelligibility and naturalness while preserving speaker identity and expressiveness.\n\nCollaborating closely with product and platform teams to integrate models into real-time video and audio communication systems.\n\nAnalyzing and optimizing model performance across speech quality, latency, robustness, and scalability dimensions.\n\nStaying current with the latest developments in speech processing research, and contribute to the community through patents, and internal knowledge sharing.\n\nWhat we’re looking for\nHold a PhD or equivalent experience in a relevant field in Streaming, Accent Conversion, Voice Conversion, TTS, or ASR. More than 2 years of relevant industry experience considered a plus.\n\nShow proficiency in deep learning frameworks like PyTorch or TensorFlow.\n\nDemonstrate effective programming skills in Python, C/C++, or similar languages.\n\nHave an understanding of sequence modeling architectures (Transformers, RNNs, diffusion models, or conformers).\n\nDemonstrate experience developing and deploying low-latency, real-time speech or audio models with streaming architectures and optimized pipelines.\n\nShow familiarity with model compression and acceleration techniques, including quantization, pruning, and distillation.\n\nExhibit experience working with real-time audio systems in networked communication environments.\n\nPublish in top-tier conferences such as Icassp, Interspeech, NeurIPS, and ICLR.\n\nWays of Working\nOur structured hybrid approach is centered around our offices and remote work environments. The work style of each role, Hybrid, Remote, or In-Person is indicated in the job description/posting.\nBenefits\nAs part of our award-winning workplace culture and commitment to delivering happiness, our benefits program offers a variety of perks, benefits, and options to help employees maintain their physical, mental, emotional, and financial health; support work-life balance; and contribute to their community in meaningful ways. Click Learn for more information.\nAbout Us\nZoomies help people stay connected so they can get more done together. We set out to build the best collaboration platform for the enterprise, and today help people communicate better with products like Zoom Contact Center, Zoom Phone, Zoom Events, Zoom Apps, Zoom Rooms, and Zoom Webinars.\nWe’re problem-solvers, working at a fast pace to design solutions with our customers and users in mind. Find room to grow with opportunities to stretch your skills and advance your career in a collaborative, growth-focused environment.\nOur Commitment\nAt Zoom, we believe great work happens when people feel supported and empowered. We’re committed to fair hiring practices that ensure every candidate is evaluated based on skills, experience, and potential. If you require an accommodation during the hiring process, let us know-we’re here to support you at every step.\nIf you need assistance navigating the interview process due to a medical disability, please submit an Accommodations Request Form and someone from our team will reach out soon. This form is solely for applicants who require an accommodation due to a qualifying medical disability. Non-accommodation-related requests, such as application follow-ups or technical issues, will not be addressed.\nOur interviews are supported by BrightHire, a tool that helps us create a consistent and thoughtful interview experience and may include recordings. Please refer to our candidate privacy statement for more information of how we use your data.","description_format":"text","description_chars":4640,"description_truncated":false,"requirements":{"experience_years_min":2,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"phd","optional":false},"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"Singapore","iso":"SG","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["E-learning","Virtual Events","Call Center Services","Contact Center Software"],"lifecycle":[{"event":"open","at":"2026-10-08T23:51:15Z"}],"visa":[],"liveness":{"score":8,"band":"cold","label":"Long shot","p_open":1,"p_active":0.301,"p_room":0.28,"age_days":71,"expected_fill_days":31,"reasons":["conf:2","stale_co","velocity","win:tail","crowd:junior,brand"],"computed_at":"2026-10-10T05:45:15Z"},"pay":null,"html_url":"https://alion.io/job/zoom-audio-ai-engineer","json_url":"https://alion.io/job/zoom-audio-ai-engineer.json","meta":{"generated_at":"2026-10-11T19:10:44Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler_verified","counted_by":"address","units_charged":1,"used_today":5759,"day_limit":null,"remaining_today":null,"minute_limit":300,"resets_at":"2026-10-12T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":3099},"rest":"https://alion.io/mcp/rest/get_company?id=3099"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fzoom-audio-ai-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fzoom-audio-ai-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fzoom-audio-ai-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/zoom-audio-ai-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fzoom-audio-ai-engineer"}]}