{"id":1261594,"url":"https://alion.io/job/amazon-machine-learning-performance-engineer-annapurna-labs","title":"Machine Learning Performance Engineer, Annapurna Labs","company":{"id":19,"name":"Amazon","domain":"amazon.com","url":"https://alion.io/company/amazon","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Career site","truth_index":{"grade":"A","score":100,"open_postings":4,"ghost_share":0,"stale_share":0,"repost_share":0,"time_to_fill_p50_days":18,"computed_at":"2026-10-05T05:45:15Z"}},"role":"AI/ML","role_family":"AI/ML","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Tel Aviv, Israel"],"countries":["IL"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":90000,"max_usd":259000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":585},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"AWS Trainium","optional":false},{"name":"C++","optional":false},{"name":"Diffusion Models","optional":false},{"name":"JAX","optional":false},{"name":"Machine Learning","optional":false},{"name":"Python","optional":false},{"name":"PyTorch","optional":false},{"name":"PyTorch C++","optional":false},{"name":"TensorFlow","optional":false},{"name":"TensorFlow C++","optional":false},{"name":"C","optional":true},{"name":"CUDA","optional":true},{"name":"CUDA Toolkit","optional":true},{"name":"CUTLASS","optional":true},{"name":"LLM","optional":true},{"name":"Mojo","optional":true},{"name":"MPI","optional":true},{"name":"Triton","optional":true}],"status":"closed","first_seen_at":"2026-09-22T00:00:00Z","employer_posted_date":"2026-09-25","last_verified_at":"2026-10-04T10:57:42Z","board_verified":false,"closed_at":"2026-10-04T10:57:42Z","days_open":12,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":12},"description":"The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and generative AI workloads on Amazon's custom machine learning accelerators - Inferentia and Trainium. These chips power workloads for thousands of AWS customers, from large language model training runs to real-time inference serving billions of daily predictions.\nWe are building the first Neuron performance engineering team in Tel Aviv. As a Machine Learning Performance Engineer, you'll help shape the direction of this team from the ground up - profiling and optimizing workloads across the full ML software stack, writing high-performance kernels, and improving the Neuron SDK that external developers depend on. You'll work at the boundary between software and hardware, collaborating directly with compiler, runtime, and chip design engineers to close performance gaps customers care about.\nThe team is new and small, which means broad scope, direct ownership, and real influence over the technical direction we take. If you enjoy digging into performance bottlenecks and turning analysis into measurable wins, this role is for you.\nKey job responsibilities\nDesign and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.\nProfile ML workloads end-to-end to identify bottlenecks - memory, compute, or communication - and drive optimizations through to a measured improvement.\nEnhance the programming model and tooling that kernel and model developers rely on, improving usability and debugging workflows.\nIdentify and drive optimization opportunities across the Neuron software stack (compiler, runtime, frameworks).\nDocument software designs, operational runbooks, and performance findings so the broader team can build on your work.\nA day in the life\nYou might start your morning reviewing profiling data from a customer's large diffusion model training job, tracing a utilization gap back to a specific kernel. After a design discussion with compiler engineers about a new operator fusion strategy, you spend the afternoon writing and benchmarking a kernel prototype. Later, you review a teammate's pull request for a runtime optimization and share your findings in a short write-up for the broader Neuron organization. Your work directly translates into faster model execution and lower cost for AWS customers running ML workloads at scale.\nAbout the team\nThe Neuron Performance Engineering team in Tel Aviv is part of Annapurna Labs within AWS. Our mission is to make sure every ML workload running on Inferentia and Trainium chips reaches its full performance potential. We partner closely with compiler, framework, and hardware teams across Annapurna Labs, and we work directly with AWS customers to understand their models and unblock their adoption.\nWe are a newly formed group which is part of the larger Neuron organization. If you want to shape a team's technical culture from its earliest days while working on problems that matter to the future of AI infrastructure, we'd love to hear from you.\nBasic qualifications\n- 3+ years of non-internship professional software development experience\n- Knowledge of Python and/or C++ programming\n- Knowledge of computer architecture, operating systems, and parallel computing\n- Experience with PyTorch, TensorFlow, and/or JAX\nPreferred qualifications\n- Master's degree in Computer Science, Engineering, Mathematics, or a related field\n- Experience optimizing performance for LLM, Vision, or other deep-learning models\n- Experience with kernel writing or parallel programming (CUDA, Triton, CUTLASS, Pallas, Mojo, SIMD, MPI)\n- Experience with compiler optimization or hardware-software co-design\nOur inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.","description_format":"text","description_chars":4210,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"master","optional":true},"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Advertising","Cloud Platforms (IaaS & PaaS)","Logistics","Streaming & OTT Platforms"],"lifecycle":[{"event":"open","at":"2026-09-25T20:33:40Z"},{"event":"close","at":"2026-10-04T10:57:42Z"}],"visa":[],"liveness":null,"pay":null,"html_url":"https://alion.io/job/amazon-machine-learning-performance-engineer-annapurna-labs","json_url":"https://alion.io/job/amazon-machine-learning-performance-engineer-annapurna-labs.json","meta":{"generated_at":"2026-10-06T00:45:41Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1065,"day_limit":5000,"remaining_today":3935,"minute_limit":60,"resets_at":"2026-10-07T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":19},"rest":"https://alion.io/mcp/rest/get_company?id=19"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Famazon-machine-learning-performance-engineer-annapurna-labs"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Famazon-machine-learning-performance-engineer-annapurna-labs"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Famazon-machine-learning-performance-engineer-annapurna-labs"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/amazon-machine-learning-performance-engineer-annapurna-labs\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Famazon-machine-learning-performance-engineer-annapurna-labs"}]}