1,190,436open jobs
66,700companies
211,487added this week
Browse all
Location
In office (Pune)
Seniority
Middle · 3+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 4, 2026. First seen by Alion on Sep 11, 2026. GlobalFoundries scores B on the Alion truth index.

Overview
Company
Impact
Profile match
GlobalFoundries is a semiconductor foundry headquartered in Malta, New York, near Albany, that manufactures chips for other companies on FD-SOI, FinFET, RF, silicon germanium, silicon photonics, power and GaN process technologies, and also offers MIPS processor designs. Founded in 2009 and listed on Nasdaq, it runs fabs in New York, Vermont, Dresden and Singapore with engineering and design centres in Austin, Santa Clara, Bengaluru and Penang. It hires process, equipment, yield and tapeout engineers, maintenance technicians and apprentices, field application engineers and interns in engineering and finance.

Sr Staff Engineer - AI/ML Compiler & Runtime Software Engineer

AI SDK Team

Location: Pune / Bangalore - India

Join the RISC-V Revolution!

About GlobalFoundries

GlobalFoundries is a leading full-service semiconductor foundry providing a unique combination of design, development, and fabrication services to some of the world’s most inspired technology companies. With a global manufacturing footprint spanning three continents, GlobalFoundries makes possible the technologies and systems that transform industries and give customers the power to shape their markets. For more information, visit www.gf.com

Introduction

We are seeking a highly skilled Sr Staff Engineer in AI/ML compiler and runtime software to join our Platform Software and AI SDK team. The team is building the foundational software stack to enable Physical AI workloads on next-generation RISC-V IP and SoC platforms.

This role sits at the intersection of AI compiler technology, edge AI deployment, runtime systems, and hardware acceleration. You will work on IREE-based compiler and runtime flows, LLVM/MLIR infrastructure, custom MLIR dialects and passes, code generation, quantization, and NPU acceleration to enable efficient execution of AI models on edge devices and silicon platforms.

This is a unique opportunity to contribute across the full silicon-to-software lifecycle, combining compiler engineering, AI runtime development, and hardware-software co-design to deliver high-performance, low-latency, and power-efficient AI execution for real-time edge and Physical AI use cases.

What You’ll Do

  • Architect, design, and develop AI/ML compiler and runtime software for RISC-V based IP, NPU, and SoC platforms.

  • Develop and enhance IREE-based compiler flows, including MLIR lowering, code generation, runtime integration, and deployment paths for edge AI workloads. Create and maintain custom MLIR dialects, compiler passes, lowering pipelines, and transformation flows to map AI workloads efficiently to custom NPU and accelerator hardware.

  • Work across AI framework import paths including PyTorch, ONNX, and TFLite, and enable lowering through torch-mlir, TOSA, Linalg, and related MLIR dialects.

  • Optimize neural network workloads for edge deployment, including operator fusion, tiling, memory planning, quantization, layout transformation, and accelerator-aware scheduling.

  • Enable efficient execution of AI models across CPU, vector, matrix, and NPU acceleration paths, balancing latency, throughput, memory footprint, and power efficiency.

  • Collaborate closely with architecture, hardware, firmware, FPGA, validation, and product teams to bring up AI workloads on simulators, FPGA platforms, emulation environments, and silicon.

  • Analyze model performance, identify compiler/runtime bottlenecks, and drive optimizations across graph-level, operator-level, and kernel-level execution paths.

  • Define software architecture and technical direction for AI SDK components, including compiler pipelines, runtime interfaces, model deployment flows, and accelerator integration.

  • Build test infrastructure, validation flows, benchmark suites, and CI pipelines for AI compiler/runtime correctness, performance, and regression tracking.

  • Provide technical leadership to engineers working on AI compiler, runtime, model deployment, and edge AI software development.

  • Work with internal and customer-facing teams to support software enablement, debugging, performance tuning, and deployment of AI workloads on target platforms.

Ideally, you’ll have

  • 3-12 years of hands-on software engineering experience, with strong experience in compiler, runtime, embedded software, or AI/ML systems.

  • Strong hands-on experience with IREE, LLVM, and MLIR compiler infrastructure. Experience developing MLIR dialects, compiler passes, lowering pipelines, pattern rewrites, code generation flows, or backend integration for custom hardware.

  • Good understanding of IREE code generation flow, dispatch formation, executable generation, HAL/runtime concepts, and target-specific lowering.

  • Strong exposure to AI compiler/runtime stacks used for edge AI or accelerator-backed inference.

  • Experience with AI model formats and frameworks such as PyTorch, ONNX, TensorFlow Lite/TFLite, and related conversion or import flows.

  • Working knowledge of torch-mlir, TOSA, Linalg, tensor dialects, bufferization, quantization dialects, and MLIR-based model lowering concepts.

  • Strong understanding of neural network execution and optimization, including quantization, operator fusion, tensor layouts, memory planning, tiling, vectorization, and kernel selection.

  • Experience enabling or optimizing workloads for AI accelerators, NPUs, DSPs, vector processors, matrix engines, or custom SoC IP.

  • Strong C/C++ programming skills, with good Python scripting ability for compiler tooling, testing, automation, and model workflow integration.

  • Experience working in Linux development environments, including cross-compilation, debugging, profiling, build systems, and runtime bring-up. Strong debugging and problem-solving skills across compiler IR, generated code, runtime behavior, and hardware/software interaction.

  • Ability to work with architecture and hardware teams to understand accelerator capabilities and translate them into compiler/runtime enablement.

  • Proven ability to technically lead complex software modules, mentor engineers, and drive execution across cross-functional teams.

You might also have

  • Experience working on RISC-V, ARM, x86, DSP, GPU, or custom accelerator software stacks.

  • Familiarity with RISC-V Vector, matrix acceleration concepts, custom instructions, or accelerator-specific code generation.

  • Experience with edge AI deployment on real devices, development boards, FPGA platforms, emulators, simulators, or early silicon.

  • Familiarity with FPGA prototyping, Linux bring-up, board-level debugging, or pre-silicon software validation.

  • Exposure to LLM and edge inference stacks such as llama.cpp, GGML/GGUF, ONNX Runtime, TensorFlow Lite, TVM, XNNPACK, or similar frameworks, with understanding of quantization, memory footprint optimization, kernel performance, and deployment constraints on resource-limited devices.

  • Experience with AI model benchmarking and optimization for vision, audio, transformers, GenAI, robotics, automotive, industrial, or real-time embedded workloads.

  • Understanding of hardware-software co-design, memory hierarchy, DMA, scratchpad memory, cache behavior, and accelerator data movement. Experience with runtime systems, kernel libraries, microkernels, custom dispatch flows, or accelerator runtime APIs.

  • Familiarity with CI/CD and agile tools such as Jenkins, Git, CMake, Bazel, Jira, or similar engineering infrastructure.

  • Experience working in customer-facing enablement, silicon bring-up, platform software, or SDK delivery environments.

  • Excellent communication and interpersonal skills, with the ability to explain complex compiler and AI runtime topics clearly to software, hardware, and product stakeholders.

GlobalFoundries is an equal opportunity employer, cultivating a diverse and inclusive workforce. We believe having a multicultural workplace enhances productivity, efficiency and innovation whilst our employees feel truly respected, valued and heard.

As an affirmative employer, all qualified applicants are considered for employment regardless of age, ethnicity, marital status, citizenship, race, religion, political affiliation, gender, sexual orientation and medical and/or physical abilities.

All offers of employment with GlobalFoundries are conditioned upon the successful completion of background checks, medical screenings as applicable and subject to the respective local laws and regulations.

Information about our benefits you can find here: https://gf.com/about-us/careers/opportunities-asia

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,190,436 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Pune
≈ $15k – $38k per year (Estimated) • In office • Full-Time • Chennai
JavaScript
TypeScript
Frontend
Redux
React.js
Mobile
State Management
DevOps
Rest API
CI/CD
Git
Management
Agile
Scrum
QA
Jest
Apply
≈ $16k – $42k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Hyderabad
Python
JavaScript
TypeScript
SQL
C#
C++
AI/ML
Copilot
Prompt Engineering
Function Calling
AI Agents
LLM
RAG
Human-in-the-Loop
LLM Guardrails
Agentic Workflows
Frontend
GraphQL
React.js
DevOps
Progressive Delivery
Management
Microsoft Teams
Agile
Apply
Software Engineer 2 4 days ago
In office • 6+ years exp • Master's Degree • Bengaluru
Python
Go
JavaScript
Java
C#
Node JS
Databases
PostgreSQL
Apache Kafka
Azure Cosmos DB
Azure SQL Database
Microsoft Fabric
AI/ML
Spark
DevOps
Logstash
Azure
Analytics
Power BI
Azure Data Factory
Apply
Full Stack Developer 3 hours ago
≈ $15k – $39k per year (Estimated) • Remote (India) • 4+ years exp • Bachelor's Degree
Python
JavaScript
TypeScript
Node JS
Databases
MySQL
Redis
ClickHouse
RabbitMQ
Apache Kafka
Frontend
Vue.js
Angular
React.js
DevOps
GCP
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
Apply
Python Developer 3 hours ago
≈ $21k – $60k per year (Estimated) • Remote (India) • 5+ years exp • Mumbai
Python
Python
Flask
SQLAlchemy
FastAPI
Django
Celery
Databases
PostgreSQL
Redis
RabbitMQ
DevOps
Rest API
CI/CD
Git
AWS
Docker
AWS Fargate
AWS Lambda
Amazon EC2
Amazon ECS
Apply
In office • Contractor • 1+ year exp
Python
PowerShell
DevOps
CI/CD
Docker
Apply
$70k – $84k per year • Hybrid • Full-Time • 3+ years exp
Python
SQL
Analytics
Power BI
Microsoft Excel
Apply
$34k – $50k per year • Hybrid • Internship • Bachelor's Degree • Stamford
Python
JavaScript
TypeScript
SQL
C#
Databases
MySQL
PostgreSQL
AI/ML
Copilot
ChatGPT
OpenAI
Frontend
Angular
React.js
DevOps
Terraform
Azure
Jenkins
Git
AWS
GitHub
GitLab
Apply
≈ $30k – $62k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Bengaluru
Python
Java
SQL
AI/ML
LangChain
Claude
MLFlow
Fine-tuning
Scikit-learn
Prompt Engineering
AI Agents
Llama
PyTorch
RAG
LLMOps
Machine Learning
DevOps
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $132k – $261k per year (Estimated) • In office • Dallas
Python
TypeScript
SQL
AI/ML
Embeddings
AI Agents
DevOps
Azure
AWS
AWS Lambda
AWS Step Functions
API Gateway
Analytics
ETL/ELT
Apply
≈ $27k – $59k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Bengaluru • Pune
Python
DevOps
RTOS
CI/CD
Robotics
EtherCAT
Chips/EDA
OpenOCD
IoT
FreeRTOS
Apply
≈ $17k – $44k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Bengaluru
Python
Java
DevOps
CI/CD
Apply
In office • Full-Time • Bachelor's Degree • Singapore
Python
JavaScript
TypeScript
SQL
Python
Flask
AI/ML
RAG
Frontend
Vue.js
Angular
React.js
DevOps
Rest API
CI/CD
Git
Apply
$145k – $236k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Apply
≈ $51k – $106k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Singapore
Management
Microsoft Office
Apply
PHP Developer 1 day ago
In office • Full-Time • Pune
JavaScript
PHP
SQL
PHP
Laravel
WordPress
Magento
Drupal
Frontend
JQuery
DevOps
Rest API
Windows
SOAP
Apply
In office • Pune
Apply
≈ $12k – $27k per year (Estimated) • Hybrid • Full-Time • 7+ years exp • Pune
Apex
DevOps
Rest API
SOAP
Apply
In office • Pune
Management
Microsoft Office
Apply
Advanced Data Analyst 4 hours ago
In office • Pune
Python
SQL
Analytics
Power BI
Apply
See all jobs
This is one of many
1,190,436 more open roles from verified company boards, updated every day.