435,278open jobs
15,177companies
61,432added this week
Browse all
Salary
$42k – $106k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Senior
Overview
Company
Impact
Profile match

Glance AI is an AI commerce platform shaping the next wave of e-commerce with inspiration-led shopping, less about searching for what you want and more about discovering who you could be. Operating in 140 countries, Glance AI transforms every screen into a stage for instant, personal, and joyful discovery, where inspiration becomes something you can explore, feel, and shop in the moment.

Its proprietary models, seamlessly integrated with Google’s most advanced AI platforms, Gemini and Imagen on Vertex AI, deliver hyper-realistic, deeply personal shopping experiences across categories such as fashion, beauty, travel, accessories, home décor, pets, and more. Designed to seamlessly integrate into everyday consumer technology, Glance AI reimagines the future of e-commerce with inspiration-led discovery and shopping.

With an open architecture built for effortless adoption across hardware and software ecosystems, Glance AI is creating a platform that can become a staple in everyday consumer technology. It partners with the world’s leading smartphone makers, connected TV manufacturers, telecom providers, and global brands - meeting people where they are: on mobile, smart TVs, and brand websites.

Through Glance AI’s rich first-party data and unparalleled consumer access, it harnesses InMobi’s global scale, insights, and targeting capabilities to create high-impact, performance-driven shopping journeys for brands worldwide. Part of the InMobi Group, a global technology and advertising leader reaching over 2 billion devices and serving more than 30,000 enterprise brands worldwide, Glance AI is backed by Google, Jio Platforms, and Mithril Capital.

About the Role

As a GPU Systems Engineer, you’ll lead design and optimization efforts across our GPU inference stack.

You will architect the libraries and runtime systems that enable Stable Diffusion, multimodal transformers, and emerging video generation models to run efficiently at scale.

You’ll guide cross-functional teams, influence hardware selection, and set the technical vision for GPU optimization practices across the company.

Key Responsibilities

  • Architect high-performance inference runtimes, kernel dispatchers, and memory planners for large diffusion and transformer workloads.
  • Lead investigations into cross-GPU performance bottlenecks, communication overheads, and scheduling inefficiencies.
  • Drive multi-GPU parallelism strategies - model, pipeline, and tensor parallelization.
  • Establish company-wide GPU optimization standards, tooling, and SLIs.
  • Collaborate with research to design scalable implementations of novel architectures.
  • Mentor engineers in profiling, tuning, and low-level optimization.
  • Partner with hardware vendors and infra teams to maximize cluster utilization.

Required Qualifications

  • 5+ years in high-performance computing, GPU runtime systems, or ML infrastructure.
  • Proven expertise in CUDA / Triton / C++, with deep understanding of GPU scheduling, occupancy, register usage, and tensor cores.
  • Experience building and maintaining distributed inference or training systems.
  • Ability to design abstractions balancing flexibility and performance.
  • Strong knowledge of NCCL, NVLink, PCIe, and interconnects.
  • Familiar with profiling automation and performance dashboards.
  • Excellent technical leadership and mentoring capabilities.

Preferred Qualifications

  • Background in compiler-aided optimization (TVM, XLA, MLIR, Triton).
  • Experience tuning Stable Diffusion or transformer inference pipelines.
  • Exposure to heterogeneous compute backends (AMD ROCm, TPU, ASICs).
  • Experience working with hardware-software co-design initiatives.
  • Open-source or research contributions in GPU optimization

"Glance collects and processes personal data such as your name, contact details, resume and other information that may contain personal data for the purpose of processing your application. Glance utilizes Greenhouse, a third-party platform. Please review Greenhouse's Privacy Policy to understand how the data collected from you is processed and managed. By clicking on 'Submit Application', you acknowledge and agree to the above privacy terms. Should you have any privacy concerns, you may contact us through the details mentioned in your application confirmation email."

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
435,278 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$93k – $233k per year (Estimated) • In office • Full-Time • 5+ years exp • Sydney
Python
Go
C++
AI/ML
vLLM
CUDA Toolkit
Triton Inference Server
Embeddings
Quantization
Multimodal AI
Function Calling
AI Agents
SGLang
TensorRT
TensorRT-LLM
TGI
LLM
RAG
Reranking
NVIDIA NIM
CUDA
Triton
Hugging Face
NCCL
NVLink
cuDNN
Speculative Decoding
KV Cache
Agentic Workflows
Tool Use
DevOps
CI/CD
GitOps
Kubernetes
Platform Engineering
Apply
$19k – $50k per year (Estimated) • In office • Full-Time • 3+ years exp • Master's Degree • Shanghai
Python
C++
MATLAB
MATLAB
Simulink
Apply
$120k – $263k per year (Estimated) • In office • Full-Time • Richmond • Sydney
Python
SQL
Databases
Apache Kafka
Google BigQuery
BigQuery
AI/ML
Embeddings
Multimodal AI
LLM
DevOps
CI/CD
AWS
Apply
$164k – $335k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • Melbourne
JavaScript
Rust
C++
Frontend
React.js
Mobile
React Native
Expo
DevOps
CI/CD
Design
Canva
Apply
$163k – $332k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • Sydney
JavaScript
Rust
C++
Frontend
React.js
Mobile
React Native
Expo
DevOps
CI/CD
Design
Canva
Apply
$35k – $66k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
AI/ML
Vertex AI
Gemini
Design
Adobe Photoshop
Blender
Adobe Premiere Pro
Adobe Illustrator
Figma
Adobe After Effects
Apply
Consumer Insights 7 days ago
In office • Bachelor's Degree • Bengaluru
Python
AI/ML
Vertex AI
AI Agents
Gemini
Analytics
Tableau
Apply
$150k – $220k per year • Equity • In office • 6+ years exp • PhD • New York
AI/ML
Vertex AI
AI Agents
Gemini
Apply
Sr UX Designer 8 days ago
$28k – $63k per year (Estimated) • In office • Bachelor's Degree • Bengaluru
AI/ML
Vertex AI
Gemini
Apply
UX Researcher 11 days ago
$19k – $54k per year (Estimated) • In office • 3+ years exp • Bengaluru
AI/ML
Vertex AI
Gemini
Analytics
A/B Testing
Apply
$23k – $50k per year (Estimated) • In office • Full-Time • PhD • Bengaluru
Apply
Remote/Hybrid • Full-Time • PhD • Bengaluru
AI/ML
Reinforcement Learning
Time Series Forecasting
Interpretability
Apply
$17k – $38k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Bengaluru • Gurgaon
Management
Outlook
Apply
$91k – $151k per year • In office • Full-Time • 1+ year exp • Bachelor's Degree • Atlanta • Bengaluru • Hyderabad
Analytics
Tableau
Power BI
Microsoft Excel
Apply
Buyer I 8 hours ago
$13k – $31k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Analytics
Tableau
Apply
See all jobs
This is one of many
435,278 more open roles from verified company boards, updated every day.