Salary
≈ $92k – $187k per year (Estimated)
Location
In office (Austin)
Seniority
Middle · 4+ years exp
First seen by Alion on Sep 25, 2026. Google scores B on the Alion truth index.
Overview
Company
Impact
Profile match
Google is an American technology company founded in 1998 by Larry Page and Sergey Brin and now the principal subsidiary of Alphabet, headquartered in Mountain View, California. It operates the world's dominant search engine and the advertising system built around it, along with YouTube, Android, Chrome, Gmail, Maps, Workspace and Google Cloud, reaching billions of users across nearly every internet-connected market. The company designs its own silicon in the Tensor Processing Unit line, develops the Gemini foundation models through Google DeepMind, and derives most of its revenue from advertising while cloud has become its fastest growing segment.
About the job
Our AI Infrastructure Engineering Support team is dedicated to ensuring our customers get the most out of their Google Cloud hardware investment. As a Platform Application Engineer (Hardware Engineer), you will be focused on solving customer observations by, driving deep hardware analysis, debug, and issue resolution through to root cause. You will dive deep into complex technical challenges, troubleshoot critical issues across the platform, and provide resolutions in both short-term and long-term platform solutions. In this role, you will represent the customer solution, collaborating tightly with engineering and product teams to drive continuous improvement in our products and services.Google Cloud accelerates every organization’s ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google’s cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.
Individual pay is determined by factors including job-related skills, experience, and relevant education or training.US: $159000 - $230000 (USD) + 15% bonus target + equity + benefits
Learn more about benefits at Google.
Responsibilities
- Manage customer’s problems through effective diagnosis, resolution, or implementation of new investigation tools to increase productivity on AI/ML infrastructure.
- Work closely with multiple Product, Quality, and Engineering teams to improve the product, and interact with our Site Reliability Engineering (SRE) teams to understand behaviours.
- Debug platform hardware and silicon-related issues to drive root-cause resolution and develop permanent improvements.
- Drive understanding of AI/ML workloads and underlying hardware architectures by troubleshooting, reproducing, determining the cause for customer reported issues, and building tools for faster diagnosis.
- Act as a consultant and subject matter expert for internal stakeholders in Engineering and Quality organizations to resolve complex deployment and operational obstacles in AI infrastructure environments.
Qualifications
Minimum qualifications:
- Bachelor's degree in Computer Science, Management Information Systems, or other technical field, or equivalent practical experience.
- 4 years of debug or validation experience with CPU, dGPU, or TPU.
- 4 years of experience with technical infrastructure (deployment, maintenance, and troubleshooting), and quality and reliability of technical infrastructure.
- 3 years of experience with hardware debug (silicon debug, platform debug, IO interface, or memory analysis).
- Experience debugging technical issues across the stack (hardware faults, low-level software, networking, virtualization, kernel drivers, firmware, or performance).
- Experience with Linux/Unix systems and debugging issues across the hardware/software boundary on enterprise-grade server infrastructure.
Preferred qualifications:
- Experience working with distributed systems, and familiarity with common solutions, design patterns, or best practices.
- Experience working directly with AI/ML computing hardware, including GPUs or other accelerators.
- Familiarity with containerization and orchestration technologies like Kubernetes or Slurm in an on-prem or cloud environment.
- Experience with systems automation, and with systems design and debug.
- Experience with ML frameworks (e.g., TensorFlow, PyTorch), and understanding of the AI/ML training and inference lifecycle.
- Advanced understanding of memory and high-speed input/output (IO) technologies.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
911,780 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
DevOps
Similar stack
Same company
Austin
$76k – $114k per year • In office • Secret • Full-Time • 2+ years exp • Bachelor's Degree • Manhattan Beach
Python
C++
DevOps
Linux
Unix
Analytics
Microsoft Excel
Management
Agile
Apply
Senior Kubernetes Engineer
6 months ago
≈ $112k – $218k per year (Estimated) • In office • Full-Time • Dallas
Python
AI/ML
LLM
DevOps
Terraform
Helm
OpenTelemetry
FluxCD
Kustomize
Prometheus
SLURM
CI/CD
GitOps
ArgoCD
Kubernetes
Grafana
HPC
Apply
HPC Systems Engineer
13 days ago
≈ $113k – $221k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Dallas
AI/ML
InfiniBand
DevOps
Terraform
Ansible
SLURM
CI/CD
Kubernetes
HPC
Apply
HPC Platform Engineer
12 days ago
≈ $110k – $215k per year (Estimated) • In office • Full-Time • 5+ years exp • Dallas
Python
AI/ML
InfiniBand
DevOps
Puppet
Ansible
Chef
Prometheus
SLURM
Docker
Kubernetes
Nagios
OpenStack
Telegraf
HPC
Cybersecurity
GDPR
HIPAA
Analytics
Apache NiFi
Apply
Azure Cloud Infrastructure Engineer
6 hours ago
$120k – $235k per year • In office • 8+ years exp • Bachelor's Degree • Washington
Python
JavaScript
C#
C++
AI/ML
Model Context Protocol
AI Agents
DevOps
Terraform
GCP
Azure DevOps
GitHub Actions
Prometheus
Pulumi
Azure
CI/CD
AWS
Kubernetes
Grafana
Platform Engineering
Bicep
CDKTF
Azure AKS
Management
Microsoft Office
Apply
Специалист поддержки рабочих мест
1 day ago
≈ $8.5k – $16k per year (Estimated) • In office • Bachelor's Degree • Omsk
DevOps
Linux
Windows
Wi-Fi
Management
Microsoft Office
Apply
Senior Golang Разработчик (GigaChat)
1 day ago
≈ $30k – $73k per year (Estimated) • Hybrid • 1+ year exp • Moscow
SQL
Databases
PostgreSQL
AI/ML
Function Calling
LLM
OpenAI
Tool Use
DevOps
Rest API
gRPC
WebSockets
CI/CD
Git
Docker
Linux
Apply
≈ $22k – $43k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Moscow
SQL
Databases
PostgreSQL
DevOps
Ansible
Zabbix
Jenkins
Grafana
Bitbucket
Unix
Management
Confluence
Jira
Apply
Senior Python разработчик (GigaChat)
1 day ago
≈ $30k – $74k per year (Estimated) • In office • Moscow
Python
DevOps
Kubernetes
Linux
Apply
Сетевой инженер ЦОД
1 day ago
≈ $14k – $36k per year (Estimated) • In office • Bachelor's Degree • Yekaterinburg
DevOps
Zabbix
Unix
VPN
VLAN
BGP
OSPF
MPLS
Management
Р7-Офис
Microsoft Office
Apply
Senior Staff Software Engineer, High Performance Networking, Platforms Infrastructure Engineering
3 days ago
≈ $159k – $316k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Sunnyvale
C++
AI/ML
Vertex AI
LLM
TPU
Edge AI
Machine Learning
DevOps
GCP
HPC
Linux
Apply
≈ $132k – $257k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Sunnyvale
Python
C++
AI/ML
Machine Learning
DevOps
GCP
Linux
Apply
≈ $127k – $247k per year (Estimated) • In office • 6+ years exp • Bachelor's Degree • Kirkland
AI/ML
TensorFlow
PyTorch
TPU
DevOps
GCP
SLURM
Kubernetes
Linux
Unix
Apply
≈ $102k – $208k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Sunnyvale
DevOps
GCP
Apply
Wireless System Engineer, Bluetooth
4 days ago
≈ $84k – $180k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Mountain View
Python
Java
C++
DevOps
Wi-Fi
Apply
Customer Service Warehouse Associate
6 hours ago
$36k – $40k per year • In office • Full-Time • Austin
Apply
Event Marketing Communications Leader
1 day ago
$196k – $286k per year • Equity • Remote (United States) • Full-Time • 15+ years exp • San Jose • Austin • Atlanta • Boston • Bellevue
Apply
$120k – $158k per year • Equity • Hybrid • Full-Time • 2+ years exp • Austin
Apply
≈ $107k – $226k per year (Estimated) • Remote (United States) • Full-Time • 5+ years exp • PhD • Austin
Apply
$72k – $95k per year • In office • Full-Time • 4+ years exp • High School Diploma • Austin
Apply
This is one of many
911,780 more open roles from verified company boards, updated every day.

