1,287,628open jobs
74,529companies
207,347added this week
Browse all
Salary
≈ $140k – $266k per year (Estimated)
Location
In office (Houston, Tlaquepaque)
Seniority
Senior
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Oct 6, 2026. HP Inc. scores B on the Alion truth index.

Overview
Company
Impact
Profile match
HP Inc. is an American technology company created in 2015 when the original Hewlett-Packard split its personal systems and printing businesses away from its enterprise arm. It is one of the largest personal computer vendors in the world, shipping consumer and commercial laptops, desktops and workstations under the HP, Pavilion, Envy, EliteBook, OMEN and Poly brands, alongside a printing division covering home, office, large-format and industrial 3D printing. Headquartered in Palo Alto, California, the company earns a large share of profit from supplies and services attached to its installed base and has pushed into hybrid-work peripherals and subscription hardware.
Senior Machine Learning Platform Engineer - Model Hosting & MLOps

Description -

Role summary

We are hiring a Senior Machine Learning Platform Engineer to build the infrastructure and engineering workflows that take custom models from development into reliable production use. This is a hands-on role for someone who has hosted LLMs on GPU infrastructure, developed the cloud platform around model serving, and built MLOps pipelines that make releases repeatable and observable.

You will work with model developers, application engineers, security, and cloud platform teams to turn a model artifact into a secure, scalable inference service. The work spans serving architecture, infrastructure as code, deployment automation, model lifecycle management, and production operations across AWS and Azure, with potential integration into on-premises environments. You will also help shape agent workflows that use hosted models, so serving choices support the needs of multi-step applications. Success means teams can deploy, update, monitor, and troubleshoot models through clear, reusable platform patterns.

What you will do

  • Design and build hosting for custom ML and AI models across AWS and Azure, with particular focus on GPU-backed LLM inference and real-time endpoints; support batch inference where appropriate.
  • Package models and their dependencies into reproducible serving workloads; choose and implement suitable managed services, containers, or Kubernetes-based patterns based on throughput, latency, security, reliability, and cost.
  • Provision and configure Kubernetes clusters or other suitable hosting infrastructure, including compute, storage, API access, identity and access controls, secrets, networking integration, observability, and environment configuration.
  • Help design secure connections and deployment patterns between cloud and on-premises environments as hosting needs evolve.
  • Develop infrastructure as code and deployment automation so model-hosting environments can be provisioned, reviewed, promoted, and maintained consistently.
  • Build MLOps workflows for model registration, versioning, validation, release, rollback, and retirement. Connect training or model preparation to deployment through automated pipelines and appropriate quality gates.
  • Partner with application teams to design and prototype agent workflows, including model and tool orchestration, state handling, failure recovery, and evaluation.
  • Translate agent workload patterns into hosting decisions about model selection, context length, concurrency, latency, cost, tool access, and end-to-end tracing.
  • Establish production monitoring for service health, latency, throughput, errors, GPU and other resource use, and model behavior. Share responsibility for diagnosing incidents and improving capacity, reliability, and cost.
  • Create reusable deployment templates, reference architectures, documentation, and onboarding paths that help other teams ship models safely.
  • Partner with model and application teams on practical tradeoffs such as online versus batch inference, managed versus self-hosted serving, scaling, evaluation, data handling, and operational ownership.

Required experience

  • Hands-on experience hosting LLM inference on GPU infrastructure in a production environment. You can explain what you personally built, how models reached production, and how you managed throughput, latency, utilization, reliability, and cost.
  • Experience building the surrounding model-serving platform for custom models, such as inference runtimes, deployment patterns, endpoint access, scaling, and operational tooling.
  • Strong software engineering skills, especially Python, with experience building services, automation, and maintainable production code.
  • Experience developing infrastructure for ML workloads across AWS and Azure, with deep hands-on delivery in at least one and practical ability to work in the other. You have used infrastructure as code such as Terraform or an equivalent tool.
  • Experience provisioning, configuring, and maintaining Kubernetes or another production hosting platform for containerized inference workloads.
  • Experience building or operating ML deployment pipelines with versioned artifacts, automated validation, CI/CD, environment promotion, and rollback.
  • Familiarity with LLM agent patterns, including model invocation, tool calls, and multi-step workflows, and how they affect serving capacity, reliability, and access controls.
  • Working knowledge of production concerns for inference services: scaling, latency, availability, logging and metrics, shared incident response, access control, and cost.
  • Ability to work across model development, application, platform, and security teams; turn ambiguous requirements into a working design; and document the resulting operational approach.

Helpful experience

  • Advanced GPU inference optimization, including capacity planning, batching, autoscaling, memory use, model loading, and performance tuning.
  • Serving generative or other compute-intensive custom models beyond LLMs; selecting and tuning inference servers and runtime configurations.
  • AWS SageMaker or EKS; Azure Machine Learning or AKS; or equivalent managed and self-hosted model platforms.
  • Hybrid or on-premises model hosting, including connectivity, security boundaries, hardware constraints, and operational handoff.
  • Model registries, experiment tracking, data or feature pipelines, scheduled retraining, model evaluation, and drift or quality monitoring.
  • Secure enterprise deployment patterns such as private networking, IAM/RBAC, secrets management, auditability, and handling sensitive data.
  • Building shared ML platform capabilities or self-service workflows used by multiple engineering teams.
  • Hands-on development of agent workflows, including tool integration, evaluation, tracing, or guardrails.

Who will thrive here

You are an engineer who has taken responsibility for what happens after a model is trained: how it is packaged, deployed, secured, scaled, observed, updated, and supported. You are comfortable writing code and infrastructure, investigating production failures, and making clear tradeoffs with partner teams.

Pay & Benefits

The pay range for this role is $147,050 to $230,850 USD annually with additional

opportunities for pay in the form of bonus and/or equity (applies to United

States of America candidates only). Pay varies by work location, job-related

knowledge, skills, and experience.

Benefits:

HP offers a comprehensive benefits package for this position, including:

  • Health insurance
  • Dental insurance
  • Vision insurance
  • Long term/short term disability insurance
  • Employee assistance program
  • Flexible spending account
  • Life insurance
  • Generous time off policies, including;
  • 4-12 weeks fully paid parental leave based on tenure
  • 11 paid holidays
  • Additional flexible paid vacation and sick leave
  • US benefits overview https://hpbenefits.ce.alight.com/

The compensation and benefits information is accurate as of the date of this

posting. The Company reserves the right to modify this information at any time,

with or without notice, subject to applicable law.

Job -

Software

Schedule -

Full time

Shift -

No shift premium (United States of America)

Travel -

No

Relocation -

No

Equal Opportunity Employer (EEO)-

HP, Inc. provides equal employment opportunity to all employees and prospective employees, without regard to race, color, religion, sex, national origin, ancestry, citizenship, sexual orientation, age, disability, or status as a protected veteran, marital status, familial status, physical or mental disability, medical condition, pregnancy, genetic predisposition or carrier status, uniformed service status, political affiliation or any other characteristic protected by applicable national, federal, state, and local law(s).

Please be assured that you will not be subject to any adverse treatment if you choose to disclose the information requested. This information is provided voluntarily. The information obtained will be kept in strict confidence.

For more information, review HP’sEEO Policy or read about your rights as an applicant under the law here: “ Know Your Rights: Workplace Discrimination is Illegal "

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,287,628 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Houston
AI Workflow Engineer 2 hours ago
$107k – $171k per year • Remote (United States) • Contractor • 3+ years exp • Bachelor's Degree • Chicago
AI/ML
Prompt Engineering
LLM
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
Edge AI
Machine Learning
Apply
$209k – $313k per year • Equity • In office • Full-Time • 4+ years exp • Bachelor's Degree • New York
AI/ML
Machine Learning
Apply
$160k – $185k per year • Equity • Remote (United States) • 3+ years exp • PhD
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Computer Vision
TFLite
TensorFlow
Keras
PyTorch
Machine Learning
Apply
$119k – $168k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Irvine
Python
Rust
SQL
C++
Julia
C++
TensorFlow C++
PyTorch C++
Databases
Snowflake
Databricks
AI/ML
OpenCV
Scikit-learn
SciPy
Diffusion Models
Computer Vision
Transfer Learning
TensorFlow
NumPy
Keras
PyTorch
Torchvision
scikit-image
Self-Supervised Learning
Contrastive Learning
Machine Learning
DevOps
Azure
Git
Management
Agile
Apply
$172k – $230k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Seattle • Santa Monica
Python
Java
AI/ML
Fine-tuning
JAX
Prompt Engineering
TensorFlow
PyTorch
LLM
RAG
Hugging Face
Machine Learning
DevOps
GCP
Azure
AWS
Apply
≈ $136k – $229k per year (Estimated) • Remote (United States) • TS/SCI • 5+ years exp
Python
Go
JavaScript
TypeScript
Frontend
Vue.js
Angular
React.js
DevOps
Rest API
Terraform
Ansible
GCP
VMWare
CloudFormation
Pulumi
Azure
CI/CD
AWS
Docker
Kubernetes
KVM
Hyper-V
Management
Agile
Apply
$95k – $125k per year • In office • Full-Time • 1+ year exp • Bachelor's Degree • New York
Python
SQL
PowerShell
Databases
PostgreSQL
Snowflake
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
dbt
LLM
DevOps
Terraform
GCP
Azure
CI/CD
Git
AWS
Platform Engineering
Analytics
Tableau
ETL/ELT
Apply
$139k – $174k per year • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Warren
Python
SQL
Scala
Databases
Databricks
Delta Lake
AI/ML
Copilot
Cursor
Claude
Spark
AI Agents
LLM
RAG
LLM Guardrails
DevOps
GCP
Azure
CI/CD
AWS
Apply
$185k – $284k per year • Remote (United States) • Full-Time • 7+ years exp • Bachelor's Degree • Sunnyvale
Python
Go
Java
TypeScript
C++
Databases
PostgreSQL
DuckDB
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
dbt
Streamlit
Galileo
DevOps
CI/CD
Robotics
ROS
Analytics
Looker
Apply
≈ $49k – $114k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Belo Horizonte
Python
Verilog
VHDL
DevOps
SLURM
CI/CD
Jenkins
Linux
Chips/EDA
Cadence Palladium
Management
Jira
Apply
$131k – $205k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston • Austin
Python
JavaScript
SQL
C++
AI/ML
AI Agents
Frontend
React.js
DevOps
Azure
Docker
Kubernetes
Windows
Management
Agile
Apply
$147k – $231k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Houston • Austin
Python
JavaScript
SQL
C++
AI/ML
AI Agents
Edge AI
Frontend
React.js
DevOps
Azure
Docker
Kubernetes
Windows
Management
Agile
Apply
$131k – $205k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston • Austin
Python
JavaScript
SQL
C++
AI/ML
AI Agents
ONNX
TensorRT
OpenVINO
Edge AI
Frontend
React.js
DevOps
Azure
Docker
Kubernetes
Management
Agile
Apply
$110k – $160k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Austin
Python
C++
C++
CMake
STL
AI/ML
Computer Vision
DevOps
CI/CD
Docker
Linux
Apply
≈ $138k – $264k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston • Tlaquepaque
Python
Java
C++
Python
FastAPI
C++
TensorFlow C++
PyTorch C++
AI/ML
Spark
Model Context Protocol
Scikit-learn
AWS Bedrock
TensorFlow
PyTorch
Amazon SageMaker
Edge AI
Machine Learning
DevOps
Terraform
Azure
CI/CD
AWS
Kubernetes
Management
Agile
Apply
≈ $116k – $231k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Houston
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Apache Kafka
Azure SQL Database
AI/ML
Spark
MLFlow
Scikit-learn
MLlib
TensorFlow
Keras
PyTorch
Ray
Machine Learning
DevOps
Rest API
Azure DevOps
Azure
CI/CD
Git
Kubernetes
Azure AKS
Analytics
ETL/ELT
Azure Data Factory
Apply
$131k – $205k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Houston
AI/ML
Machine Learning
Management
Agile
Scrum
Kanban
Apply
Accounting Assistant 9 hours ago
≈ $45k – $86k per year (Estimated) • In office • Full-Time • 2+ years exp • High School Diploma • Houston
Analytics
Microsoft Excel
Management
Outlook
Apply
≈ $97k – $226k per year (Estimated) • Hybrid • Full-Time • 7+ years exp • Houston
Apply
≈ $38k – $66k per year (Estimated) • In office • Full-Time • 2+ years exp • High School Diploma • Houston
Management
Agile
Apply
See all jobs
This is one of many
1,287,628 more open roles from verified company boards, updated every day.