1,264,521open jobs
72,980companies
210,977added this week
Browse all
Salary
≈ $115k – $240k per year (Estimated)
Location
In office (Glasgow)
Seniority
Staff

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Oct 2, 2026. JPMorganChase scores A on the Alion truth index.

Overview
Company
Impact
Profile match
JPMorganChase is the largest bank in the United States by assets and one of the most systemically important financial institutions in the world, with a lineage running back through more than a thousand predecessor firms to the 1799 founding of the Bank of the Manhattan Company. It combines a dominant investment bank and markets business with Chase, the largest retail banking franchise in America, plus commercial banking and asset and wealth management. Headquartered in New York, the group is unusual among banks for the scale of its technology spending, running one of the largest engineering organisations of any financial institution and deploying its own internal AI platform across the firm.

Build the platforms that make advanced AI practical at scale. In this role, you’ll shape standards, tooling, and reliable inference foundations that help engineering teams move faster with confidence. You’ll work hands-on with modern large language model serving stacks and performance tuning, while partnering closely with platform and product stakeholders. If you enjoy solving deep systems problems and enabling others through great developer experience, you’ll find meaningful impact and growth here.

Job Summary:

As a Senior Lead Software Engineer in Corporate Technology - AI, Machine Learning and Data Platform, you will lead the design and delivery of secure, stable, and scalable platform capabilities that simplify adoption and day-to-day use. You will set technical direction for tooling and runtime foundations, with a focus on production-grade large language model inference and Kubernetes-based deployment patterns. You will partner across engineering teams to improve reliability, developer experience, and operational outcomes through automation and standards. You will mentor engineers and reinforce inclusive, high-accountability ways of working.

Job Responsibilities

  • Lead the design and delivery of platform standards and tooling such as command line interfaces, software development kits, libraries, templates, and automated checks to simplify adoption and day-to-day use
  • Engineer and operate production large language model inference services using modern serving engines such as vLLM, TensorRT-LLM, SGLang, LLM-D, or equivalent systems
  • Drive Kubernetes-based deployment patterns, scaling strategies, networking approaches, and troubleshooting practices to support reliable platform operations
  • Optimize inference performance by applying a strong understanding of GPU memory behavior, including key-value cache sizing, memory bandwidth trade-offs, and compute bottlenecks
  • Evaluate and apply inference-time quantization approaches, balancing latency, throughput, cost, and output quality for real-world workloads
  • Implement secure, high-quality production code and automation that strengthens resiliency, observability, and operational readiness
  • Establish and maintain architecture and design artifacts, ensuring constraints and non-functional requirements are enforced through implementation and automation
  • Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test strategy acceleration, incident/root-cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation

Required Qualifications, Capabilities, and Skills

  • Hands-on experience building standards and tooling such as command line interfaces, software development kits, libraries, templates, and automated checks to improve platform adoption
  • Deep, hands-on experience with large language model inference systems such as vLLM, TensorRT-LLM, SGLang, LLM-D, or equivalent production serving engines
  • Hands-on experience building and operating production services on public cloud platforms such as AWS
  • Ability to design, deploy, and troubleshoot cloud infrastructure components used by platform services (e.g., compute, storage, networking, identity and access) in AWS
  • Demonstrated Kubernetes expertise across deployments, scaling, networking, and troubleshooting
  • Working knowledge of GPU memory architecture, including key-value cache sizing and behavior, and performance trade-offs between memory bandwidth and compute bottlenecks
  • Understanding of inference-time quantization trade-offs and how they impact latency, throughput, and real-world serving behavior
  • Ability to produce architecture and design artifacts and translate them into secure, scalable implementations
  • Strong understanding of software development lifecycle practices, including continuous integration and delivery, resiliency, and security expectations
  • Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, test creation, troubleshooting, or documentation) with demonstrated ability to critically evaluate, validate, and refine AI-generated outputs for correctness, performance, and security
  • Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; ability to guide peers on safe and effective usage within team practices

Preferred Qualifications, Capabilities, and Skills

  • Experience building or operating shared platform capabilities used by multiple engineering teams
  • Familiarity with model lifecycle tooling and patterns for safe deployment, rollback, and monitoring of inference services
  • Experience designing SLOs, error budgets, and operational controls for high-throughput platform services
  • Familiarity with service mesh or advanced Kubernetes traffic management patterns for inference workloads
  • Experience improving developer experience through self-service workflows and clear engineering standards
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,264,521 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Glasgow
Sr Engineer, AI 1 day ago
$153k – $276k per year • Equity • In office • Full-Time • 4+ years exp • Bachelor's Degree • Bellevue
Python
AI/ML
LangChain
Fine-tuning
Embeddings
Prompt Engineering
AI Agents
LLM
RAG
Apply
Inference Engineer 2 hours ago
≈ $165k – $364k per year (Estimated) • Hybrid • Full-Time • San Francisco
Python
Rust
C++
AI/ML
vLLM
CUDA Toolkit
Quantization
SGLang
TensorRT
TensorRT-LLM
CUDA
NCCL
Speculative Decoding
KV Cache
Apply
$146k – $189k per year • Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Raleigh
AI/ML
Prompt Engineering
LLM
Together AI
Apply
Senior ML Engineer 5 months ago
≈ $69k – $173k per year (Estimated) • Hybrid • Full-Time • Stockholm
Python
SQL
Python
Flask
AI/ML
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
Senior ML Engineer 5 months ago
≈ $69k – $173k per year (Estimated) • Hybrid • Full-Time • Gothenburg
Python
SQL
Python
Flask
AI/ML
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
In office • Contractor • 1+ year exp
Python
PowerShell
DevOps
CI/CD
Docker
Apply
In office • Contractor
SQL
C#
C#
.NET
DevOps
Azure DevOps
Azure
CI/CD
GitHub
Apply
Scrum Master 2 days ago
In office • Contractor
AI/ML
AI Agents
Hallucination
DevOps
CI/CD
Management
Microsoft Teams
Agile
Scrum
Apply
In office • Contractor
DevOps
CI/CD
Apply
≈ $31k – $64k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Bengaluru
Python
Java
SQL
AI/ML
LangChain
Claude
MLFlow
Fine-tuning
Scikit-learn
Prompt Engineering
AI Agents
Llama
PyTorch
RAG
LLMOps
Machine Learning
DevOps
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $123k – $313k per year (Estimated) • In office • London
Databases
Databricks
AI/ML
LangGraph
LangChain
Model Context Protocol
MLFlow
AI Agents
LLM
Google ADK
A2A
Multi-Agent Systems
DevOps
CI/CD
AWS
Kubernetes
Amazon EKS
Management
Agile
Apply
≈ $181k – $408k per year (Estimated) • In office • 4+ years exp • Master's Degree • New York
Python
AI/ML
AI Agents
LLM
Ray
Agentic Workflows
DevOps
Terraform
CloudFormation
SLURM
CI/CD
AWS
Kubernetes
Amazon EC2
Linux
Apply
≈ $122k – $255k per year (Estimated) • In office • Master's Degree • London
Python
JavaScript
TypeScript
SQL
Databases
PostgreSQL
Redis
Snowflake
Databricks
OpenSearch
AI/ML
Model Context Protocol
MLFlow
Fine-tuning
Embeddings
AI Agents
LLM
RAG
Amazon SageMaker
A2A
GraphRAG
Knowledge Graph
LLM Guardrails
Multi-Agent Systems
Machine Learning
Frontend
Svelte
Next.js
React.js
DevOps
Splunk
Azure
CI/CD
AWS
Kubernetes
Bitbucket
Amazon EKS
Management
Confluence
Apply
≈ $161k – $364k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Jersey City
Python
AI/ML
Reinforcement Learning
Scikit-learn
Prompt Engineering
NLP
TensorFlow
PyTorch
LLM
RAG
Time Series Forecasting
OpenAI
GPT-4
Machine Learning
DevOps
GCP
Azure
AWS
Docker
Kubernetes
Apply
≈ $111k – $231k per year (Estimated) • In office • Master's Degree • Glasgow
Python
Databases
Databricks
AI/ML
Copilot
Claude Code
Fine-tuning
Prompt Engineering
AI Agents
LLM
RAG
LLM Guardrails
Machine Learning
DevOps
CI/CD
Kubernetes
Management
Agile
Apply
≈ $49k – $114k per year (Estimated) • In office • Full-Time • Northampton • London • Glasgow
Apply
≈ $25k – $49k per year (Estimated) • Hybrid • Full-Time • London • Manchester • Leeds • Birmingham • Glasgow
Apply
≈ $60k – $107k per year (Estimated) • Hybrid • Full-Time • Master's Degree • Warrington • Glasgow • Newcastle upon Tyne • Middlesbrough
Management
Agile
Apply
In office • Glasgow
Java
Databases
Snowflake
Apache Kafka
Frontend
GraphQL
DevOps
Rest API
CI/CD
AWS
Management
Agile
Apply
≈ $55k – $123k per year (Estimated) • In office • Glasgow
Java
Java
Spring Boot
Databases
Apache Kafka
AI/ML
Machine Learning
DevOps
Splunk
CI/CD
Grafana
Linux
Unix
Management
Agile
Apply
See all jobs
This is one of many
1,264,521 more open roles from verified company boards, updated every day.