1,283,717open jobs
74,382companies
213,626added this week
Browse all
Salary
≈ $151k – $300k per year (Estimated)
Location
In office (Palo Alto)
Seniority
Middle · 3+ years exp

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Aug 20, 2026. Chase UK scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Chase UK is the digital retail bank launched in the United Kingdom by JPMorgan Chase, offering app-based current and saver accounts, cashback on card spending and investing through JP Morgan Personal Investing, with its registered office at Canary Wharf in London. It launched in 2021 as the first international expansion of the Chase consumer banking brand, has no physical branches and provides round-the-clock support through its app and by phone. Its London-based teams include software engineers, product managers, data scientists, designers and customer service staff, and the bank bought the digital wealth manager Nutmeg in 2021.

We have an exciting and rewarding opportunity for you to take your software engineering career to the next level.

As a Software Engineer III at JPMorganChase within the AI/ML data platform team you serve as a seasoned member of an agile team to build and operate scalable, reliable ML training systems and pipelines on AWS and other cloud platforms. You will productionize training workloads (often GPU-based), improve performance and cost efficiency, and enable repeatable, well-governed training across environments

Job Responsibilities

  • Design, build, and maintain end-to-end ML training platform.

  • Run and optimize GPU training workloads (single-node and distributed), improving throughput, utilization and reproducibility.

  • Build and operate training infrastructure on Kubernetes (e.g., EKS and other manage Kubernetes platforms), including resource management and workload troubleshooting.

  • Enable Gen AI/LLM training and fine-tuning workflows (e.g., supervised fine-tuning), including evaluation harnesses, artifact/version governance, and scalable GPU execution patterns aligned to enterprise controls.

  • Implement observability for training systems: metrics, logs, dashboards, alerting, and operational runbooks.

  • Partner with data engineering and platform teams to define interfaces, standards, and guardrails (security, access, cost controls)

  • Improve developer experience for training: standardized containers, CI/CD, templates, documentation, and self-service workflow

  • Leverages enterprise-authorized AI coding assist tools within the work environment to improve code quality, delivery speed, and productivity across complex deliverables (e.g., code generation/refactoring, unit test creation, documentation), while validating outputs through peer review, automated testing, and secure coding standards; contributes learnings and reusable patterns to improve broader team effectiveness.

  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.

Required qualifications, capabilities, and skills

  • Formal training or certification on software engineering concepts and 3+ years applied experience

  • Demonstrated experience running ML training in cloud environments and debugging issues across infrastructure & code.

  • Strong Python skills with solid engineering practices (testing, code reviews, modular design, dependency management).

  • Experience building automation/CI for ML codebases (build, test, release, deployment/promotion workflows).

  • Hands on experience with deep learning training workflows and at least one major framework (eg., PyTorch or TensorFlow).

  • Understanding of training performance and stability: data loading bottlenecks, mixed precision, checkpointing, reproducibility, and evaluation methodology.

  • Experience with distributed training and related concepts (e.g., DDP/FSDP/DeepSpeed concepts, collective communication basics, scaling and bottleneck analysis).

  • Ability to profile and optimize training systems (CPU/GPU utilization, memory, I/O throughput, networking, scheduling).

  • Experience with Kubernetes fundamentals for running compute-intensive workloads and AWS (eg., EKS/ECR, S3, IAM, VPC/networking, Cloudwatch, EC2)

  • Hands-on experience using enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, test creation, troubleshooting, or documentation) with demonstrated ability to critically evaluate, validate, and refine AI-generated outputs for correctness, performance, and security.

  • Understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; ability to guide peers on safe and effective usage within team practices.

Preferred qualifications, capabilities, and skills

  • Experience running training workloads across multiple cloud platforms and managing portability, performance, and governance across environments.

  • Familiarity with cloud-native networking/storage patterns for high-throughput training and artifact management.

  • Experience optimizing training input pipelines (sharding, prefetching, caching, format choices such as Parquet/WebDataset) and working with large datasets.

  • Familiarity with distributed compute frameworks (Spark, Ray, Dask) for feature/dataset generation.

  • Familiarity with workflow orchestration tools (Airflow-like systems, Argo Workflows-like patterns) and model registry concepts.

  • Experience optimizing training cost/performance (right-sizing, scheduling policies, interruptible capacity strategies where applicable budge guardrails, quota planning).

  • Strong observability practice for training systems: metrics/logs/traces, GPU telemetry, dashboards, and alert tuning.

FEDERAL DEPOSIT INSURANCE ACT: This position is subject to Section 19 of the Federal Deposit Insurance Act. As such, an employment offer for this position is contingent on JPMorganChase’s review of criminal conviction history, including pretrial diversions or program entries.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,283,717 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Palo Alto
$150k – $165k per year • In office • Full-Time • 5+ years exp • Atlanta
AI/ML
Function Calling
LLM
RAG
Agentic Workflows
Apply
AI/ML Engineer 2 hours ago
$118k – $148k per year • Hybrid • 8+ years exp
Python
JavaScript
Node JS
Databases
Google BigQuery
BigQuery
AI/ML
Vertex AI
Fine-tuning
TensorFlow
Gemini
DevOps
GCP
Git
IAM
Management
Jira
Agile
Apply
$151k – $251k per year • In office • Secret • 6+ years exp • Bachelor's Degree • McLean
Python
JavaScript
TypeScript
Node JS
Python
FastAPI
AI/ML
Hadoop
NumPy
OpenAI
Anthropic
Machine Learning
Frontend
Vue.js
Svelte
Next.js
React.js
Remix
React Router
DevOps
Vercel
Azure
AWS
Analytics
ETL/ELT
A/B Testing
Apply
$243k – $313k per year • Equity • Hybrid • 7+ years exp • Bachelor's Degree • San Mateo
JavaScript
AI/ML
Machine Learning
Frontend
Bootstrap
DevOps
Rest API
gRPC
CI/CD
Apply
$175k – $215k per year • In office • Full-Time • 3+ years exp • Master's Degree • New York
Python
C++
AI/ML
Computer Vision
Machine Learning
DevOps
Git
Unix
Apply
$253k – $380k per year • Equity • Remote (United States) • Full-Time • 10+ years exp • Bachelor's Degree • Palo Alto
Python
AI/ML
Dagster
Computer Vision
PyTorch
Ray
Machine Learning
DevOps
HPC
Apply
$157k – $235k per year • Equity • In office • Full-Time • 2+ years exp • Bachelor's Degree • Santa Monica • Bellevue • Seattle • Palo Alto
Python
Java
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Spark
Scikit-learn
AI Agents
Flink
TensorFlow
PyTorch
Ray
Machine Learning
Apply
≈ $66k – $132k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Palo Alto
Apply
≈ $66k – $132k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Palo Alto
Apply
≈ $53k – $112k per year (Estimated) • In office • 2+ years exp • Palo Alto
Apply
See all jobs
This is one of many
1,283,717 more open roles from verified company boards, updated every day.