614,885open jobs
34,185companies
85,883added this week
Browse all
Salary
$142k – $249k per year (Estimated)
Location
Remote (United States)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

Position Summary

We’re hiring a Senior MLOps Engineer with deep machine learning engineering experience to build and operate the production platform powering ML/LLM-driven healthcare workflows. You’ll design reliable, secure, and compliant systems for model development, evaluation, deployment, monitoring, and continuous improvement-working closely with ML, data, security, and product teams.

This role is ideal for someone who has shipped ML systems in production and is excited about LLM orchestration, RAG, evaluations, guardrails, and observability in a regulated environment.

Key responsibilities

MLOps & ML Platform

  • Design and operate ML platforms that support end-to-end workflows: data ingestion, feature engineering, training, evaluation, deployment, and monitoring.
  • Build and maintain CI/CD for ML (testing, packaging, versioning, reproducibility, automated rollbacks, approvals).
  • Implement MLOps best practices: model registry, experiment tracking, lineage, governance, and reproducible training environments.
  • Develop scalable training infrastructure (distributed training, GPU scheduling, cost controls, auto-scaling).
  • Create and maintain feature pipelines / feature stores, ensuring consistency between training and inference (training-serving skew prevention).
  • Establish model monitoring and observability: performance, drift, bias/fairness signals (where relevant), latency, throughput, and data quality.
  • Build and own end-to-end LLM delivery pipelines: prompt/versioning, retrieval, orchestration, evaluation, deployment, monitoring, and iterative improvement.
  • Create robust LLM evaluation harnesses (offline + online): golden datasets, automated regression testing, human-in-the-loop review workflows, and risk scoring.
  • Build cost controls: token/cost budgeting, caching strategies, autoscaling, and performance tuning.

Deployment, reliability, and operations

  • Productionize ML Models on GCP using containers and orchestration (e.g., GKE, Cloud Run), and build CI/CD for ML/LLM systems with automated tests and safe rollouts.
  • Implement observability: tracing, metrics, logs, dashboards, alerting for model/system health (latency, token usage, error rates, retrieval quality, hallucination indicators, drift where relevant).
  • Build cost controls: token/cost budgeting, caching strategies, autoscaling, and performance tuning.

Data, governance, and compliance (Healthcare)

  • Design systems with security and privacy by default: IAM, least privilege, secrets management, audit logs, encryption, data retention, and PHI/PII handling.
  • Implement governance: model/prompt lineage, dataset provenance, evaluation traceability, and approval workflows aligned with healthcare compliance expectations.

Integrate guardrails: content filters, policy checks, prompt injection defenses, structured output validation, and fallback strategies.

Requirements

  • 6+ years in software/platform engineering, including 4+ years operating ML systems in production (or equivalent depth).
  • Strong experience in ML engineering: training pipelines, evaluation, deployment patterns, monitoring, and iteration loops.
  • Strong engineering skills in Python, plus production-grade experience building APIs/services.
  • Demonstrated hands-on experience with LLM systems in production and ML engineering: training pipelines, evaluation, deployment patterns, monitoring, and iteration loops.
  • Strong experience with GCP services and cloud-native patterns.
  • Experience with Vertex AI (pipelines, endpoints, feature store, model registry, evaluation) and/or managed vector search on GCP.
  • Experience with containerization and orchestration (Docker, Kubernetes/GKE and/or Cloud Run).

Benefits

Why Join Us?

Joining C the Signs is not just about building AI; it’s about shaping the future of healthcare. If you are a technical leader with an unshakable belief in the power of AI to save lives and the ability to make it happen at scale, this is your opportunity to create a tangible, global impact.

Benefits:

  • Competitive salary and benefits package.
  • Flexible working arrangements (remote or hybrid options available).
  • The opportunity to work on life-changing AI technology that directly impacts patient outcomes.
  • Join a team that combines cutting-edge innovation with a mission to save lives and improve health equity.
  • Continuous learning opportunities with access to the latest tools and advancements in AI and healthcare.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
614,885 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
Integration Developer 8 hours ago
$105k – $189k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Princeton
Python
Java
SQL
Apex
Apex
MuleSoft
Databases
Apache Kafka
Frontend
GraphQL
DevOps
Rest API
GCP
GitHub Actions
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
SLI/SLO/SLA
IAM
QA
Swagger
Apply
Data Engineer 8 hours ago
$91k – $180k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Princeton
Python
SQL
Databases
PostgreSQL
Snowflake
Databricks
Oracle
Apache Kafka
Microsoft Fabric
AI/ML
Airflow
dbt
Prefect
Great Expectations
DevOps
Rest API
GCP
Azure
CI/CD
AWS
SLI/SLO/SLA
IAM
Amazon Kinesis
Analytics
ETL/ELT
DataStage
Fivetran
Azure Data Factory
Dimensional Modeling
Apply
Data Engineer 5 hours ago
$168k – $245k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Jose
Python
SQL
Databases
Snowflake
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Dagster
Prefect
Great Expectations
DevOps
CI/CD
Git
Vector
FinOps
Analytics
Data Vault
Dimensional Modeling
Apply
$19k – $51k per year (Estimated) • In office • 3+ years exp • Ryazan
Python
Go
SQL
Databases
PostgreSQL
ClickHouse
RabbitMQ
AI/ML
Model Context Protocol
AI Agents
LLM
DevOps
gRPC
Docker
QA
Selenium
Playwright
Apply
$33k – $73k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Bachelor's Degree • Bengaluru
Python
Apply
$85k – $149k per year (Estimated) • Remote • Full-Time • 5+ years exp
Design
Figma
Apply
$94k – $201k per year (Estimated) • Remote • 6+ years exp
Apply
Design Lead 6 months ago
$143k – $250k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree
Apply
$150k – $263k per year (Estimated) • Remote • Full-Time • 5+ years exp • Boston
Python
AI/ML
Spark
DeepSpeed
Fine-tuning
Scikit-learn
Multimodal AI
Accelerate
TensorFlow
Pandas
NumPy
PyTorch
LLM
Ray
Hugging Face
TPU
DevOps
GCP
AWS
Apply
AI Data Engineer 12 months ago
$112k – $239k per year (Estimated) • Remote • Boston
Python
Java
Scala
AI/ML
Spark
Airflow
Fine-tuning
LLM
DevOps
GCP
Azure
AWS
Cybersecurity
HIPAA
Analytics
ETL/ELT
Apply
See all jobs
This is one of many
614,885 more open roles from verified company boards, updated every day.