368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$163k – $271k per year (Estimated)
Location
Remote (United States)
Seniority
Senior · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Workiva (NYSE: WK) is a global SaaS and a leading provider of a cloud-based connected and reporting compliance platform that enables the use of connected data and automation of reporting across finance and accounting, risk, and compliance.

Join Workiva as aSr Machine Learning Engineering Manager - AI Quality and Governance and help establish how we build, evaluate, release, and operate trustworthy AI products at scale. You will lead a multidisciplinary team of software, machine learning, and quality engineers responsible for two connected missions: advancing end-to-end quality across Workiva's AI platform and products, and building shared evaluation and governance capabilities that make our AI systems measurable, observable, reliable, and ready for enterprise use.

Your team's scope spans generative AI and agentic products, including AI platform services, agent frameworks and runtimes, conversational experiences, and RAG/knowledge systems. You will partner across Product, Engineering, Data Science, Security, Risk, and Legal to establish practical quality standards and embed evaluation and governance throughout the AI development lifecycle.

What You'll Do

Leadership & Team Development

  • Lead, mentor, and develop a multidisciplinary team of software, ML, and quality engineers

  • Build a culture of technical excellence, quality ownership, experimentation, and continuous improvement

  • Establish clear team priorities while balancing platform investments, product needs, and enterprise risk

  • Recruit engineers with complementary expertise across software quality, ML evaluation, platform engineering, and governance automation

AI Product Quality

  • Define and drive a comprehensive quality strategy for Workiva's AI platform and products, spanning unit, integration, end-to-end, performance, resilience, security, and production testing

  • Establish measurable quality bars, release-readiness criteria, and automated quality gates for AI and agentic capabilities

  • Advance testing approaches for nondeterministic systems, including RAG pipelines, agents, prompts, models, tools, and multi-step workflows

  • Detect regressions, model or data drift, unsafe behavior, and degraded customer experiences before and after release

AI Evaluation Platform

  • Lead architecture and delivery of a scalable, self-service evaluation platform for generative AI, RAG, and agentic systems

  • Enable teams to create, manage, version, and reuse evaluation datasets, golden test sets, task-specific metrics, graders, and benchmarks

  • Support deterministic checks, statistical metrics, model-based graders, human evaluation, adversarial testing, and domain-expert review

  • Build capabilities for offline evaluation, pre-release regression testing, online experimentation, production sampling, and continuous evaluation

  • Ensure evaluation results are reproducible, explainable, actionable, and integrated into developer workflows, CI/CD pipelines, and operational dashboards

AI Governance & Assurance

  • Translate Workiva's Responsible AI principles into practical engineering controls and platform capabilities

  • Build governance into the AI lifecycle through traceability, lineage, versioning, documentation, risk classification, approval workflows, and auditable evidence

  • Partner with Security, Legal, Privacy, Compliance, and Risk teams to define controls that support enterprise and regulated use cases

  • Enable inventories and traceability across models, prompts, datasets, evaluations, tools, knowledge sources, and deployed AI features

Cross-Functional Leadership

  • Collaborate with Product, Program Management, UX, UXR, Data Science, Security, Legal, Risk, and engineering leaders to define quality expectations and roadmaps

  • Influence engineering teams across Workiva to adopt shared evaluation standards, testing practices, observability, and release controls

  • Communicate complex technical tradeoffs, quality signals, and risk findings clearly to technical and non-technical audiences

Operational Excellence

  • Ensure the evaluation and governance platform is secure, scalable, reliable, observable, and cost-effective
  • Define service-level objectives and meaningful operational and quality metrics

  • Champion production readiness, incident response, root-cause analysis, and continuous operational improvement

What You'll Need

Minimum Qualifications

  • Bachelor's degree in Computer Science, Engineering, Data Science, or related field (or equivalent experience)

  • 10+ years in software engineering, ML engineering, quality engineering, or related roles, including 4+ years leading an engineering team

  • Strong software engineering and systems-design fundamentals, with experience delivering and operating production SaaS or platform capabilities

  • Demonstrated experience establishing automated quality practices for distributed, cloud-based products

  • Practical understanding of the generative AI development lifecycle and challenges of evaluating nondeterministic systems

  • Experience with generative AI concepts: LLMs, RAG, embeddings, vector/hybrid search, agents, tool use, and prompt orchestration

  • Experience defining measurable quality criteria using data, experimentation, telemetry, and production signals

  • Experience with cloud-native architectures on AWS, Azure, or GCP.

  • Proven ability to lead senior individual contributors, navigate tehhnical disagreements, and build high-performance cultures

  • Strong communication and cross-functional leadership skills

Preferred Qualifications

  • Master's degree in Computer Science, Engineering, ML, Data Science, or related field.

  • Experience building or operating AI/ML evaluation, experimentation, observability, model-governance, or ML platform capabilities

  • Experience evaluating RAG and agentic systems, including retrieval quality, groundedness, task completion, tool use, and safety

  • Familiarity with evaluation techniques: golden datasets, statistical metrics, model-based graders, human evaluation, red teaming, A/B testing, and drift/regression detection

  • Working knowledge of ML/AI lifecycle practices: dataset management, model/prompt versioning, experiment tracking, deployment, monitoring, and feedback loops

  • Experience translating Responsible AI, model-risk, privacy, security, or regulatory requirements into scalable engineering controls

  • Familiarity with AI risk/governance frameworks (NIST AI RMF, ISO/IEC 42001, or comparable)

  • Experience with Kubernetes, microservices, CI/CD, infrastructure as code, and modern DevOps/MLOps practices

  • Experience supporting enterprise software in regulated or high-assurance environments

Working Conditions

  • Willingness to travel up to 15% for team and corporate meetings

  • Reliable internet access for remote work

How You’ll Be Rewarded

Salary range in the US: $193,000.00 - $308,000.00

A discretionary bonus typically paid annually

Restricted Stock Units granted at time of hire

401(k) match and comprehensive employee benefits package

The salary range represents the low and high end of the salary range for this job in the US. Minimums and maximums may vary based on location. The actual salary offer will carefully consider a wide range of factors, including your skills, qualifications, experience and other relevant factors.

Why Join Workiva

Workiva is the platform designed to bring confidence, control, and a competitive edge to the world’s most complex organizations. Our AI-powered platform unifies finance, risk, and sustainability on a single, secure foundation-ensuring data is trusted, traceable, and ready to act on. With an unbroken path from source to output, leaders gain confidence in their numbers, visibility into current and emerging risks, and the ability to move with speed and precision in a constantly changing world.

At Workiva, you’ll bring technology to market that executives, boards, and regulators depend on. The work you do here helps organizations navigate uncertainty, maintain trust, and make decisions that stand up to scrutiny. If you’re energized by meaningful challenges, inspired by collaborative teams, and motivated to help organizations turn uncertainty into advantage, we’d love to meet you.

Employment decisions are made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other protected characteristic.

Workiva is committed to working with and providing reasonable accommodations to applicants with disabilities. To request assistance with the application process, please email [email protected].

Workiva employees are required to undergo comprehensive security and privacy training tailored to their roles, ensuring adherence to company policies and regulatory standards.

Workiva supports employees in working where they work best - either from an office or remotely from any location within their country of employment.

#LI-MJ2
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
United States
$24k – $64k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Chennai
Python
Scala
SQL
TypeScript
JavaScript
Java
Python
pySpark
Java
Spring Boot
Databases
Apache Kafka
Databricks
AI/ML
AI Agents
Copilot
Google ADK
LLM
NLP
Prompt Engineering
Spark
Devin
Model Context Protocol
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
OpenShift
GitHub
Analytics
ETL/ELT
Apply
$42k – $104k per year (Estimated) • In office • Internship • 5+ years exp
Bash
PowerShell
Python
DevOps
Amazon EC2
Amazon EKS
Ansible
AWS
AWS Lambda
Azure
CentOS Stream
Chef
CI/CD
CloudFormation
Configuration Management
Datadog
FinOps
GCP
Git
GitHub Actions
GitLab CI
Grafana
Hyper-V
Istio
Jenkins
Kubernetes
KVM
Linkerd
Platform Engineering
Prometheus
Puppet
Service Mesh
Splunk
Terraform
Ubuntu
VMWare
Windows Server
Amazon CloudWatch
Amazon ECS
Amazon S3
API Gateway
AWS Step Functions
GitHub
GitLab
IAM
Cybersecurity
GDPR
ISO 27001
SOC 2
Apply
Data Engineer 1 day ago
In office • Full-Time • Singapore
Node JS
Python
SQL
JavaScript
Python
Beautiful Soup
Databases
Apache Kafka
MySQL
PostgreSQL
RabbitMQ
SQLite
AI/ML
Hadoop
Spark
DevOps
AWS
AWS Lambda
Azure
CI/CD
GCP
Amazon ECS
Amazon EventBridge
Amazon S3
Analytics
ETL/ELT
QA
Selenium
Apply
$170k – $318k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
JavaScript
Python
TypeScript
Python
pySpark
AI/ML
Prompt Engineering
Spark
DevOps
AWS
Azure
CI/CD
GCP
Git
Jenkins
GitHub
GitLab
Analytics
ETL/ELT
Apply
up to $48k per year (net) • Remote/Hybrid • Moscow
Node JS
Python
TypeScript
JavaScript
Node JS
Nest.JS
Python
Django
FastAPI
Databases
Apache Kafka
pgvector
Pinecone
PostgreSQL
Qdrant
RabbitMQ
Redis
AI/ML
Chain-of-Thought
Claude
Claude Code
Copilot
Cursor
LangChain
LlamaIndex
LLM
Prompt Engineering
RAG
Anthropic
Function Calling
OpenAI
Structured Outputs
Frontend
GraphQL
Next.js
React.js
Redux
Redux Toolkit
Zustand
Mobile
State Management
DevOps
AWS
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Kubernetes
Prometheus
Yandex Cloud
GitHub
GitLab
Apply
$146k – $274k per year (Estimated) • Equity • In office • 6+ years exp • Bachelor's Degree • Denver
Python
SQL
Apex
Apex
MuleSoft
Databases
Amazon Redshift
Databricks
Google BigQuery
Snowflake
DevOps
Azure
Analytics
Power BI
ETL/ELT
Apply
$130k – $228k per year (Estimated) • Remote • 6+ years exp • Bachelor's Degree • Canada
Python
SQL
Apex
Apex
MuleSoft
Databases
Amazon Redshift
Databricks
Google BigQuery
Snowflake
DevOps
Azure
Analytics
Power BI
ETL/ELT
Apply
Remote • Internship • United States
C#
C++
Go
JavaScript
Python
SQL
TypeScript
Frontend
Angular
React.js
DevOps
Amazon EC2
AWS
Apply
Remote • Internship • Bachelor's Degree • United States
Python
AI/ML
AI Agents
LlamaIndex
LlamaParse
DevOps
AWS
Docker
Git
Kubernetes
Rest API
Apply
Remote • Internship • United States
Apply
$60k – $61k per year • In office • Full-Time • 2+ years exp • High School Diploma • United States
Apply
$62k – $63k per year • In office • Full-Time • 2+ years exp • High School Diploma • United States
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
$120k – $140k per year • In office • Full-Time • Charlotte • Raleigh • Dallas • Boston • New York
SQL
Analytics
ETL/ELT
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
DevOps
SLI/SLO/SLA
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.