1,173,434open jobs
66,192companies
208,404added this week
Browse all
Salary
$155k – $190k per year
Location
Remote (United States)
Seniority
Architect · 10+ years exp

Confirmed on the employer's own hiring board on Oct 4, 2026. First seen by Alion on Sep 15, 2026. Signify Technology scores C on the Alion truth index.

Overview
Company
Impact
Profile match
Signify Technology is a global staffing agency connecting exceptional talent in emerging technologies across leadership, software engineering, data, and cloud. We are minority-owned (MSDUK certified) with offices in London and Austin.
Job title: Principal DevOps Architect

Job type: Permanent

Salary: $155,000 - $190,000

Role Location: Remote, US

Visa requirements: Not specified

The company:

Our client is a global leader and rapidly growing SaaS technology company that provides comprehensive software solutions to the clinical research industry.

Their vision is to reshape the global clinical research industry with innovative solutions that help advance medicine and save lives. Their cloud-based solutions are dedicated to solving complex problems and simplifying clinical research processes to be more organized, efficient, and cost-effective. They are based out of San Antonio, TX, but are truly a remote and telecommuting company.

Role and responsibilities:

The Principal DevOps Architect is the senior technical authority for the cloud platform that runs our client's global, multi-tenant clinical-trial services, and drives our client's adoption of infrastructure as code and AI tooling. This is a hands-on individual-contributor architect role rather than a people-management role: the successful candidate will design the platform architecture, set the standards and reference implementations other teams build on, and still write Terraform, build pipelines, and stand up the AI platform themselves. They will define how reliability is measured against SLOs, how releases ship, and how the AI/ML platform is built and governed, and will keep the environment HIPAA-compliant and SOC 2 Type 2 audit-ready. They will influence products, software, and QA through architecture and example rather than direct authority.

Technical Leadership & Architecture

  • Own the platform architecture and technical roadmap for infrastructure, deployment, observability, and the AI platform, and advise the VP of Software Architecture and engineering leadership.
  • Set the engineering standards, patterns, and golden paths for infrastructure as code, CI/CD, and AI tooling, and drive their adoption through reference implementations and architecture reviews.
  • Act as hands-on technical authority and mentor to engineers across teams, leading by example in code, infrastructure, and incident response, without direct management responsibility.
  • Drive our client's move to infrastructure as code and AI tooling and prevent uncontrolled spread of unvetted AI tools.
  • Recommend build/buy decisions for platform tooling and evaluate vendors, with final decisions owned by the VP and CTO.

Infrastructure as Code & Cloud Platform

  • Manage all cloud infrastructure as code in Terraform as the single source of truth: reusable modules, remote state, peer-reviewed infrastructure pull requests, drift detection, and automated plan/apply in CI/CD.
  • Enforce policy-as-code (for example, OPA or Sentinel) so infrastructure changes meet security and cost guardrails before they merge, with separation between the author and approver of a change.
  • Design and deploy AWS infrastructure for performance, availability, recoverability, and security across development, UAT, staging, and production, meeting the CIS Critical Security Controls.
  • Maintain and improve the multi-tenant database and hosting architecture in a cloud-hosted environment, including replication and per-tenant isolation.

CI/CD & Release Engineering

  • Build and operate CI/CD pipelines for large-scale applications on AWS using GitHub Actions or equivalent, with automated build, test, and deployment.
  • Own release management and rollback: safe deployment strategies (blue/green, canary), fast and reliable rollbacks, and release gates for QA and customer acceptance.
  • Package and run containerized workloads on Docker and Kubernetes (EKS), and automate configuration with tools such as Ansible and Packer.

Reliability & Observability

  • Lead the SLI/SLO/SLA program and manage error budgets to balance reliability against delivery speed.
  • Operate modern observability using cloud-native tooling and OpenTelemetry (metrics, logs, and distributed traces) to drive down MTTD and MTTR.
  • Analyze production events to improve reliability, operability, and customer experience, and lead blameless post-incident reviews.
  • Participate in on-call rotations, triage and resolve incidents, and provide workarounds or escalation to service owners.

AI / ML Platform Operations

  • Provision and operate the AI/ML platform, including Anthropic Claude models served through AWS Bedrock and, where used, the Anthropic API, along with inference endpoints and our client's internal MCP services, all managed as infrastructure as code.
  • Build guardrails and observability for AI workloads: prompt and response logging, rate limits, model and token cost controls, and usage auditing across Claude and any other models in use.
  • Enforce the data boundary for AI systems so Protected Health Information is handled only through approved, BAA-covered providers (AWS Bedrock under the AWS BAA, and Anthropic under an Anthropic BAA where the Anthropic API is used), and confirm no PHI leaves the compliant boundary.
  • Operate and support the agentic and MCP tooling used across engineering and operations (for example Claude Code and internal MCP services) and govern how that tooling accesses code and data with least-privilege tool access and human-in-the-loop escalation.
  • Stand up LLMOps practices: prompt versioning and regression detection, evaluation frameworks for non-deterministic output, token-level cost attribution, and audit logging of agent actions.
  • Route inference across a multi-provider model portfolio and manage protocol choices (such as MCP) to avoid vendor lock-in.
  • Use AI-assisted tooling to accelerate operations where appropriate, such as incident triage, runbook generation, and log analysis.

Security, Compliance & Cost

  • Operate and evidence the platform controls required for SOC 2 Type 2 and HIPAA continuously and support external audits.
  • Own secrets management so no credentials, keys, or tokens live in source code, container images, or Terraform state.
  • Run supply-chain security for the pipeline: image and dependency scanning, SBOM generation, and vulnerability remediation to defined SLAs.
  • Manage cloud cost (FinOps): tagging, budgets, and right-sizing across compute, storage, and AI/Bedrock spend.

Collaboration

  • Coordinate with product, development, support, operations, and QA so installation and integration are automated and well-documented.
  • Communicate and work effectively across a distributed, multi-time-zone team, and cultivate cross-team collaboration and trust.

Job requirements:

  • Bachelor's degree in software engineering or an equivalent combination of technical education and work experience.
  • 10+ years in SRE/DevOps/platform engineering delivering CI/CD, REST API deployment, containerization, IaaS/PaaS, data pipelines, and application observability, including time at a senior individual-contributor or architect level (Staff, Principal, or Architect).
  • Proven technical authority across teams: setting architecture and standards and influencing delivery through expertise and example rather than direct management.
  • Demonstrated experience driving adoption of a new practice or platform (infrastructure as code, a CI/CD overhaul, or an AI/ML platform) across multiple teams.
  • Hands-on experience with Terraform, including writing reusable modules that other teams consume through self-service, remote state, and change management in a CI/CD pipeline.
  • Experience building and operating CI/CD pipelines for large-scale applications on AWS (GitHub Actions, Jenkins, GitLab, or AWS-native).
  • Experience running containerized workloads on Docker and Kubernetes.
  • Experience with monitoring and troubleshooting using cloud-native tooling and OpenTelemetry; New Relic experience a plus.
  • Linux system administration, Unix scripting, and automation.
  • 2+ years in one or more of PHP, MySQL, and SQL is a plus, matching our client's stack.
  • Experience working in a HIPAA / HITECH / HITRUST / PHI / PII or PCI DSS environment.

What sets you apart:

  • Experience operating an AI/ML or GenAI platform in production, including Anthropic Claude via AWS Bedrock or the Anthropic API, or comparable model-serving infrastructure, with cost and guardrail controls.
  • LLMOps maturity: evaluation pipelines for non-deterministic output, prompt versioning, token cost attribution, and incident response for AI-specific failures (data modification, unwanted workflow triggering, or exfiltration by an agent).
  • Experience supporting LLM applications, agents, or MCP services (for example Claude Code or custom MCP servers), and governing how they access regulated data.
  • Familiarity with AI security and governance: OWASP LLM Top 10, key management (BYOK), and alignment to frameworks such as the NIST AI RMF.
  • Platform-as-product mindset with a focus on internal developer experience and self-service.
  • Experience leading or supporting a SOC 2 Type 2 examination and familiarity with security benchmarks such as CIS, OWASP, PCI DSS, and FedRAMP.
  • Experience with policy-as-code, secrets management, and supply-chain security (SBOM, image scanning).
  • Experience with FinOps and managing large AWS infrastructure and its cost.
  • Experience in clinical research or healthcare technology.
  • Self-starter who establishes best practices, ships on time, and communicates complex technical information clearly to any audience.

Benefits:

  • Our client sponsors health insurance, long-term disability, and life insurance.
  • Unlimited paid time off.
  • 10 paid holidays.
  • Paid parental leave.
  • Work anniversary bonus.
  • Participation in the Employee of the Quarter program.
  • Monthly $100 connectivity stipend reimbursement.
  • Our client matches employee 401(k) contributions at 100% of the first 3% invested and 50% of the next 2% invested.

Accessibility Statement:

Read and apply for this role in the way that works for you by using our Recite Me assistive technology tool. Click the circle at the bottom right side of the screen and select your preferences.

We make an active choice to be inclusive towards everyone every day.? Please let us know if you require any accessibility adjustments through the application or interview process.

Our Commitment to Diversity, Equity, and Inclusion:

Signify’s mission is to empower every person, regardless of their background or?circumstances, with an equitable chance to achieve the careers?they deserve. Building a diverse future, one placement at a?time.

Check out our DE&I page here

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,173,434 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
In your city
$181k – $263k per year • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco • New York • Seattle
Python
Databases
Cassandra
DynamoDB
ScyllaDB
AI/ML
Claude Code
AI Agents
DevOps
Terraform
GCP
CircleCI
CI/CD
Jenkins
AWS
Kubernetes
Platform Engineering
Chaos Engineering
FinOps
SLI/SLO/SLA
IAM
Cybersecurity
ISO 27001
SOC 2
Apply
≈ $112k – $232k per year (Estimated) • In office • 5+ years exp • Houston
Python
Java
C#
DevOps
Terraform
OpenShift
GitHub Actions
CloudFormation
Dynatrace
CI/CD
Jenkins
AWS
Kubernetes
Grafana
Incident Management
GitLab
IAM
Apply
≈ $109k – $225k per year (Estimated) • In office • 5+ years exp • Jersey City
PowerShell
C#
C#
.NET
DevOps
Ansible
VMWare
CI/CD
Platform Engineering
Configuration Management
KVM
Hyper-V
Cybersecurity
Qualys Cloud Platform
Apply
≈ $151k – $330k per year (Estimated) • In office • 8+ years exp • New York
AI/ML
Model Context Protocol
AI Agents
LLM
Structured Outputs
Apply
≈ $104k – $214k per year (Estimated) • In office • 5+ years exp • Plano
Python
Databases
Snowflake
Databricks
DevOps
AWS
Incident Management
IAM
Linux
Windows
DNS
Cybersecurity
Least Privilege
Active Directory
Analytics
Tableau
Apply
≈ $83k – $167k per year (Estimated) • Hybrid • Bachelor's Degree • Sterling
JavaScript
Java
SQL
Java
Spring Boot
DevOps
Rest API
CI/CD
Git
Management
Agile
Apply
≈ $15k – $29k per year (Estimated) • In office • 5+ years exp • Pune
SQL
Apply
≈ $13k – $32k per year (Estimated) • In office • 7+ years exp • Bengaluru
Python
SQL
DevOps
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
Jenkins
AWS
Amazon CloudWatch
Cybersecurity
OWASP Top 10
Management
ServiceNow
ITSM
QA
Selenium
JMeter
Cypress
Playwright
Postman
Rest-Assured
Locust
Apply
≈ $62k – $131k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Lagos
Python
SQL
Analytics
Power BI
Apply
In office • 7+ years exp
ABAP
ABAP
ABAP on HANA
CDS Views
DevOps
Rest API
Apply
$215k per year • Hybrid • 10+ years exp • Orlando
DevOps
Terraform
GCP
Pulumi
Azure
CI/CD
AWS
Kubernetes
Platform Engineering
FinOps
Incident Management
Apply
SRE/DevOps Engineer 3 days ago
$40k – $60k per year • In office • 3+ years exp • Brazil
Python
DevOps
Terraform
GCP
GitHub Actions
Datadog
CI/CD
Jenkins
AWS
Docker
Kubernetes
IAM
DNS
Cybersecurity
PCI DSS
Apply
≈ $117k – $227k per year (Estimated) • Hybrid • Bachelor's Degree • Austin
DevOps
GCP
Kubernetes
Apply
$61k – $81k per year (gross) • In office • 4+ years exp • Bachelor's Degree • Amsterdam
Python
DevOps
CI/CD
Docker
Linux
Apply
≈ $90k – $177k per year (Estimated) • In office • London
Python
SQL
C++
DevOps
Kubernetes
Linux
Unix
TCP/IP
Apply
See all jobs
This is one of many
1,173,434 more open roles from verified company boards, updated every day.