666,674open jobs
39,006companies
99,917added this week
Browse all
Salary
$35k – $101k per year (Estimated)
Location
Remote (Mexico)
Employment
Full-Time
Overview
Company
Impact
Profile match
Helpware is a business process outsourcing company that provides customer support, technical support, back-office, sales, content moderation and patient services, alongside AI implementation and custom software development through its Helpware CX, Helpware AI and Helpware Tech divisions. Incorporated as Helpware Inc. in 2015, it runs about 19 locations with teams in the US, Mexico, Ukraine, Germany, Albania, Poland, Puerto Rico, Guam and the Philippines, including Manila and Cebu. It hires customer service and technical support representatives, patient enrollment specialists, team leaders, telesales agents, recruiters, QA engineers and full stack developers.

About the role

Helpware is building a multi-tenant B2B AI platform for customer support operations - a product portfolio spanning Agent Assist, AI-powered QA, document processing, support automation, workforce management, and others. We're hiring a Sr. PlatformOps Engineer to own the reliability, security, and operability of this infrastructure, and to carry classic DevOps responsibilities: CI/CD, environments, and developer enablement.

This is a platform-ownership role. Your job is to help the development team make architectural designs operationally real, keep them healthy as tenant count grows, and enforce them structurally (in pipelines and shared tooling).

What you'll do

Platform infrastructure (AWS)

  • Contribute with cloud architectural designs, focusing on Well-Architected Framework pillars: Security, Operational Excellence, Performance Efficiency, Reliability and Cost Optimization.
  • Own the AWS footprint: ECS Fargate services, Lambda and Batch workloads, RDS PostgreSQL, ElastiCache Redis, S3, CloudWatch, networking/IAM, Bedrock, etc.
  • Manage everything as code (Terraform preferred), with reviewable, repeatable environment provisioning for dev/test/staging/prod.
  • Operate our Observability and self-hosted Langfuse stack: deployment, upgrades, backup/restore, and access control (SSO-backed, role-scoped, itself audit-logged).

DevOps & developer enablement

  • Build and maintain CI/CD pipelines, including the platform's conformance gates: contract tests for the logging standard and the logging conformance linter.
  • Own secrets management, artifact/image pipelines, and deployment strategies (blue/green or rolling) across services.

Monitoring and Support

  • Own our OpenTelemetry pipeline (SDK → Collector → CloudWatch for app/infra; Langfuse for LLM traces), keeping the three-layer separation between infra observability, LLM observability, and the business audit log intact.
  • Maintain per-tenant usage and cost infrastructure, platform-level dashboards, alerting, and SLOs.
  • Keep trace retention and masking aligned with the data governance standard (masked content only in LLM traces, defined retention windows per layer).
  • Support product teams day-to-day: environment issues, deployment questions, performance debugging.

Security & compliance

  • Maintain the security posture underpinning SOC 2 / HIPAA / GDPR commitments: least-privilege IAM, network isolation, encryption at rest and in transit, audit trails.
  • Ensure prompt/completion data, support the PII masking architecture at the infrastructure level.

What we're looking for

  • 5+ years in platform, SRE, or DevOps roles running production workloads on AWS (and other Cloud Platforms, preferred)
  • Deep hands-on experience with ECS (or EKS), RDS, ElastiCache, S3, CloudWatch, IAM, and VPC networking.
  • Strong infrastructure-as-code practice (Terraform or CDK) and CI/CD ownership (GitHub Actions, GitLab CI, or similar).
  • Experience with Claude Code and AI Spec Driven Design
  • Experience operating stateful open-source systems in production - running ClickHouse, or comparable columnar/analytics stores (Druid, Pinot) or demonstrably transferable database operations depth (Postgres at scale, Elasticsearch, Kafka).
  • Working knowledge of OpenTelemetry: collectors, exporters, sampling, and instrumentation conventions.
  • A security-first habit: you treat IAM policies, secrets, and data boundaries as design work, not afterthoughts.

Nice to have

  • Experience with multiple Cloud providers (AWS, Azure, GCP).
  • Experience self-hosting LLM/GenAI tooling (Langfuse, LiteLLM, vLLM) or operating GenAI workloads (Bedrock, token cost management).
  • Scripting/automation fluency (Python or Go preferred; the platform's services are Python/FastAPI).
  • Prior work in a compliance-driven environment (SOC 2 audits, HIPAA, GDPR data residency).
  • Experience with multi-tenant SaaS operations: tenant isolation, per-tenant metering, noisy-neighbor management.
  • FinOps experience - AWS cost allocation, tagging strategies, per-tenant cost attribution.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
666,674 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$28k – $60k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
Databases
PostgreSQL
Redis
Apache Kafka
DevOps
Terraform
Ansible
GCP
Helm
Azure DevOps
Prometheus
Pulumi
Azure
CI/CD
ArgoCD
Jenkins
AWS
Docker
Kubernetes
Grafana
Service Mesh
Incident Management
GitHub
Apply
$22k – $43k per year (Estimated) • Remote • 3+ years exp • Nizhny Novgorod
JavaScript
TypeScript
SQL
Node JS
Node JS
Nest.JS
BullMQ
Databases
PostgreSQL
Redis
AI/ML
LLM
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
Material UI
shadcn/ui
DevOps
Rest API
Docker Compose
Git
Docker
Kubernetes
Apply
$20k – $51k per year (Estimated) • Remote • Moscow
JavaScript
TypeScript
Node JS
Databases
PostgreSQL
NATS
Frontend
GraphQL
DevOps
gRPC
Docker
Kubernetes
Apply
AI инженер 3 hours ago
$23k – $60k per year (Estimated) • In office • Moscow
Python
AI/ML
LangGraph
LangChain
ChatGPT
Claude Code
AI Agents
Gemini
LLM
RAG
OpenAI
Anthropic
Management
n8n
Telegram
Apply
$28k – $59k per year (Estimated) • In office • Bengaluru
Python
Bash
DevOps
Terraform
Azure DevOps
GitHub Actions
CloudFormation
Crossplane
Azure
CI/CD
GitOps
Windows Server
ArgoCD
Jenkins
AWS
Kubernetes
Ubuntu
Platform Engineering
Backstage
Amazon EKS
Azure AKS
Apply
$26k – $61k per year (Estimated) • Remote • Full-Time • 5+ years exp
AI/ML
LLM
RAG
Hallucination
DevOps
CI/CD
AWS
Cybersecurity
PCI DSS
SOC 2
HIPAA
Management
Intercom
Freshdesk
QA
Selenium
Cypress
Playwright
Postman
Rest-Assured
Apply
$13k per year (gross) • In office • Full-Time • 1+ year exp • Guadalajara
Apply
$9.5k – $25k per year (Estimated) • In office • Full-Time • Cebu City
Cybersecurity
HIPAA
Marketing
Salesforce
HubSpot
Apply
In office • Full-Time • 1+ year exp • Cebu City
Cybersecurity
HIPAA
Marketing
Salesforce
HubSpot
Apply
$37k – $77k per year (Estimated) • In office • Full-Time • 2+ years exp • Lexington
Marketing
Salesforce
Apply
See all jobs
This is one of many
666,674 more open roles from verified company boards, updated every day.