743,556open jobs
44,602companies
107,147added this week
Browse all
Location
Remote (LATAM)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 24, 2026. Azumo scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Azumo is a leading software development company that specializes in nearshore services. The company offers a range of solutions, including software development, dedicated teams, staff augmentation, and virtual CTO services. Azumo is particularly known for its expertise in artificial intelligence, mobile app development, data engineering, and cloud services.

Cloud DevOps Engineer

Azumo builds and operates production AI systems for companies ranging from seed-stage startups to Meta. We are hiring a Cloud DevOps Engineer to own the infrastructure those systems run on: the clusters, the pipelines that deploy to them, and the monitoring that says whether any of it is healthy. The role is fully remote across Latin America, aligned to your client's working day.

You will not be handed a runbook. Azumo has shipped production infrastructure since 2016, and the work here is in systems that are already live: what breaks, what a cluster costs to run, and what nobody has automated yet because it was always easier to do by hand.

Where this role sits

At Azumo, ownership is split by what gets built: the pipelines, the method behind a decision, the product, and what an AI system does once it's live. This role owns what all of it runs on - clusters, deployment, and the infrastructure those systems depend on to stay up.

A system that is down, a slow query nobody escalates, a cluster that costs more than it should: they are all yours. Recovery first, root cause after, and the change that keeps it from happening again.

What you will build

  • Clusters that hold. Kubernetes in production - provisioning, upgrades, networking, and the workload configuration that decides whether a bad deploy takes one service down or all of them.
  • Infrastructure as code. Terraform or equivalent, with environments reproducible from the repository rather than from memory, and drift that gets detected instead of discovered.
  • Delivery pipelines. CI/CD that developers and QA can rely on, container images that are built the same way every time, and deployments to production that are unremarkable.
  • Observability and response. Monitoring and alerting that fires on what matters, plus the investigation, the root cause and the follow-up change when something does break.
  • Infrastructure for AI workloads. The clusters and inference services the AI Engineer lane deploys onto, whether that's hosted APIs or models we run ourselves, with the cost and scaling profile each one brings.
  • Cost and capacity. What the infrastructure costs to run, what it would cost at three times the traffic, and which of the two problems is worth solving now.
  • Hardening. Linux, Kubernetes, containers and the service mesh secured as a default rather than as a project, with secrets, access and audit trails that survive a client's review.
  • Tooling, documentation and support. The internal tooling your own team runs on, the documentation that makes any of it operable by someone else, and the developers and QA you unblock during the release cycle.
  • Work inside the client's environment, when that's the engagement. Some of this work runs on Azumo's own infrastructure and some inside a client's - their repositories, their cloud account, their change process. Azumo is SOC 2 certified, client code stays in client repositories, and some engagements carry additional requirements such as HIPAA.

How we work

Our engineers build with AI every day. Claude Code, Codex, and similar tools are part of the standard toolchain here, not an experiment. We run an automated audit across the whole codebase on day one and every day after, grading security, cost, and architecture findings by severity with the exact file and line, so a small team can move quickly without quality drifting. We stay vendor-neutral across OpenAI, Anthropic, and open-weight models, and we run Valkyrie, our own production layer, when a single interface to any model is the right call.

About Azumo

Azumo is a San Francisco based software development company that has been building intelligent applications since 2016. We provide nearshore AI engineering teams to organizations that need production AI faster than they can hire for it: as an embedded engineering team, as AI staff augmentation alongside an existing team, or as a full project build. Our engineers work from Latin America, aligned to United States time zones, and have delivered for Twitter, Meta, Discovery Channel, Omnicom, UnitedHealth, and CENTEGIX.

We hire for seniority and test for it before anyone joins a client team. We support engineers in going deep on the modern AI stack, and we give time back to open-source work, community teaching, and philanthropy.

Apply at https://azumo.com/join-our-team or write to us at [email protected].

Requirements

Basic qualifications

  • 5+ years as a DevOps, SRE or systems engineer running production infrastructure, with the fundamentals that go with it: Linux administration, networking, Git, and scripting in Bash plus Python or Go.
  • Kubernetes in production, at depth: not only deploying to a cluster someone else built, but provisioning, upgrading and debugging one when it misbehaves.
  • Infrastructure as code as your working practice: Terraform or equivalent, with state, modules and environment parity you can defend.
  • CI/CD pipelines you built and maintain - GitHub Actions, GitLab or equivalent - and container images you are responsible for.
  • Cloud deployment experience on AWS, Azure or GCP, including the managed Kubernetes service and the cost model that comes with it.
  • Monitoring and incident work: Datadog, CloudWatch or equivalent, and a track record of taking an incident from alert to root cause to a change that prevented the repeat.
  • Security hardening of Linux, containers and Kubernetes as part of how you build, not as a separate phase.
  • Working discipline around infrastructure cost and capacity. You can explain what a cluster costs to run and what you did about it.
  • Active use of AI-assisted coding tools such as Claude Code, Cursor, or GitHub Copilot in real delivery work.
  • Clear written and spoken English, C1 or above, and the confidence to explain a technical trade-off directly to a client.
  • Bachelor's degree in Computer Science, a related field, or equivalent professional experience.

Preferred qualifications

  • Service mesh and traffic management: Istio, Linkerd or equivalent, and the failure modes they introduce as well as the ones they solve.
  • Helm, Kustomize or an equivalent approach to templating and environment configuration.
  • Database operations at production scale: PostgreSQL, MongoDB, RDS, DynamoDB or equivalent.
  • Running or scaling inference workloads, and the cost and latency trade-offs against hosted APIs.
  • Delivery under a compliance regime such as SOC 2 or HIPAA.
  • Cloud certifications, or contributions to open-source infrastructure tooling and published technical writing.

Benefits

100% remote-first culture (work anywhere in Latin America)

Paid time off (PTO)

U.S. Holidays

Solid AI Training and certification

Mentored career development

Profit sharing

$US remuneration

Maternity coverage

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
743,556 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Buenos Aires
≈ $33k – $91k per year (Estimated) • Remote (Japan) • Full-Time • Tokyo
Python
PHP
Ruby
Ruby
Ruby on Rails
AI/ML
Cursor
Claude
Claude Code
LLM
DevOps
Terraform
AWS
SLI/SLO/SLA
GitHub
Management
Slack
n8n
Apply
≈ $114k – $210k per year (Estimated) • Remote (United States) • 7+ years exp
SQL
Apply
≈ $82k – $165k per year (Estimated) • Remote (United States, ET hours) • Full-Time • Philadelphia
Python
TypeScript
DevOps
Terraform
Azure DevOps
Azure
CI/CD
AWS
GitLab
Linux
Apply
$110k – $150k per year • Remote (United States) • Full-Time • 5+ years exp
Python
Databases
PostgreSQL
Redis
Snowflake
Databricks
Microsoft Fabric
DevOps
Terraform
Azure DevOps
GitHub Actions
Istio
Datadog
Envoy
Linkerd
Prometheus
GitLab CI
Azure
CI/CD
GitOps
Jenkins
AWS
Docker
Kubernetes
Grafana
Service Mesh
Self-Healing
Amazon EKS
Azure AKS
Amazon S3
Cybersecurity
Crowdstrike
HashiCorp Vault
Cryptography
Vault
Apply
≈ $124k – $242k per year (Estimated) • Remote (United States) • 8+ years exp • New York
Python
Go
JavaScript
Java
TypeScript
SQL
Node JS
Databases
Weaviate
Chroma
Pinecone
Amazon Redshift
AI/ML
LangGraph
AutoGen
LangChain
Vertex AI
Fine-tuning
Prompt Engineering
Function Calling
AI Agents
Semantic Kernel
AWS Bedrock
CrewAI
Gemini
LLM
RAG
Google ADK
OpenAI
Anthropic
LLM Guardrails
Multi-Agent Systems
Frontend
React.js
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Amazon EKS
Management
Slack
Google Workspace
Apply
≈ $55k – $123k per year (Estimated) • Remote (LATAM) • Full-Time • Buenos Aires
AI/ML
Copilot
Cursor
Claude Code
AI Agents
LLM
Anomaly Detection
OpenAI
Anthropic
OpenAI Codex
LLM Guardrails
Tool Use
DevOps
Terraform
Azure
CI/CD
Git
AWS
Bicep
Cybersecurity
PCI DSS
SOC 2
HIPAA
Threat Modeling
SIEM
Apply
≈ $25k – $74k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bucharest
Python
Java
Bash
DevOps
Terraform
Ansible
Red Hat
OpenShift
Azure DevOps
VMWare
Prometheus
Azure
CI/CD
Git
Docker
Kubernetes
Grafana
Platform Engineering
OpenStack
IAM
Linux
Cybersecurity
Checkmarx
Active Directory
Management
Agile
Scrum
Apply
≈ $74k – $148k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Atlanta
Python
PowerShell
Bash
AI/ML
AI Agents
Red Teaming
DevOps
CI/CD
IAM
Linux
Windows
Cybersecurity
MITRE ATT&CK
SIEM
Apply
≈ $58k – $111k per year (Estimated) • In office • Full-Time
Python
JavaScript
C#
C++
Management
Jira
Apply
Project Engineer 3 hours ago
≈ $41k – $98k per year (Estimated) • In office • Full-Time
Python
JavaScript
C#
C++
Apply
≈ $55k – $123k per year (Estimated) • Remote (LATAM) • Full-Time • Buenos Aires
AI/ML
Copilot
Cursor
Claude Code
AI Agents
LLM
Anomaly Detection
OpenAI
Anthropic
OpenAI Codex
LLM Guardrails
Tool Use
DevOps
Terraform
Azure
CI/CD
Git
AWS
Bicep
Cybersecurity
PCI DSS
SOC 2
HIPAA
Threat Modeling
SIEM
Apply
Remote (Argentina, Colombia, Mexico, Chile, Dominican Republic) • Full-Time • Bachelor's Degree
AI/ML
Claude
ChatGPT
Gemini
Perplexity
Design
Webflow
Marketing
Meta Ads
Google Ads
HubSpot
LinkedIn
Reddit
Instagram
Apply
≈ $46k – $122k per year (Estimated) • Remote (LATAM) • Full-Time • Bachelor's Degree
AI/ML
Claude Code
OpenAI
Anthropic
OpenAI Codex
DevOps
Terraform
GitHub Actions
Azure
CI/CD
AWS
Docker
Bicep
Cybersecurity
SOC 2
HIPAA
Apply
Data Scientist 3 days ago
≈ $43k – $116k per year (Estimated) • Remote (LATAM) • Full-Time • Bachelor's Degree
Python
AI/ML
Copilot
Cursor
Weights & Biases
Claude Code
LoRA
MLFlow
Vertex AI
Fine-tuning
Scikit-learn
Multimodal AI
PEFT
QLoRA
Kubeflow
Transformers
NumPy
PyTorch
LLM
OpenAI
Anthropic
Amazon SageMaker
OpenAI Codex
Machine Learning
DevOps
GCP
Azure
Git
AWS
Cybersecurity
SOC 2
HIPAA
Apply
Data Engineer 3 days ago
≈ $42k – $115k per year (Estimated) • Remote (LATAM) • Full-Time • Bachelor's Degree
Python
SQL
Databases
PostgreSQL
Snowflake
Databricks
pgvector
Pinecone
FAISS
Qdrant
Apache Kafka
Google BigQuery
Amazon Redshift
Azure SQL Database
BigQuery
AI/ML
Copilot
Cursor
Spark
Claude Code
Dagster
dbt
Prefect
Flink
RAG
OpenAI
Anthropic
OpenAI Codex
DevOps
Terraform
GitHub Actions
Azure
CI/CD
Git
AWS
Docker
Bicep
Cybersecurity
SOC 2
HIPAA
Analytics
Dimensional Modeling
Apply
Remote (Argentina) • Full-Time • Buenos Aires
Python
JavaScript
Java
C#
Node JS
C#
.NET
Databases
RabbitMQ
Apache Kafka
DevOps
OpenShift
Loki
Prometheus
Jenkins
AWS
Kubernetes
Grafana
AppDynamics
Amazon EKS
Management
ServiceNow
Apply
$18k – $20k per year • Remote (Argentina, MST hours) • Full-Time • Bachelor's Degree • Buenos Aires
1C
Apply
Hybrid • Full-Time • Buenos Aires
JavaScript
TypeScript
AI/ML
Vertex AI
LLM
Agentic Workflows
Frontend
Redux
Angular
Storybook
NgRx
Mobile
Firebase
State Management
DevOps
GCP
Google Cloud Run
Design
Figma
Apply
≈ $14k – $31k per year (Estimated) • Remote (Argentina) • Full-Time • 2+ years exp • Bachelor's Degree • Buenos Aires
AI/ML
Prompt Engineering
Management
Zapier
Apply
≈ $22k – $43k per year (Estimated) • In office • Internship • Buenos Aires
Apply
See all jobs
This is one of many
743,556 more open roles from verified company boards, updated every day.