691,726open jobs
40,578companies
98,121added this week
Browse all
Location
Remote/Hybrid (Mexico City, Mexico)
Employment
Full-Time
Overview
Company
Impact
Profile match
Nubank is a Brazilian digital bank founded in 2013 that began by issuing a no-fee purple credit card to customers frustrated by the branch banking oligopoly, and it has grown into one of the largest financial institutions in Latin America by customer count. It now serves well over a hundred million customers across Brazil, Mexico and Colombia with accounts, cards, lending, investments and insurance, operating entirely without branches. Headquartered in São Paulo and listed on the New York Stock Exchange, it was an early Berkshire Hathaway investment and is unusual among digital banks for being consistently profitable at scale.

About Nu

Nu is the leading digital bank in Latin America, serving 140 million customers across Brazil, Mexico, and Colombia. The company has been leading an industry transformation by leveraging data and proprietary technology to develop innovative products and services.

Guided by its mission to fight complexity and empower people, Nu caters to customers’ complete financial journey, promoting financial access and advancement with responsible lending and transparency. The company is powered by an efficient and scalable business model that combines low cost to serve with growing returns.

Nu’s impact has been recognized in multiple awards, including Time 100 Most Influential Companies, Fast Company’s Most Innovative Companies, and Forbes World’s Best Banks.

Visit ourInstitutional Page

About the role

We are looking for a Reliability Engineer to drive reliability, resilience, and operational excellence for Nubank's regulatory infrastructure. In this role, you will own critical reliability outcomes for systems that support real-time interbank transfers, regulatory integrations and key regulatory platforms. You will work across infrastructure, software, networking, security, and operations to keep mission-critical services available, observable, and resilient under demanding regulatory and operational constraints.

What you'll do

  • Own end-to-end reliability for IT regulatory infrastructure services, including connectivity, transaction processing flows, and settlement-related operational workflows.

  • Define, measure, and continuously improve SLIs, SLOs, and operational health indicators for critical transactions and platforms.

  • Lead incident response for different severity production events, coordinate recovery, and drive high-quality postmortems with concrete follow-through on corrective actions.

  • Design and evolve observability for the platform, with emphasis on tracing, queue health, signature validation, infrastructure signals, and early detection of degraded regulatory links or transaction bottlenecks.

  • Plan and execute disaster recovery and business continuity exercises, including failover validation, contingency readiness, and RTO/RPO verification.

  • Reduce operational toil through automation, tooling, and improved runbooks for recurring failure and operational procedures.

  • Drive capacity planning and performance engineering to ensure the platforms can safely absorb peak transaction volumes and evolving business demand.

  • Partner closely with security, networking, middleware, and software engineering teams to improve platform hardening, resilience, and change management safety.

  • Participate in an on-call rotation.

What we're looking for

  • Experience with SRE, DevOps, or production engineering experience operating mission-critical, high-availability systems.

  • Hands-on experience with IT infrastructure platforms, virtualized and hyperconverged environments such as VxRail, and physical production infrastructure.

  • Hands-on experience with Linux systems, including performance tuning, kernel parameters, and security hardening.

  • Track record leading incident response and postmortem processes for customer-impacting services.

  • Solid knowledge of networking fundamentals: TCP/IP and routing.

  • Proficiency in at least one scripting or programming language (Python, Shell scripting) for automation and tooling.

  • Experience with observability stacks (Prometheus, Grafana) and distributed tracing.

  • AWS infrastructure experience across services like EC2, S3, CloudWatch, KMS, EKS/ECS, VPCs, RDS/DynamoDB, and SQS/SNS.

  • IT Regulatory audit experience, including evidence gathering, control validation, audit prep, and remediation follow-through.

  • Experience with infrastructure-as-code tools (like Puppet) and Git-based workflows.

  • Working knowledge of AI-assisted engineering tools (Claude, Cursor, Claude Code).

  • Strong written and verbal communication skills in Spanish and English.

Nice to have

  • Experience with Mexican payment rails (SPEI)

  • SQL knowledge for operational analysis and troubleshooting.

Relevant cloud (AWS Certified Cloud Practitioner certification or higher) or IT infrastructure certifications.

Location for this opportunity

  • Mexico City, Mexico

Our Benefits

  • Chance of earning equity at Nu

  • Extended maternity and paternity leaves

  • Health and life insurance

  • Dental and Vision Insurance

  • NuCare - Our mental health and wellness assistance program

  • Nucleo - Our learning platform of courses

  • NuLanguage - Our language learning program

  • Holiday Bonus ("Aguinaldo") of 30 days of pay per year

  • 17 days of paid vacation with 25% vacation bonus

  • Gym partnership

  • Food card

  • Work-from-home Allowance

  • Parental Consultancy

  • Relocation Assistance Package, if applicable

Work Model for this Role

Our recruitment process may involve the use of artificial intelligence-enabled tools, such as automated interview transcription and analysis, to support the evaluation process. Artificial intelligence is not used to make final hiring decisions; all decisions are made by human reviewers.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
691,726 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Mexico City
Jr Data Engineer 2 days ago
$53k – $110k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Toronto
Python
SQL
Databases
Google BigQuery
BigQuery
AI/ML
Claude
AI Agents
Human-in-the-Loop
DevOps
GCP
Git
Google Cloud Run
Analytics
ETL/ELT
Apply
$59k – $145k per year (Estimated) • In office • Contractor • Bachelor's Degree • Lansing
Python
PHP
SQL
Databases
MySQL
Mobile
Twilio
DevOps
AWS
Nginx
Amazon EC2
Amazon S3
Linux
Apache HTTP Server
Cybersecurity
HIPAA
Active Directory
LDAP
Analytics
SSIS
Management
ServiceNow
Apply
Success Manager 2 days ago
Remote
AI/ML
Claude
ChatGPT
DevOps
DNS
Management
Slack
Google Workspace
Marketing
HubSpot
Apply
Remote
AI/ML
Claude
ChatGPT
DevOps
SLI/SLO/SLA
DNS
Management
Zapier
Marketing
HubSpot
Apply
$50k – $70k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • Chicago
SQL
AI/ML
Claude
Gemini
Anthropic
Analytics
Microsoft Excel
Management
Slack
Google Sheets
Apply
$78k – $159k per year (Estimated) • Remote/Hybrid • Full-Time • São Paulo • Campinas • Palo Alto • Toronto • Belo Horizonte
Clojure
Clojure
Datomic
Databases
Redis
DynamoDB
Apache Kafka
Apply
$47k – $106k per year (Estimated) • Remote/Hybrid • Full-Time • São Paulo
Apply
FP&A Sr Expert 2 days ago
$45k – $98k per year (Estimated) • Remote/Hybrid • Full-Time • São Paulo
Apply
$164k – $273k per year (Estimated) • Remote/Hybrid • Full-Time • Toronto • Palo Alto • Miami
Clojure
AI/ML
AI Agents
RAG
Human-in-the-Loop
Context Engineering
Apply
Treasury Expert 3 days ago
$47k – $102k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Mexico City
Python
Apply
$30k – $42k per year • Equity 0.1–0.1% • In office • Full-Time • 6+ years exp • Bachelor's Degree • Mexico City
Apply
$29k – $69k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Mexico City
Apply
$45k – $109k per year (Estimated) • In office • Full-Time • 3+ years exp • Master's Degree • Mexico City
DevOps
GCP
Management
Microsoft Office
Apply
$49k – $113k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Mexico City
DevOps
SLI/SLO/SLA
Apply
$67k – $143k per year (Estimated) • In office • 7+ years exp • Mexico City
Apply
See all jobs
This is one of many
691,726 more open roles from verified company boards, updated every day.