726,270open jobs
43,325companies
103,498added this week
Browse all
Salary
$80k – $179k per year (Estimated)
Location
Remote (São Paulo, Brazil)
Seniority
Staff
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 23, 2026.

Overview
Company
Impact
Profile match
A frontier AI lab building predictive models for enterprise decisions. We pre-train a Graph Foundation Model on the relational economy, then fine-tune inside your workspace — for credit, fraud, marketing, growth, and any task with a graph underneath it.

About the role

At Avra, every technical IC is a Member of Technical Staff (MTS). The title doesn't put anyone in a silo: you own systems and outcomes, not steps in a function, and you keep building depth in your area. Seniority shows up in your scope, level, and compensation, not in titles.

In this role, you'll join the Platform team to own where our models execute. Customers consume our models through large batches of millions of records and through real-time APIs, and they make business decisions on every response. You'll run governed model releases reliably and efficiently - in our cloud and on customer-hosted Kubernetes - and make inference fast, predictable, and cheap enough to serve both enterprise and mid-market customers.

What you'll do

  • Evolve Sophos, our online and batch inference runtime, built on Kubernetes.

  • Run large batch inference on ephemeral jobs, with multi-dimensional admission control (CPU, memory, GPU) through Kueue.

  • Build and extend the Sophos controller and its Kubernetes custom resources.

  • Optimize each model's inference engine and feature processing, using vectorized, columnar operations.

  • Serve graphs and data efficiently from Lance-based storage.

  • Own execution of training, post-training, and fine-tuning jobs, in our cloud and in customer dataplanes.

  • Drive autoscaling, GPU serving, performance, and cost optimization, with telemetry for every model we run.

  • Solve open problems such as deterministic job sizing, checkpointing and recovery for batch runs, per-customer encryption and isolation, resilience to difficult input files, and automatic profiling when a new model is accepted.

How we measure success

  • 99.9% serving availability.

  • p95/p99 latency for online inference and throughput for batch.

  • Cost per prediction and per training job.

  • GPU utilization: paid capacity versus capacity actually used.

  • Training and batch jobs that finish on time and succeed without manual retries.

What we're looking for

  • Experience running model serving or large-scale batch compute on Kubernetes.

  • Experience building Kubernetes controllers or operators.

  • Skill at profiling and optimizing data-heavy Python pipelines.

  • A clear sense of cost: you treat compute efficiency as a product feature.

  • Production-quality code and reviews, and a willingness to operate what you build.

Nice to have

  • Ray, Ray Serve, or KubeRay in production.

  • Kueue or other batch scheduling and admission-control systems.

  • GPU serving and performance optimization.

  • Arrow, Parquet, Lance, or other columnar formats.

  • Shipping software to customer-hosted Kubernetes.

  • GCP/AWS and GKE/EKS, and financial services or regulated environments.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
726,270 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
São Paulo
$67k – $160k per year (Estimated) • Remote/Hybrid • Full-Time • Munich • Berlin
Python
PowerShell
DevOps
Terraform
GitLab CI
Azure
CI/CD
Docker
Kubernetes
Configuration Management
Bicep
FinOps
IAM
Cybersecurity
ISO 27001
PCI DSS
SOC 2
GDPR
Zero Trust
Apply
$45k per year (net) • In office • Bachelor's Degree • Moscow
JavaScript
Java
TypeScript
C#
Java
Testcontainers
C#
ASP.NET Core
IdentityServer
MediatR
gRPC for .NET
xUnit
FluentValidation
MassTransit
Refit
Databases
PostgreSQL
Redis
RabbitMQ
AI/ML
Cursor
Claude Code
Frontend
Vue.js
Vuetify
Mobile
Material Design
DevOps
gRPC
OpenTelemetry
Prometheus
CI/CD
Docker
Kubernetes
Grafana
GitLab
Cybersecurity
Keycloak
Apply
Intern - Data Analyst 3 hours ago
$41k – $114k per year (Estimated) • In office • Internship • Christchurch
Python
Java
SQL
Databases
Databricks
Analytics
Power BI
Microsoft Excel
Apply
AI Engineer 3 hours ago
In office • Full-Time • Bachelor's Degree • Gurgaon
Python
SQL
Databases
Databricks
Microsoft Fabric
AI/ML
LangGraph
LangChain
Spark
Embeddings
Prompt Engineering
Multimodal AI
AI Agents
TensorFlow
PyTorch
Gemini
LLM
RAG
Hallucination
Semantic Search
OpenAI
Hugging Face
LLMOps
Semantic Search
LLM Guardrails
Machine Learning
DevOps
Azure
CI/CD
AWS
Apply
$28k – $64k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Gurgaon • Bengaluru
Python
SQL
Databases
Databricks
DevOps
GCP
Azure
CI/CD
AWS
Analytics
ETL/ELT
Dimensional Modeling
Apply
$78k – $175k per year (Estimated) • Remote • Full-Time • São Paulo
Python
AI/ML
CUDA Toolkit
SkyPilot
PyTorch
Ray
CUDA
Apply
$77k – $172k per year (Estimated) • Remote • Full-Time • São Paulo
AI/ML
Ray
DevOps
Terraform
GCP
Helm
OpenTelemetry
AWS
Kubernetes
Amazon EKS
Google GKE
Incident Management
Apply
In office • Full-Time • 2+ years exp • São Paulo
Python
SQL
Analytics
A/B Testing
Apply
Remote/Hybrid • Full-Time • 2+ years exp • São Paulo
AI/ML
AI Agents
LLM
Apply
Operator 1 month ago
$21k – $48k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • São Paulo
AI/ML
Claude
OpenAI
DevOps
AWS
Management
Slack
Agile
Apply
$16k – $43k per year (Estimated) • In office • Full-Time • São Paulo
Management
Agile
Microsoft Office
Apply
$40k – $60k per year • Equity • In office • Full-Time • 3+ years exp • São Paulo
AI/ML
AI Agents
Apply
$30k – $45k per year • Equity • Remote • Full-Time • São Paulo
AI/ML
AI Agents
Apply
$34k – $90k per year (Estimated) • Remote • Full-Time • São Paulo
Apply
In office • Full-Time • 2+ years exp • PhD • São Paulo
AI/ML
AI Agents
Agentforce
Management
Slack
Marketing
Salesforce
Apply
See all jobs
This is one of many
726,270 more open roles from verified company boards, updated every day.