725,615open jobs
43,268companies
103,521added this week
Browse all
Salary
$78k – $175k per year (Estimated)
Location
Remote (São Paulo, Brazil)
Seniority
Staff
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 23, 2026.

Overview
Company
Impact
Profile match
A frontier AI lab building predictive models for enterprise decisions. We pre-train a Graph Foundation Model on the relational economy, then fine-tune inside your workspace — for credit, fraud, marketing, growth, and any task with a graph underneath it.

About the role

At Avra, every technical IC is a Member of Technical Staff (MTS). The title doesn't put anyone in a silo: you own systems and outcomes, not steps in a function, and you keep building depth in your area. Seniority shows up in your scope, level, and compensation, not in titles.

In this role, you'll join our ML Systems team, which owns Avra's ML core and the governance of every model we ship. Research produces candidate models and evidence; you build the reliable path from data and training to a governed, reproducible release that can run in our cloud or in any customer environment. ML Systems is an internal platform: its users are our researchers and platform engineers, and its success is measured by the leverage it creates for them.

What you'll do

  • Build CUDA kernels and compute primitives for training and serving graph neural networks (GNNs).

  • Evolve Monad, our sampler and distributed-training library, including neighbor sampling and training performance.

  • Specify our binary data formats (Lance, Arrow, CSR/CSC), and own materializations and feature backfills for training and evaluation.

  • Define data contracts and consumption requirements with the teams that build our customer and proprietary datasets.

  • Build and operate experiment tracking, checkpoints, and evaluation infrastructure, with reproducibility by default.

  • Own the model registry, lineage, versioning, and compatibility across models, embeddings, and downstream models.

  • Define and run release gates, so every model running in production, batch, or on-premise maps to a governed release.

  • Make it possible to audit exactly which data, code, configuration, and evidence produced each release.

How we measure success

  • Time-to-experiment: how quickly a researcher goes from a hypothesis to materialized data, compute, and tracking.

  • Time-to-governed-release: how quickly a validated candidate becomes an authorized release.

  • Training throughput per GPU on our foundation model training runs.

  • 100% of production models with complete release records and lineage - no ad hoc models in any environment.

  • Every release reproducible from its registered data, code, and configuration.

What we're looking for

  • Strong systems engineering skills and production-quality Python.

  • Experience with distributed training (e.g., Ray, PyTorch distributed) and multi-node GPU workloads.

  • Experience with columnar data formats and large-scale data materialization.

  • Familiarity with ML lifecycle tooling: experiment tracking, model registries, evaluation, and reproducibility.

  • A product mindset: you treat an internal platform as a product with real users. You don't need to be a data scientist.

Nice to have

  • CUDA kernel development or GPU performance optimization.

  • Graph neural networks or graph sampling at scale.

  • Lance, Arrow, or other columnar/indexed storage formats.

  • Multi-cloud GPU compute (e.g., SkyPilot).

  • Model governance or audit requirements in financial services or other regulated environments.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
725,615 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
São Paulo
AI-инженер 1 day ago
$19k – $51k per year (Estimated) • In office • Minsk
Python
SQL
Python
FastAPI
Databases
PostgreSQL
Weaviate
pgvector
Qdrant
ElasticSearch
OpenSearch
AI/ML
DVC
Weights & Biases
LangChain
LlamaIndex
LoRA
vLLM
MLFlow
XGBoost
Fine-tuning
Embeddings
Function Calling
Haystack
Ollama
PEFT
TGI
Falcon
Transformers
PyTorch
LLM
RAG
Reranking
Hugging Face
Human-in-the-Loop
Structured Outputs
Tool Use
DevOps
Rest API
gRPC
CI/CD
Git
Docker
Apply
$20k – $51k per year (Estimated) • In office • Full-Time • 5+ years exp • Malaysia
Python
SQL
Databases
Amazon Redshift
Microsoft Fabric
AI/ML
Copilot
Cursor
Claude
ChatGPT
Claude Code
AI Agents
AWS Bedrock
LLM
RAG
OpenAI
OpenAI Codex
Multi-Agent Systems
DevOps
Azure
AWS
Platform Engineering
Apply
$141k – $218k per year • Equity • In office • 5+ years exp • Bachelor's Degree
Python
Databases
PostgreSQL
Apache Kafka
AI/ML
LLM
Frontend
GraphQL
DevOps
Rest API
CI/CD
Docker
GitHub
Apply
$124k – $196k per year • Equity • In office • Bachelor's Degree
Python
Databases
Apache Kafka
DevOps
Terraform
GitHub Actions
Datadog
Azure
CI/CD
GitOps
ArgoCD
Kubernetes
Proxmox VE
Apply
Senior AI Engineer 1 day ago
$137k – $207k per year • Equity • In office • 5+ years exp • Bachelor's Degree
Python
AI/ML
Embeddings
Computer Vision
LLM
RAG
Anomaly Detection
Agentic Workflows
Apply
$80k – $179k per year (Estimated) • Remote • Full-Time • São Paulo
Python
AI/ML
Ray Serve
Fine-tuning
Ray
Post-training
DevOps
GCP
AWS
Kubernetes
Amazon EKS
Google GKE
Cybersecurity
Sophos
Apply
$77k – $172k per year (Estimated) • Remote • Full-Time • São Paulo
AI/ML
Ray
DevOps
Terraform
GCP
Helm
OpenTelemetry
AWS
Kubernetes
Amazon EKS
Google GKE
Incident Management
Apply
In office • Full-Time • 2+ years exp • São Paulo
Python
SQL
Analytics
A/B Testing
Apply
Remote/Hybrid • Full-Time • 2+ years exp • São Paulo
AI/ML
AI Agents
LLM
Apply
Operator 1 month ago
$21k – $48k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • São Paulo
AI/ML
Claude
OpenAI
DevOps
AWS
Management
Slack
Agile
Apply
$16k – $43k per year (Estimated) • In office • Full-Time • São Paulo
Management
Agile
Microsoft Office
Apply
$40k – $60k per year • Equity • In office • Full-Time • 3+ years exp • São Paulo
AI/ML
AI Agents
Apply
$30k – $45k per year • Equity • Remote • Full-Time • São Paulo
AI/ML
AI Agents
Apply
$34k – $90k per year (Estimated) • Remote • Full-Time • São Paulo
Apply
In office • Full-Time • 2+ years exp • PhD • São Paulo
AI/ML
AI Agents
Agentforce
Management
Slack
Marketing
Salesforce
Apply
See all jobs
This is one of many
725,615 more open roles from verified company boards, updated every day.