1,059,896open jobs
62,082companies
176,030added this week
Browse all
Salary
≈ $53k – $152k per year (Estimated)
Location
Remote (Dubai, United Arab Emirates, EMEA time zones hours)
Seniority
Middle · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 1, 2026. First seen by Alion on Jul 10, 2026. Pragmatike scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Pragmatike is a technology outsourcing company operating from Vietnam that supplies software engineering teams to clients in Europe and North America. Its work covers web and mobile development, quality assurance, cloud infrastructure and data engineering, delivered as dedicated teams working inside the client's own process rather than as fixed-scope projects. Based in Hanoi, it competes on the cost and engineering supply advantages Vietnam has built over the past decade, and it advertises client vacancies under its own name while recruiting continuously across the local developer market.

Location: Fully remote (EMEA timezone)

Start date: ASAP

Languages: Fluent English required

Industry: Cloud Computing / AI / European Deep-Tech SaaS

About the Role

Pragmatike is recruiting on behalf of a fast-scaling, well-funded distributed cloud infrastructure startup building next-generation AI-native cloud services. The company is redefining how compute is delivered by providing GPU-powered infrastructure for AI/ML workloads, secure storage, and high-speed data transfer through a decentralized architecture that significantly reduces environmental impact compared to traditional cloud providers.

We are seeking a AI Infrastructure Engineer with strong experience in production-grade model serving and infrastructure for AI systems. This is a highly technical, hands-on role focused on building scalable, reliable, and efficient ML inference platforms powering real-time AI applications.

You will be responsible for designing and operating the core infrastructure that serves machine learning models at scale. You will work closely with infrastructure, platform, and applied AI teams to ensure high availability, low latency, and cost-efficient inference systems. Strong ownership, production mindset, and experience with distributed GPU systems are essential.

Your Responsibilities

  • Build and operate production-grade model serving infrastructure using frameworks such as vLLM, TGI, Triton, or equivalent

  • Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models

  • Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers

  • Optimize GPU utilization, memory efficiency, network throughput, and model artifact storage performance

  • Design observability systems for tracking inference latency, throughput, GPU usage, cost metrics, and system health

  • Manage model registries and CI/CD pipelines enabling automated and reproducible model deployments

  • Own the full lifecycle of ML systems from development through production, including operational support and on-call responsibilities

  • Define engineering best practices and contribute to platform scalability in a fast-moving startup environment

Required Qualifications

  • 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar infrastructure roles focused on ML systems

  • Hands-on experience with model serving frameworks such as vLLM, TGI, Triton, or equivalent

  • Strong background in container orchestration and operating GPU-based workloads in production

  • Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines

  • Proficiency in Python and infrastructure-as-code tools (e.g., Terraform, Helm, or similar)

  • Strong understanding of distributed systems, performance tuning, and production reliability engineering

  • Ability to effectively use AI coding assistants to accelerate development and debugging workflows

  • Ownership mindset with the ability to operate independently in a remote-first environment

Preferred Qualifications

  • Experience with ML platforms such as Kubeflow, MLflow, or KubeAI

  • Knowledge of GPU scheduling, CUDA/ROCm optimization, or multi-tenant inference systems

  • Experience with cost optimization across different GPU types and inference workloads

  • Background in early-stage startups or greenfield infrastructure projects

  • Proven experience building production systems from scratch rather than maintaining legacy platforms

Why Join Us

  • Take ownership of critical infrastructure powering a rapidly scaling AI-native cloud platform

  • Build foundational ML inference systems from the ground up in a high-growth, well-funded startup

  • Work at the intersection of distributed systems, GPU computing, and sustainable cloud architecture

  • Gain deep expertise in next-generation AI infrastructure and large-scale model serving systems

  • Influence core engineering decisions and define best practices that will scale with the company.

Pragmatike is committed to a fair, transparent, and inclusive recruitment process. We do not discriminate based on age, disability, gender, gender identity or expression, marital or civil partner status, pregnancy or maternity, race, religion or belief, sex, or sexual orientation.

In accordance with GDPR, your personal data will be processed lawfully, fairly, and securely, and used solely for recruitment purposes, including sharing it with our client(s) for employment consideration.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,059,896 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Dubai
$165k – $200k per year • Equity • Remote (United States) • Top Secret • 5+ years exp
Python
AI/ML
Fine-tuning
Embeddings
RAG
Apply
$29k – $36k per year (gross) • Remote (India) • Full-Time • 10+ years exp
Python
Databases
Apache Kafka
AI/ML
Model Context Protocol
AI Agents
DevOps
Terraform
CI/CD
GitOps
Kubernetes
Platform Engineering
Bamboo
Bitbucket
Management
Confluence
Jira
Apply
AI Engineer 3 hours ago
$135k – $168k per year • Remote (United States) • 5+ years exp
Python
JavaScript
TypeScript
SQL
PowerShell
AI/ML
Copilot
Prompt Engineering
AI Agents
RAG
Copilot Studio
Machine Learning
DevOps
Azure
GitHub
Apply
≈ $63k – $107k per year (Estimated) • Remote (Poland) • Full-Time • 2+ years exp • PhD • Warsaw
Python
SQL
Databases
Google BigQuery
BigQuery
AI/ML
CatBoost
MLFlow
XGBoost
Vertex AI
LightGBM
Kubeflow
Machine Learning
DevOps
Terraform
GCP
CI/CD
Docker
Kubernetes
SLI/SLO/SLA
Linux
Apply
Remote (United States, Philippines) • Contractor • 1+ year exp
Python
JavaScript
SQL
AI/ML
AI Agents
LLM
Human-in-the-Loop
Analytics
A/B Testing
Management
Zapier
WhatsApp
Apply
≈ $17k – $44k per year (Estimated) • In office • Full-Time • Bengaluru
Python
Java
Python
pylint
PyQt
DevOps
Azure
CI/CD
Grafana
Linux
Windows
Cybersecurity
SonarQube
Management
Agile
Apply
AI Engineer 2 days ago
≈ $53k – $127k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Barcelona
Python
Go
JavaScript
SQL
Ruby
Node JS
Databases
PostgreSQL
Apache Kafka
Google BigQuery
BigQuery
AI/ML
Spark
Embeddings
JAX
Function Calling
Flink
TensorFlow
PyTorch
LLM
RAG
Recommender Systems
Machine Learning
DevOps
Rest API
GCP
GitHub Actions
Prometheus
CI/CD
AWS
Docker
Grafana
Analytics
A/B Testing
Management
Agile
Apply
≈ $18k – $39k per year (Estimated) • In office • 4+ years exp • Pune
Python
JavaScript
Ruby
PowerShell
Node JS
DevOps
Terraform
CloudFormation
AWS
AWS Lambda
Management
Agile
Apply
≈ $31k – $64k per year (Estimated) • In office • 8+ years exp • Bengaluru
AI/ML
RLHF
Reinforcement Learning
NLP
TensorFlow
PyTorch
Recommender Systems
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
HIPAA
Apply
up to $29k per year • Remote (likely EAEU) • Full-Time • 5+ years exp
JavaScript
Java
Kotlin
SQL
Frontend
npm
Mobile
JUnit
DevOps
Rest API
Kibana
CI/CD
Git
QA
TestNG
Selenium
Playwright
Rest-Assured
Apply
≈ $28k – $74k per year (Estimated) • Remote (India) • Full-Time • 4+ years exp
Python
AI/ML
Copilot
Cursor
Claude Code
vLLM
Fine-tuning
Quantization
Scikit-learn
Knowledge Distillation
AI Agents
NLP
Transformers
PyTorch
LLM
Triton
Hugging Face
OpenAI Codex
TorchServe
ONNX Runtime
Model Distillation
DevOps
Docker
Kubernetes
Cybersecurity
GDPR
DLP
Apply
≈ $28k – $73k per year (Estimated) • Remote (India) • Full-Time • 4+ years exp
Python
Go
AI/ML
Copilot
Cursor
Claude Code
Model Context Protocol
Fine-tuning
Prompt Engineering
Function Calling
AI Agents
LLM
RAG
OpenAI Codex
Structured Outputs
Tool Use
Machine Learning
DevOps
GCP
OpenTelemetry
Azure
AWS
Docker
Kubernetes
IAM
Cybersecurity
GDPR
SIEM
Apply
$200k – $400k per year • Remote (United States) • Full-Time • 7+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Cybersecurity
GDPR
Apply
Staff AI Engineer 2 months ago
$200k – $400k per year • Remote (United States) • Full-Time • 7+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Python
AI/ML
vLLM
CUDA Toolkit
Fine-tuning
AI Agents
TGI
PyTorch
LLM
RAG
CUDA
Triton
DevOps
Kubernetes
Cybersecurity
GDPR
Apply
$200k – $400k per year • Remote (United States) • Full-Time • 5+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Python
Java
Rust
TypeScript
DevOps
Kubernetes
Cybersecurity
GDPR
Apply
In office • Full-Time • Dubai
Management
Outlook
Microsoft Office
Apply
Hybrid • Full-Time • Bachelor's Degree • Dubai
Apply
≈ $50k – $114k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Dubai
Apply
≈ $81k – $172k per year (Estimated) • In office • Full-Time • Dubai
Apply
≈ $48k – $121k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Dubai
Apply
See all jobs
This is one of many
1,059,896 more open roles from verified company boards, updated every day.