994,162open jobs
59,296companies
165,357added this week
Browse all
Salary
≈ $53k – $152k per year (Estimated)
Location
Remote (EMEA, Armenia, EMEA time zones hours)
Seniority
Middle · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 1, 2026. First seen by Alion on Jul 10, 2026. Pragmatike scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Pragmatike is a technology outsourcing company operating from Vietnam that supplies software engineering teams to clients in Europe and North America. Its work covers web and mobile development, quality assurance, cloud infrastructure and data engineering, delivered as dedicated teams working inside the client's own process rather than as fixed-scope projects. Based in Hanoi, it competes on the cost and engineering supply advantages Vietnam has built over the past decade, and it advertises client vacancies under its own name while recruiting continuously across the local developer market.

Location: Fully remote (EMEA timezone)

Start date: ASAP

Languages: Fluent English required

Industry: Cloud Computing / AI / European Deep-Tech SaaS

About the Role

Pragmatike is recruiting on behalf of a fast-scaling, well-funded distributed cloud infrastructure startup building next-generation AI-native cloud services. The company is redefining how compute is delivered by providing GPU-powered infrastructure for AI/ML workloads, secure storage, and high-speed data transfer through a decentralized architecture that significantly reduces environmental impact compared to traditional cloud providers.

We are seeking a ML Ops Engineer with strong experience in production-grade model serving and infrastructure for AI systems. This is a highly technical, hands-on role focused on building scalable, reliable, and efficient ML inference platforms powering real-time AI applications.

You will be responsible for designing and operating the core infrastructure that serves machine learning models at scale. You will work closely with infrastructure, platform, and applied AI teams to ensure high availability, low latency, and cost-efficient inference systems. Strong ownership, production mindset, and experience with distributed GPU systems are essential.

Your Responsibilities

  • Build and operate production-grade model serving infrastructure using frameworks such as vLLM, TGI, Triton, or equivalent

  • Design and implement robust deployment pipelines with blue/green and canary rollout strategies for ML models

  • Develop and maintain auto-scaling systems, multi-model serving architectures, and intelligent request routing layers

  • Optimize GPU utilization, memory efficiency, network throughput, and model artifact storage performance

  • Design observability systems for tracking inference latency, throughput, GPU usage, cost metrics, and system health

  • Manage model registries and CI/CD pipelines enabling automated and reproducible model deployments

  • Own the full lifecycle of ML systems from development through production, including operational support and on-call responsibilities

  • Define engineering best practices and contribute to platform scalability in a fast-moving startup environment

Required Qualifications

  • 4+ years of experience in ML Ops, Platform Engineering, SRE, or similar infrastructure roles focused on ML systems

  • Hands-on experience with model serving frameworks such as vLLM, TGI, Triton, or equivalent

  • Strong background in container orchestration and operating GPU-based workloads in production

  • Experience with MLOps tooling including model registries, experiment tracking, and automated deployment pipelines

  • Proficiency in Python and infrastructure-as-code tools (e.g., Terraform, Helm, or similar)

  • Strong understanding of distributed systems, performance tuning, and production reliability engineering

  • Ability to effectively use AI coding assistants to accelerate development and debugging workflows

  • Ownership mindset with the ability to operate independently in a remote-first environment

Preferred Qualifications

  • Experience with ML platforms such as Kubeflow, MLflow, or KubeAI

  • Knowledge of GPU scheduling, CUDA/ROCm optimization, or multi-tenant inference systems

  • Experience with cost optimization across different GPU types and inference workloads

  • Background in early-stage startups or greenfield infrastructure projects

  • Proven experience building production systems from scratch rather than maintaining legacy platforms

Why Join Us

  • Take ownership of critical infrastructure powering a rapidly scaling AI-native cloud platform

  • Build foundational ML inference systems from the ground up in a high-growth, well-funded startup

  • Work at the intersection of distributed systems, GPU computing, and sustainable cloud architecture

  • Gain deep expertise in next-generation AI infrastructure and large-scale model serving systems

  • Influence core engineering decisions and define best practices that will scale with the company.

Pragmatike is committed to a fair, transparent, and inclusive recruitment process. We do not discriminate based on age, disability, gender, gender identity or expression, marital or civil partner status, pregnancy or maternity, race, religion or belief, sex, or sexual orientation.

In accordance with GDPR, your personal data will be processed lawfully, fairly, and securely, and used solely for recruitment purposes, including sharing it with our client(s) for employment consideration.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
994,162 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
Senior AI Engineer 2 hours ago
≈ $45k – $112k per year (Estimated) • Remote (Europe) • Sofia
Python
SQL
AI/ML
LangChain
Hadoop
Spark
Scikit-learn
TensorFlow
PyTorch
Hugging Face
Edge AI
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Analytics
Tableau
Power BI
Apply
GenAI Engineer 3 hours ago
Remote (India) • 6+ years exp • India
Python
Python
Flask
FastAPI
AI/ML
AutoGen
LangChain
LlamaIndex
Vertex AI
Embeddings
Prompt Engineering
Function Calling
AI Agents
Semantic Kernel
CrewAI
Gemini
LLM
RAG
Google ADK
Hallucination
LLMOps
Human-in-the-Loop
Structured Outputs
Knowledge Graph
LLM Guardrails
Multi-Agent Systems
DevOps
Rest API
GCP
Kubernetes
Google GKE
Google Cloud Run
Management
Agile
Apply
Remote (India) • Full-Time • 6+ years exp
AI/ML
AI Agents
LLM
RAG
OpenAI
LLMOps
Copilot Studio
Machine Learning
DevOps
Azure
Management
Power Automate
Microsoft Teams
SharePoint
Apply
≈ $168k – $299k per year (Estimated) • Remote (United States) • Full-Time • 4+ years exp • High School Diploma • Thousand Oaks
Python
AI/ML
Embeddings
Function Calling
RAG
Hallucination
Agentic Workflows
Tool Use
Machine Learning
DevOps
CI/CD
AWS
Apply
$200k – $400k per year • Remote (LATAM) • Contractor
Python
JavaScript
TypeScript
AI/ML
Cursor
Claude Code
OpenAI Codex
Apply
≈ $15k – $36k per year (Estimated) • Remote (India) • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
Apex
Bash
COBOL
Apex
MuleSoft
COBOL
IBM MQ
DevOps
Rest API
Terraform
Ansible
GCP
OpenShift
Azure DevOps
GitHub Actions
GitLab CI
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Platform Engineering
Incident Management
Linux
SOAP
Management
Agile
Apply
Backend Engineer 6 days ago
≈ $10k – $30k per year (Estimated) • In office • 2+ years exp • Bengaluru
Python
Databases
MySQL
PostgreSQL
DevOps
Rest API
CI/CD
Docker
Apply
≈ $14k – $37k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Tashkent
Python
SQL
Python
SQLAlchemy
pySpark
Databases
PostgreSQL
Snowflake
ClickHouse
Apache Kafka
Greenplum
AI/ML
Spark
Airflow
Dagster
dbt
Great Expectations
DevOps
GitLab CI
CI/CD
Git
Docker
Analytics
Power BI
ETL/ELT
Superset
Data Vault
Apply
≈ $19k – $47k per year (Estimated) • In office • Bengaluru
Python
JavaScript
TypeScript
SQL
AI/ML
LLM
DevOps
Rest API
GitHub Actions
GitLab CI
CI/CD
Jenkins
Docker
Kubernetes
QA
Selenium
JMeter
Cypress
Playwright
Gatling
Postman
Rest-Assured
Pact
k6
Locust
Apply
≈ $12k – $29k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Tashkent
Python
SQL
Databases
PostgreSQL
AI/ML
Pandas
Analytics
ETL/ELT
Superset
Apply
≈ $28k – $74k per year (Estimated) • Remote (India) • Full-Time • 4+ years exp
Python
AI/ML
Copilot
Cursor
Claude Code
vLLM
Fine-tuning
Quantization
Scikit-learn
Knowledge Distillation
AI Agents
NLP
Transformers
PyTorch
LLM
Triton
Hugging Face
OpenAI Codex
TorchServe
ONNX Runtime
Model Distillation
DevOps
Docker
Kubernetes
Cybersecurity
GDPR
DLP
Apply
≈ $28k – $73k per year (Estimated) • Remote (India) • Full-Time • 4+ years exp
Python
Go
AI/ML
Copilot
Cursor
Claude Code
Model Context Protocol
Fine-tuning
Prompt Engineering
Function Calling
AI Agents
LLM
RAG
OpenAI Codex
Structured Outputs
Tool Use
Machine Learning
DevOps
GCP
OpenTelemetry
Azure
AWS
Docker
Kubernetes
IAM
Cybersecurity
GDPR
SIEM
Apply
$200k – $400k per year • Remote (United States) • Full-Time • 7+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Cybersecurity
GDPR
Apply
Staff AI Engineer 2 months ago
$200k – $400k per year • Remote (United States) • Full-Time • 7+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Python
AI/ML
vLLM
CUDA Toolkit
Fine-tuning
AI Agents
TGI
PyTorch
LLM
RAG
CUDA
Triton
DevOps
Kubernetes
Cybersecurity
GDPR
Apply
$200k – $400k per year • Remote (United States) • Full-Time • 5+ years exp • San Francisco • New York • Los Angeles • Sacramento • Seattle
Python
Java
Rust
TypeScript
DevOps
Kubernetes
Cybersecurity
GDPR
Apply
See all jobs
This is one of many
994,162 more open roles from verified company boards, updated every day.