Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs

Avomind

Avomind is a platform connecting graduates and senior alumni from leading academic institutions with fast-growing firms. Our mission is to connect high-caliber candidates with impactful opportunities.... Our global network is built up of partnersh...

About the Company

Our client is a stealth AI startup backed by one of Southeast Asia's leading technology companies and is currently building its global founding team.

The company is developing an AI-native communication platform designed to simplify everyday tasks by integrating AI directly into conversations. Instead of switching between multiple applications, users can plan, organize, compare, research, and complete tasks within a single intelligent assistant.

Serving a market of billions of users still relying on traditional productivity tools, the platform focuses on delivering reliable AI workflows, persistent context, multi-step reasoning, and seamless task execution. The mission is to create an AI assistant that significantly improves productivity while making everyday work simpler and more intuitive.

About the Role

Our client is seeking a Technical Lead, Machine Learning to lead the execution of its AI platform by translating research into scalable, production-ready machine learning systems. This role sits at the intersection of research, infrastructure, and product, with responsibility for ensuring models are trainable, deployable, observable, and optimized for real-world performance.

Working closely with research, engineering, and product teams, this position will drive the development of robust ML infrastructure while balancing performance, reliability, latency, and cost.

Key Responsibilities

  • Lead the end-to-end execution of machine learning systems, including data pipelines, training workflows, evaluation frameworks, inference architecture, and production deployment.
  • Fine-tune and optimize models using modern techniques such as LoRA, QLoRA, Supervised Fine-Tuning (SFT), Direct Preference Optimization (DPO), and model distillation.
  • Design, build, and operate scalable inference systems with a focus on latency, cost efficiency, and reliability.
  • Develop and maintain data pipelines for both synthetic and real-world training datasets.
  • Build evaluation frameworks to measure model performance, robustness, safety, and bias in collaboration with research teams.
  • Optimize production deployments through GPU utilization, memory efficiency, inference optimization, and scaling strategies.
  • Partner closely with application engineering teams to integrate machine learning systems into backend, desktop, and mobile products.
  • Continuously improve production systems through rapid iteration, monitoring, and data-driven optimization.

Requirements

  • Proven experience building and deploying production-grade machine learning systems used by real users.
  • Strong expertise working with large language models and understanding model behavior, limitations, and failure modes.
  • Experience developing scalable ML infrastructure, training pipelines, and inference systems.
  • Strong software engineering skills with the ability to write maintainable, production-quality code.
  • Experience balancing real-world production constraints, including latency, reliability, scalability, cost, and safety.
  • Strong ownership mindset with the ability to independently drive technical initiatives from design through deployment.
  • Excellent communication and collaboration skills, with experience working in cross-functional, high-performing engineering teams.

Preferred Technical Skills

Experience with the following technologies is preferred:

  • Python
  • PyTorch and/or JAX
  • GPU-based model training and inference systems

Recommended for you based on this role

Similar stack
Same company
In your city
Equity • Full-Time • Bachelor's Degree
C++
Java
Python
C++
PyTorch C++
Databases
PostgreSQL
TimescaleDB
AI/ML
LightGBM
NumPy
Pandas
Polars
PyTorch
XGBoost
DevOps
CI/CD
Docker
Git
Platform Engineering
Apply
Remote/Hybrid • Master's Degree
Node JS
Python
SQL
JavaScript
AI/ML
AI Agents
PyTorch
RAG
Frontend
Next.js
React.js
DevOps
Docker
Kubernetes
Apply
Full-Time • Singapore
Python
AI/ML
JAX
Knowledge Distillation
Multimodal AI
PyTorch
RAG
Apply
Full-Time • Singapore
Python
AI/ML
JAX
PyTorch
Apply
Full-Time • Singapore
AI/ML
DeepSpeed
Diffusion Models
Fine-tuning
JAX
Knowledge Distillation
LLM
LoRA
Multimodal AI
PyTorch
QLoRA
Quantization
Ray
Reinforcement Learning
RLHF
Spark
TensorRT
TensorRT-LLM
vLLM
PEFT
Apply
Full-Time • Singapore
SQL
Swift
AI/ML
TensorFlow
Mobile
Core ML
SPM
SwiftUI
Apply
Full-Time • Singapore
Java
Kotlin
SQL
Kotlin
Kotlin Coroutines
AI/ML
Computer Vision
TensorFlow
Mobile
ML Kit
DevOps
gRPC
Apply
Full-Time • Singapore
Node JS
Python
SQL
JavaScript
AI/ML
Embeddings
Multimodal AI
PyTorch
DevOps
Docker
Kubernetes
Apply
In office • Full-Time • Bachelor's Degree • Guangzhou
JavaScript
SQL
Apply
Senior QA Engineer 2 months ago
In office • Full-Time • Bachelor's Degree • Dubai
SQL
Databases
Apache Kafka
DevOps
CI/CD
Kubernetes
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Create free account
Free forever • Less than a minute • No credit card

Work setup

Location
Singapore
Employment
Full-Time