1,118,317open jobs
64,475companies
191,030added this week
Browse all
Salary
≈ $20k – $50k per year (Estimated)
Location
In office (Hanoi)
Seniority
Senior
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 2, 2026. First seen by Alion on Sep 16, 2026. VinFast scores B on the Alion truth index.

Overview
Company
Impact
Profile match
VinFast is a Vietnamese electric vehicle maker within the Vingroup conglomerate that designs and builds electric cars, e-scooters and motorbikes, buses, and charging and battery swap infrastructure, with its main manufacturing complex in Haiphong. Founded in 2017, it listed on the Nasdaq in 2023 through its Singapore-registered parent and sells vehicles in Vietnam, North America, Europe, India and Southeast Asia. It hires AI, embedded, AUTOSAR, sensor fusion and test engineers, data scientists, powertrain and quality engineers, and purchasing and logistics leaders, mostly in Hanoi and Haiphong.
VINFAST is a pioneering electric vehicle (EV) company committed to revolutionizing the automotive industry with sustainable and innovative mobility solutions. As a leading player in the EV market, VinFast is dedicated to delivering high-quality, cutting-edge electric vehicles that redefine the driving experience. Our team consists of passionate professionals driven by a shared vision of creating a greener and more sustainable future through innovation, technology, and excellence. We are looking for an Embedded AI Inference Engineer to develop, deploy, and optimize AI models on embedded SoCs and hardware accelerators. You will work across the complete inference pipeline - from AI model and computational graph to runtime, hardware accelerator, memory, and embedded application - with a strong focus on achieving product targets for latency, throughput, memory, power, and accuracy.

Key

Responsibilities Deploy CNN/DNN and vision models onto embedded AI platforms using runtimes such as Qualcomm QNN, TensorRT, TIDL, ONNX Runtime, or equivalent. Convert and optimize models from PyTorch/TensorFlow → ONNX → target inference runtime. Optimize models using FP16, INT8, quantization, PTQ/QAT, operator fusion, and graph optimization. Analyze operator compatibility, graph partitioning, accelerator mapping, and CPU fallback. Optimize AI workloads across heterogeneous CPU/GPU/DSP/NPU architectures. Profile and optimize latency, FPS, throughput, CPU/accelerator utilization, memory footprint, and memory bandwidth. Analyze and reduce data-movement overhead through DMA, shared memory, zero-copy, buffer management, and efficient memory pipelines. Optimize end-to-end pipelines such as: Camera → Pre-processing → AI Inference → Post-processing → Application Identify and resolve performance bottlenecks related to compute, memory bandwidth, synchronization, scheduling, and hardware utilization. Integrate and debug AI inference pipelines in embedded Linux environments using C/C++ and Python.

Requirements Required

Qualifications Bachelor's or Master's degree in Computer Science, Computer Engineering, Electronics, Embedded Systems, AI, or related fields. Strong programming skills in C/C++ and working knowledge of Python. Good understanding of CNN/DNN, tensors, computational graphs, neural-network operators, and AI inference. Hands-on experience with ONNX and at least one embedded AI inference runtime or accelerator SDK. Understanding of FP32/FP16/INT8, quantization, PTQ/QAT, and model optimization. Understanding of heterogeneous computing using CPU, GPU, DSP, NPU, or dedicated AI accelerators. Good knowledge of embedded system concepts including multithreading, synchronization, memory management, shared memory, and IPC. Experience with profiling, performance analysis, and bottleneck identification.

Preferred

Qualifications Experience in one or more of the following is highly desirable: Qualcomm QNN / Hexagon DSP / HTP NVIDIA CUDA / TensorRT / Jetson TI TIDL / C7x / MMA Camera and computer-vision pipelines OpenCV, OpenVX, GStreamer DMA and zero-copy architectures DDR/cache/memory-bandwidth optimization Real-time or high-performance embedded systems Automotive ADAS, robotics, or edge-AI products What We Are Looking For We are looking for an engineer who can go beyond simply “making the model run.” The ideal candidate can understand and optimize the complete path: Model → Graph → Runtime → Hardware Accelerator → Memory → Embedded Software → End-to-End Performance and systematically determine why an AI workload is slow, where the bottleneck is, and how to optimize it for production embedded systems.

Benefits Competitive salary Premium healthcare package, including PVI insurance & annual health check-ups 13th-month salary & performance bonuses to reward your contributions Enjoy preferential pricing for services within the Vingroup ecosystem including Vinmec, Vinpearl, and Vinschool... Opportunity to collaborate with and learn from industry-leading professionals in the automotive domain Work Location: Technopark Tower, Gia Lam, Ha Noi With respect to all your personal data shared to VinFast in the application and the entire recruitment process of VinFast, by clicking “Apply”, submitting your resumé/CV and/or participating in VinFast's recruitment process, you agree that you have read VinFast's Personal Data Protection Policy ("Policy") posted at https://vinfastauto.com/vn_vi/dieu-khoan-phap-ly or https://vinfast.vn/privacy-policy/, you agree to the Policy and consent for VinFast to process your personal data in accordance with the Policy and the applicable regulations on personal data protection. To all recruitment agencies: VinFast does not accept agency resumes. Please do not forward resumes to our careers alias or other VinFast employees. VinFast is not responsible for any fees related to unsolicited resumes.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,118,317 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Hanoi
In office • Full-Time • Bachelor's Degree • Ho Chi Minh City
Python
SQL
AI/ML
Prompt Engineering
AI Agents
LLM
RAG
Machine Learning
DevOps
Rest API
Azure
Git
GitHub
Analytics
Power BI
Management
Power Automate
Power Apps
Microsoft Teams
Apply
≈ $24k – $57k per year (Estimated) • In office • Full-Time • 5+ years exp • Hanoi
AI/ML
Model Context Protocol
Fine-tuning
AI Agents
LLM
RAG
A2A
DevOps
Kubernetes
Apply
≈ $55k – $157k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Vienna
Python
Go
Kotlin
AI/ML
Reinforcement Learning
Quantization
Computer Vision
ONNX
Machine Learning
Mobile
Core ML
DevOps
GCP
AWS
Docker
Amazon EC2
Apply
$200k – $400k per year • In office • Full-Time • San Francisco
AI/ML
vLLM
Cybersecurity
CVE
Apply
≈ $157k – $334k per year (Estimated) • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Seattle
Python
Go
Java
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
LangGraph
LangChain
Model Context Protocol
Multimodal AI
Computer Vision
AI Agents
TensorFlow
PyTorch
LLM Guardrails
Recommender Systems
Physical AI
Machine Learning
DevOps
AWS
Kubernetes
Apply
≈ $14k – $35k per year (Estimated) • Equity • In office • Full-Time • 5+ years exp • Manila
Python
Java
SQL
Scala
Databases
Snowflake
Amazon Redshift
Trino
AI/ML
Spark
AI Agents
LLM Evaluation
Machine Learning
DevOps
CI/CD
AWS
Analytics
ETL/ELT
Dimensional Modeling
Apply
$102k – $171k per year • In office • 5+ years exp • Bachelor's Degree • United States
Python
SQL
Databases
Oracle
Analytics
Tableau
Power BI
ETL/ELT
Apply
$14k per year • Remote (EAEU) • Moscow
Python
JavaScript
Python
Django
Frontend
Next.js
React.js
Apply
≈ $68k – $173k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • London
Python
SQL
Databases
PostgreSQL
pgvector
Pinecone
Qdrant
AI/ML
RAG
Lovable
Agentic Workflows
DevOps
Docker
GitHub
Management
n8n
Zapier
Apply
In office • Full-Time • Master's Degree • Seoul
Python
DevOps
Linux
Chips/EDA
Cadence Virtuoso
Apply
Senior AI Engineer 4 days ago
≈ $20k – $51k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Hanoi
Python
AI/ML
LoRA
vLLM
Fine-tuning
RLHF
Quantization
Knowledge Distillation
Function Calling
Computer Vision
AI Agents
NLP
Speech Recognition
ONNX
TensorRT
PEFT
TFLite
Transformers
TensorFlow
PyTorch
LLM
RAG
Hugging Face
LLMOps
DPO
SFT
Text-to-Speech
Context Engineering
LLM Guardrails
Edge AI
LiteRT
ONNX Runtime
Multi-Agent Systems
Tool Use
Model Distillation
Apply
≈ $20k – $49k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Hanoi
Python
AI/ML
vLLM
CUDA Toolkit
Triton Inference Server
Quantization
AI Agents
CUDA
DevOps
Terraform
GitHub Actions
CloudFormation
Prometheus
GitLab CI
CI/CD
ArgoCD
AWS
Kubernetes
Grafana
Blue-Green Deployment
Platform Engineering
Service Mesh
Amazon EKS
IAM
Apply
≈ $20k – $49k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Hanoi
Python
SQL
AI/ML
Weights & Biases
LangChain
MLFlow
Fine-tuning
Reinforcement Learning
Scikit-learn
Prompt Engineering
Computer Vision
AI Agents
NLP
Transfer Learning
Accelerate
Kubeflow
Transformers
TensorFlow
PyTorch
LLM
RAG
Hugging Face
Feature Store
Recommender Systems
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
AWS
Docker
Kubernetes
Robotics
Reinforcement Learning
Apply
≈ $20k – $50k per year (Estimated) • In office • Full-Time • Hanoi
Python
Bash
AI/ML
CUDA Toolkit
CUDA
DevOps
Ansible
Helm
Traefik
Envoy
Prometheus
SLURM
CI/CD
GitOps
ArgoCD
Kubernetes
Grafana
SaltStack
Platform Engineering
HPC
Linux
Apply
Speech AI Engineer 12 days ago
In office • Full-Time • 2+ years exp • Bachelor's Degree • Hanoi
Python
AI/ML
LoRA
Fine-tuning
Multimodal AI
AI Agents
Speech Recognition
PEFT
Transformers
PyTorch
LLM
Whisper
Self-Supervised Learning
Synthetic Data
SFT
Text-to-Speech
Voice Agents
Apply
≈ $24k – $57k per year (Estimated) • In office • Full-Time • 5+ years exp • Hanoi
AI/ML
Model Context Protocol
Fine-tuning
AI Agents
LLM
RAG
A2A
DevOps
Kubernetes
Apply
Hybrid • Full-Time • Hanoi
Python
Python
pySpark
AI/ML
Spark
Prompt Engineering
AI Agents
NLP
Transformers
LLM
RAG
DevOps
Azure
Apply
≈ $12k – $28k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Hanoi
Python
C++
DevOps
CI/CD
Linux
DNS
DHCP
Wi-Fi
Management
Agile
Scrum
Apply
In office • Full-Time • Hanoi
Apply
≈ $15k – $34k per year (Estimated) • In office • Full-Time • 3+ years exp • Hanoi
Python
C
C++
Bash
C
Valgrind
U-Boot
C++
CMake
DevOps
RTOS
CI/CD
Linux
TCP/IP
VLAN
Cybersecurity
Wireshark
Scapy
IoT
FreeRTOS
Management
Agile
Scrum
Apply
See all jobs
This is one of many
1,118,317 more open roles from verified company boards, updated every day.