380,086open jobs
9,917companies
49,708added this week
Browse all
Salary
$109k – $263k per year (Estimated)
Location
In office (Singapore)
Seniority
Architect
Overview
Company
Impact
Profile match
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.

About FuriosaAI

FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. 

Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence  for every enterprise.

About The Role

The Solution Architect will be a critical technical and strategic contributor responsible for both hands-on architecture and engineering support, as well as driving customer engagements and adoption of our advanced Tensor Contraction Processor technologies. You will leverage your extensive experience with AI accelerators, GPGPUs and server LLMs to develop robust enterprise and cloud solutions while building strong, lasting relationships with our customers.

Key Responsibilities

  • Technical Leadership & Architecture:

    • Develop and oversee end-to-end technical architectures for enterprise and cloud-based AI solutions.

    • Provide technical expertise and hands-on guidance on integrating FuriosaAI AI accelerators, and server LLMs into customer environments.

    • Collaborate with engineering and product teams to influence product roadmap and feature enhancements based on customer feedback and market trends.

  • Customer Engagement & Success:

    • Serve as a trusted advisor to enterprise and cloud customers, translating technical challenges into actionable solutions.

    • Lead pre-sales activities including technical presentations, proof-of-concepts, and tailored solution demonstrations.

    • Engage directly with key customers to understand their needs, offer technical guidance, and ensure seamless integration and deployment of our solutions.

  • Cross-Functional Collaboration:

    • Partner with Sales, Marketing, and Product teams to develop go-to-market strategies, technical collateral, and customer success stories.

    • Mentor and lead a team of solution architects and technical experts, fostering continuous learning and collaboration.

  • Market & Industry Leadership:

    • Stay abreast of emerging trends in AI hardware, enterprise cloud computing, and related technologies.

    • Represent FuriosaAI at industry events, conferences, and customer forums, building the brand as a leader in advanced AI acceleration.

Qualifications & Experience

  • Technical Expertise:

    • Proven hands-on experience with GPGPUs, AI Accelerators, and server LLMs in enterprise or cloud environments.

    • Deep understanding of modern data center architectures, distributed systems, and high-performance computing.

    • Proficiency in relevant programming languages and frameworks with experience in deploying AI models at scale.

    • State-of-the-Art AI Performance Benchmarking: Demonstrated expertise using industry-standard benchmarking tools (e.g., MLPerf, vLLM Benchmarking) to analyze and optimize AI model performance across heterogeneous compute architectures.

  • Customer-Facing Experience:

    • Demonstrated experience in pre-sales, consulting, or technical leadership roles with direct customer engagements.

    • Strong communication and presentation skills, with the ability to articulate complex technical concepts to diverse audiences.

  • Leadership & Strategic Vision:

    • Previous leadership experience managing technical teams or leading customer initiatives.

    • Ability to think strategically about market trends, customer needs, and competitive landscapes, driving innovation and product evolution.

  • Education:

    • Bachelor’s or Master’s degree in computer science, Electrical Engineering, or a related field. Advanced certifications or equivalent industry experience is a plus.

  • Language Requirements:

    • Must be fluent in oral and written English.

    • Proficiency in Korean is a plus.

What We Offer

  • Opportunity to work with cutting-edge AI technologies in a dynamic and innovative environment.

  • Competitive compensation and benefits package.

  • A collaborative culture that values creative problem-solving, continuous learning, and professional growth.

  • The chance to shape the future of AI acceleration for enterprise and cloud markets.

Contact

Why Join FuriosaAI

The defining bottleneck of the AI era is building the right hardware and software stack to run it at global scale. Furiosa is solving this challenge holistically from the ground up.

With our flagship chip, RNGD, in mass production today and our next-generation platform in development with Broadcom, we are proving that full-stack, tensor-native compute is the future of AI infrastructure. This is a pivotal moment to join our team, right as we accelerate our global expansion.

At Furiosa, you will:

Solve AI’s Most Urgent Challenge. Help build the high-performance, energy-efficient inference hardware and software required to fulfill the promise of advanced AI.

Pioneer Full-Stack Co-Design. Work with teams that are architecting solutions from silicon up through the compiler (featuring innovations like Tensor Contraction Language and Virtual ISA) and serving frameworks.

Ship Real-World Silicon, Software, and Solutions. Turn breakthrough technology into commercial deployment. RNGD is in mass production with TSMC and running live enterprise workloads for global leaders like LG AI Research and Samsung SDS.

Partner With the Industry's Best. Collaborate across an elite global ecosystem that includes TSMC, Broadcom, SK Hynix, and GUC.

Do Your Life’s Best Work. Join a brilliant, low-ego, mission-driven team in a high-trust environment that values autonomy, intellectual curiosity, and shared ambition. 

Contact

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
380,086 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Singapore
$113k – $257k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • McLean
AI/ML
A2A
AI Agents
AutoGen
CrewAI
Cursor
Function Calling
Knowledge Graph
LangChain
LangGraph
llama.cpp
LlamaIndex
LLM
LLM Evaluation
Model Context Protocol
Pydantic AI
PyTorch
RAG
TensorFlow
vLLM
Windsurf
DevOps
AWS
Azure
Apply
$162k – $354k per year (Estimated) • Equity • In office • Full-Time • San Francisco
AI/ML
Diffusion Models
DPO
Fine-tuning
FSDP
GRPO
Knowledge Distillation
LLM
Post-training
PPO
PyTorch
Reinforcement Learning
SFT
SGLang
vLLM
VLM
Robotics
Reinforcement Learning
Apply
$26k – $110k per year (Estimated) • In office • Mumbai
Java
Node JS
Python
JavaScript
Python
FastAPI
Databases
pgvector
Qdrant
PostgreSQL
AI/ML
AI Agents
ComfyUI
CrewAI
Embeddings
LangChain
LangGraph
LLM
Prompt Engineering
RAG
Replicate
vLLM
DevOps
AWS
Docker
Git
Rest API
Vector
Apply
$22k – $54k per year (Estimated) • In office • Astana
Python
Python
FastAPI
Databases
Apache Kafka
AI/ML
AI Agents
Computer Vision
CUDA
CUDA Toolkit
Function Calling
LangGraph
LLaVA
LLM
MLFlow
Multimodal AI
ONNX
OpenCV
OpenVINO
PyTorch
Quantization
Qwen
RAG
TensorRT
Transformers
Triton
Triton Inference Server
vLLM
VLM
YOLO
LangChain
DevOps
Docker
gRPC
Kubernetes
Robotics
GStreamer
Apply
AI Engineer 1 day ago
$53k – $91k per year (Estimated) • Remote/Hybrid • Full-Time • Warsaw
Python
SQL
Databases
pgvector
Qdrant
PostgreSQL
AI/ML
Embeddings
Fine-tuning
LangChain
LlamaIndex
LLM
MLFlow
Prompt Engineering
PyTorch
RAG
Scikit-learn
TensorFlow
Triton
vLLM
DevOps
AWS
CI/CD
Docker
Kubernetes
Marketing
Salesforce
Apply
In office • Seoul
Apply
In office • Contractor • Bachelor's Degree • Seoul
AI/ML
Mamba
Multimodal AI
TPU
Apply
In office • Bachelor's Degree • Seoul
Apply
In office • Bachelor's Degree • Seoul
C++
Python
Rust
C++
LLVM
PyTorch C++
TensorFlow C++
AI/ML
LLM
Multimodal AI
ONNX
PyTorch
TensorFlow
TPU
Triton
Apply
In office • Bachelor's Degree • Seoul
AI/ML
Multimodal AI
RAG
AI Agents
DevOps
GitHub
Analytics
A/B Testing
Apply
In office • Full-Time • Master's Degree • Singapore
C++
Python
AI/ML
Multimodal AI
Reinforcement Learning
Robotics
MoveIt
Perception
Reinforcement Learning
ROS
Apply
In office • Full-Time • Bachelor's Degree • Singapore
AI/ML
Automatic1111
ComfyUI
Gradio
LocalAI
Stable Diffusion
Apply
$55k – $121k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Singapore
Apply
$75k – $129k per year (Estimated) • In office • Contractor • Singapore
AI/ML
ChatGPT
Copilot
Apply
$60k – $153k per year (Estimated) • In office • Full-Time • 1+ year exp • Singapore
Apply
See all jobs
This is one of many
380,086 more open roles from verified company boards, updated every day.