624,530open jobs
30,918companies
85,628added this week
Browse all
Salary
$101k – $232k per year (Estimated)
Location
In office (Singapore)
Seniority
Architect
Overview
Company
Impact
Profile match
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.

About FuriosaAI

FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. 

Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence  for every enterprise.

About The Role

The Solution Architect will be a critical technical and strategic contributor responsible for both hands-on architecture and engineering support, as well as driving customer engagements and adoption of our advanced Tensor Contraction Processor technologies. You will leverage your extensive experience with AI accelerators, GPGPUs and server LLMs to develop robust enterprise and cloud solutions while building strong, lasting relationships with our customers.

Key Responsibilities

  • Technical Leadership & Architecture:

    • Develop and oversee end-to-end technical architectures for enterprise and cloud-based AI solutions.

    • Provide technical expertise and hands-on guidance on integrating FuriosaAI AI accelerators, and server LLMs into customer environments.

    • Collaborate with engineering and product teams to influence product roadmap and feature enhancements based on customer feedback and market trends.

  • Customer Engagement & Success:

    • Serve as a trusted advisor to enterprise and cloud customers, translating technical challenges into actionable solutions.

    • Lead pre-sales activities including technical presentations, proof-of-concepts, and tailored solution demonstrations.

    • Engage directly with key customers to understand their needs, offer technical guidance, and ensure seamless integration and deployment of our solutions.

  • Cross-Functional Collaboration:

    • Partner with Sales, Marketing, and Product teams to develop go-to-market strategies, technical collateral, and customer success stories.

    • Mentor and lead a team of solution architects and technical experts, fostering continuous learning and collaboration.

  • Market & Industry Leadership:

    • Stay abreast of emerging trends in AI hardware, enterprise cloud computing, and related technologies.

    • Represent FuriosaAI at industry events, conferences, and customer forums, building the brand as a leader in advanced AI acceleration.

Qualifications & Experience

  • Technical Expertise:

    • Proven hands-on experience with GPGPUs, AI Accelerators, and server LLMs in enterprise or cloud environments.

    • Deep understanding of modern data center architectures, distributed systems, and high-performance computing.

    • Proficiency in relevant programming languages and frameworks with experience in deploying AI models at scale.

    • State-of-the-Art AI Performance Benchmarking: Demonstrated expertise using industry-standard benchmarking tools (e.g., MLPerf, vLLM Benchmarking) to analyze and optimize AI model performance across heterogeneous compute architectures.

  • Customer-Facing Experience:

    • Demonstrated experience in pre-sales, consulting, or technical leadership roles with direct customer engagements.

    • Strong communication and presentation skills, with the ability to articulate complex technical concepts to diverse audiences.

  • Leadership & Strategic Vision:

    • Previous leadership experience managing technical teams or leading customer initiatives.

    • Ability to think strategically about market trends, customer needs, and competitive landscapes, driving innovation and product evolution.

  • Education:

    • Bachelor’s or Master’s degree in computer science, Electrical Engineering, or a related field. Advanced certifications or equivalent industry experience is a plus.

  • Language Requirements:

    • Must be fluent in oral and written English.

    • Proficiency in Korean is a plus.

What We Offer

  • Opportunity to work with cutting-edge AI technologies in a dynamic and innovative environment.

  • Competitive compensation and benefits package.

  • A collaborative culture that values creative problem-solving, continuous learning, and professional growth.

  • The chance to shape the future of AI acceleration for enterprise and cloud markets.

Contact

Why Join FuriosaAI

The defining bottleneck of the AI era is building the right hardware and software stack to run it at global scale. Furiosa is solving this challenge holistically from the ground up.

With our flagship chip, RNGD, in mass production today and our next-generation platform in development with Broadcom, we are proving that full-stack, tensor-native compute is the future of AI infrastructure. This is a pivotal moment to join our team, right as we accelerate our global expansion.

At Furiosa, you will:

Solve AI’s Most Urgent Challenge. Help build the high-performance, energy-efficient inference hardware and software required to fulfill the promise of advanced AI.

Pioneer Full-Stack Co-Design. Work with teams that are architecting solutions from silicon up through the compiler (featuring innovations like Tensor Contraction Language and Virtual ISA) and serving frameworks.

Ship Real-World Silicon, Software, and Solutions. Turn breakthrough technology into commercial deployment. RNGD is in mass production with TSMC and running live enterprise workloads for global leaders like LG AI Research and Samsung SDS.

Partner With the Industry's Best. Collaborate across an elite global ecosystem that includes TSMC, Broadcom, SK Hynix, and GUC.

Do Your Life’s Best Work. Join a brilliant, low-ego, mission-driven team in a high-trust environment that values autonomy, intellectual curiosity, and shared ambition. 

Contact

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
624,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Singapore
$160k – $300k per year (Estimated) • Remote • Full-Time • New York
AI/ML
vLLM
JAX
SGLang
TensorRT-LLM
TGI
Mistral
TensorFlow
PyTorch
LLM
InfiniBand
DevOps
SLURM
Kubernetes
HPC
Apply
$126k – $225k per year (Estimated) • Equity • Remote/Hybrid • 25+ years exp • PhD • Toronto
Python
Java
Python
Flask
FastAPI
Poetry
Gunicorn
Uvicorn
Java
Maven
Databases
PostgreSQL
Redis
Snowflake
Delta Lake
Qdrant
ElasticSearch
Apache Kafka
AI/ML
LangChain
CatBoost
Spark
Model Context Protocol
vLLM
XGBoost
Quantization
Multimodal AI
AI Agents
NLP
LangSmith
TensorRT-LLM
TGI
PyTorch
LLM
RAG
Amazon SageMaker
DevOps
GCP
Helm
Prometheus
Azure
CI/CD
AWS
Docker
Kubernetes
Grafana
TeamCity
Amazon S3
Analytics
ETL/ELT
A/B Testing
Management
Slack
Apply
$236k – $330k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Bellevue
Databases
Snowflake
AI/ML
DeepSpeed
vLLM
CUDA Toolkit
Quantization
AI Agents
SGLang
TensorRT-LLM
TensorFlow
LLM
Mixture of Experts
CUDA
Triton
Post-training
cuDNN
CUTLASS
Speculative Decoding
KV Cache
Apply
$243k – $297k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Jose
AI/ML
vLLM
CUDA Toolkit
AI Agents
SGLang
PyTorch
CUDA
ROCm
MLIR
DevOps
Prometheus
Kubernetes
Grafana
Apply
$152k – $242k per year • In office • Full-Time • 2+ years exp • Master's Degree • Santa Clara • Seattle
Python
Go
Rust
Bash
AI/ML
vLLM
SGLang
TensorRT-LLM
DevOps
SLURM
Kubernetes
Platform Engineering
HPC
Apply
$36k – $93k per year (Estimated) • In office • Seoul
Rust
C++
Apply
$35k – $102k per year (Estimated) • Remote/Hybrid • Hwaseong
Rust
C++
Apply
$35k – $102k per year (Estimated) • In office • Seoul
C
C++
C
Embedded C
DevOps
RTOS
HPC
Apply
$35k – $84k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Seoul
Python
Rust
AI/ML
CUDA Toolkit
Quantization
PyTorch
LLM
CUDA
Triton
Hugging Face
KV Cache
DevOps
Git
Apply
$33k – $95k per year (Estimated) • In office • Bachelor's Degree • Seoul
Python
Rust
DevOps
Kubernetes
Apply
$68k – $154k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Singapore
Apply
GTM Engineer 1 day ago
$100k – $150k per year • In office • Full-Time • 2+ years exp • Singapore
AI/ML
Reinforcement Learning
Post-training
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • Singapore
Python
AI/ML
Reinforcement Learning
DevOps
Docker
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • Singapore
Python
AI/ML
Reinforcement Learning
AI Agents
DevOps
Docker
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • Singapore
Python
AI/ML
Reinforcement Learning
LLM
Post-training
DevOps
Docker
Apply
See all jobs
This is one of many
624,530 more open roles from verified company boards, updated every day.