368,530open jobs
9,432companies
50,439added this week
Browse all
Location
In office (Seoul)
Employment
Contractor
Overview
Company
Impact
Profile match
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.

About FuriosaAI

FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. 

Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence  for every enterprise.

 

About the Role

Lead the integration of diverse AI models including VLA, Vision, and Multimodal architectures by utilizing our kernel programming language to ensure both accuracy and performance while keeping the stack ready for developers to use.

Key Responsibilities

  • Design and implement efficient kernels on FuriosaAI’s kernel programming stack (including vISA, TCL), targeting Tensor Contract Processor (TCP) architectures.

  • Diagnose and optimize kernel performance with profiling tools and roofline analysis for each RNGD-accelerated AI model.

  • Develop and apply automated kernel generation and optimization for AI workloads.

  • Build diagnostic tools or testbeds for robust and reliable kernel validation.

  • Drive end-to-end programming enablement on RNGDs, creating reproducible guides and reference implementations.

Minimum Qualifications

  • BS in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.

  • Experience in low-level systems programming targeting XPU (e.g., NPU, GPU) architectures.

  • Experience collaborating across engineering, research, and product teams to align software development with product requirements.

 

Preferred Qualifications

  • MS or PhD in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.

  • Experience in optimizing high-performance kernels on AI accelerators (e.g., GPU, TPU) for AI products.

  • Understanding of XPU architecture (computation patterns, data movement) and software-hardware co-optimization strategies.

  • Experience in open-source or research projects on AI model architectures such as Diffusion, Mamba, and VLA.

  • Experience in designing efficient deep learning architectures and developing algorithms for AI applications.

Why Join FuriosaAI

The defining bottleneck of the AI era is building the right hardware and software stack to run it at global scale. Furiosa is solving this challenge holistically from the ground up.

With our flagship chip, RNGD, in mass production today and our next-generation platform in development with Broadcom, we are proving that full-stack, tensor-native compute is the future of AI infrastructure. This is a pivotal moment to join our team, right as we accelerate our global expansion.

At Furiosa, you will:

Solve AI’s Most Urgent Challenge. Help build the high-performance, energy-efficient inference hardware and software required to fulfill the promise of advanced AI.

Pioneer Full-Stack Co-Design. Work with teams that are architecting solutions from silicon up through the compiler (featuring innovations like Tensor Contraction Language and Virtual ISA) and serving frameworks.

Ship Real-World Silicon, Software, and Solutions. Turn breakthrough technology into commercial deployment. RNGD is in mass production with TSMC and running live enterprise workloads for global leaders like LG AI Research and Samsung SDS.

Partner With the Industry's Best. Collaborate across an elite global ecosystem that includes TSMC, Broadcom, SK Hynix, and GUC.

Do Your Life’s Best Work. Join a brilliant, low-ego, mission-driven team in a high-trust environment that values autonomy, intellectual curiosity, and shared ambition. 

Contact

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Seoul
$37k – $110k per year (Estimated) • Equity • Remote/Hybrid • Internship • 1+ year exp • Bachelor's Degree • Madrid
Python
SQL
Databases
Databricks
AI/ML
AI Agents
Context Engineering
Edge AI
Fine-tuning
Function Calling
LangChain
LlamaIndex
LLM
Multimodal AI
OpenAI
Prompt Engineering
PyTorch
RAG
Robotics
Digital Twin
Apply
In office • 7+ years exp • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$72k – $201k per year (Estimated) • In office • Full-Time • 3+ years exp • PhD • Basel
Python
AI/ML
Multimodal AI
OpenCV
PyTorch
scikit-image
Scikit-learn
TensorFlow
AI Agents
DevOps
Git
Apply
AI Engineer 1 day ago
$25k – $103k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Gurgaon
Python
SQL
Databases
Databricks
Microsoft Fabric
AI/ML
AI Agents
Embeddings
Gemini
Hallucination
LangChain
LangGraph
LLM
Multimodal AI
Prompt Engineering
PyTorch
RAG
Semantic Search
Spark
TensorFlow
Hugging Face
LLM Guardrails
LLMOps
OpenAI
Semantic Search
DevOps
AWS
Azure
CI/CD
Apply
In office • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
In office • Bachelor's Degree • Seoul
AI/ML
Multimodal AI
RAG
AI Agents
DevOps
GitHub
Analytics
A/B Testing
Apply
In office • 3+ years exp • Bachelor's Degree • Seoul
Apply
In office • Seoul
C++
Rust
Apply
Remote/Hybrid • Hwaseong
C++
Rust
Apply
In office • Seoul
C++
C
C
Embedded C
DevOps
RTOS
HPC
Apply
Product Manager 1 day ago
In office • Full-Time • 5+ years exp • Bachelor's Degree • Seoul
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Seoul
Python
DevOps
AWS
CI/CD
GCP
Marketing
Salesforce
Apply
In office • Full-Time • Seoul
C++
SystemVerilog
Verilog
Apply
In office • Bachelor's Degree • Seoul
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.