904,400open jobs
55,732companies
149,454added this week
Browse all
Location
In office (Seoul)

Confirmed on the employer's own hiring board on Sep 28, 2026. First seen by Alion on Sep 18, 2026.

Overview
Company
Impact
Profile match
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.

About FuriosaAI

FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. 

Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence  for every enterprise.

About the Job

The compiler plays a central role in FuriosaAI's mission to build high-performance, energy-efficient AI systems. Modern deep learning models are evolving rapidly and becoming increasingly diverse, making compilation a challenging problem. Transforming these models into efficient executable programs requires careful reasoning about complex transformations while preserving program meaning and structure.

In this role, you will own the performance of critical AI kernels integrated into Furiosa-LLM, FuriosaAI's software stack for large language model serving. You will analyze end-to-end serving workloads, identify performance-critical bottlenecks, and implement highly optimized kernels using TCL (Tensor Contraction Language), FuriosaAI's programming language for kernel optimization. Because real-world serving workloads are inherently dynamic, you will develop scheduling, specialization, and algorithmic techniques that map them efficiently onto compiler abstractions and execution environments optimized for static workloads.

Responsibilities

  • Analyze end-to-end LLM serving workloads and identify performance bottlenecks that can be addressed through kernel-level optimization.
  • Design, implement, and optimize high-performance kernels in TCL for critical operations in Furiosa-LLM.
  • Develop algorithmic techniques for efficiently supporting dynamic serving workloads.
  • Integrate, benchmark, and validate optimized kernels across representative models, input shapes, and serving scenarios.
  • Collaborate with compiler and serving teams to improve compiler capabilities and ensure optimized kernels work effectively within the production software stack.

Minimum Qualifications

  • BS in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.
  • Experience in developing low-level or performance-critical software.
  • Experience in analyzing performance bottlenecks using profiling, benchmarking, and hardware performance characteristics.
  • Understanding of parallel computation, memory hierarchies, and data movement on modern architectures.

Preferred Qualifications

  • MS or PhD in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.
  • Experience in optimizing high-performance kernels on AI accelerators (e.g., GPU, TPU) for AI products.
  • Hands-on knowledge of kernel optimization techniques such as tiling, computation scheduling, operator fusion, memory layout transformation, vectorization, pipelining, and data movement optimization.
  • Experience reasoning about trade-offs among parallelism, memory bandwidth, on-chip memory capacity, compute utilization, and synchronization overhead.
  • Experience with accelerator programming or domain-specific kernel languages such as Triton, cuTile, or Pallas.

Why Join FuriosaAI

The defining bottleneck of the AI era is building the right hardware and software stack to run it at global scale. Furiosa is solving this challenge holistically from the ground up.

With our flagship chip, RNGD, in mass production today and our next-generation platform in development with Broadcom, we are proving that full-stack, tensor-native compute is the future of AI infrastructure. This is a pivotal moment to join our team, right as we accelerate our global expansion.

At Furiosa, you will:

Solve AI’s Most Urgent Challenge. Help build the high-performance, energy-efficient inference hardware and software required to fulfill the promise of advanced AI.

Pioneer Full-Stack Co-Design. Work with teams that are architecting solutions from silicon up through the compiler (featuring innovations like Tensor Contraction Language and Virtual ISA) and serving frameworks.

Ship Real-World Silicon, Software, and Solutions. Turn breakthrough technology into commercial deployment. RNGD is in mass production with TSMC and running live enterprise workloads for global leaders like LG AI Research and Samsung SDS.

Partner With the Industry's Best. Collaborate across an elite global ecosystem that includes TSMC, Broadcom, SK Hynix, and GUC.

Do Your Life’s Best Work. Join a brilliant, low-ego, mission-driven team in a high-trust environment that values autonomy, intellectual curiosity, and shared ambition. 

Contact

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
904,400 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Seoul
In office • Contractor • 3+ years exp • Seoul
JavaScript
PHP
AI/ML
AI Agents
Frontend
JQuery
DevOps
AWS
Linux
Apply
$99k – $142k per year • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Boston
SQL
Databases
Snowflake
AI/ML
dbt
DevOps
GCP
Azure
CI/CD
AWS
Analytics
ETL/ELT
Management
Agile
Apply
≈ $106k – $195k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • PhD • United States
JavaScript
Frontend
React.js
DevOps
CI/CD
Twelve-Factor App
Management
Agile
Apply
≈ $32k – $78k per year (Estimated) • Remote (Romania) • Full-Time • 5+ years exp
JavaScript
TypeScript
Node JS
Databases
PostgreSQL
pgvector
AI/ML
LangGraph
LangChain
LLM
RAG
OpenAI
Frontend
Angular
DevOps
Azure
CI/CD
ArgoCD
Docker
Kubernetes
Apply
$173k – $235k per year • Hybrid • Full-Time • 3+ years exp • San Francisco
Python
JavaScript
Databases
PostgreSQL
AI/ML
Cursor
Claude Code
AI Agents
Frontend
React.js
Robotics
Digital Twin
Apply
Software Engineer 1 day ago
≈ $107k – $198k per year (Estimated) • Equity • Remote (United States) • 3+ years exp
Python
TypeScript
SQL
AI/ML
Cursor
Claude
Claude Code
AI Agents
LLM
Apply
Founding Engineer 1 day ago
$150k – $220k per year • Remote (United States) • Full-Time • 4+ years exp • New York
Python
AI/ML
Fine-tuning
Reinforcement Learning
AI Agents
LLM
Robotics
Reinforcement Learning
Apply
≈ $37k – $63k per year (Estimated) • In office • Internship • Chicago
AI/ML
LLM
Apply
≈ $96k – $241k per year (Estimated) • In office • 6+ years exp • Master's Degree • Kigali
AI/ML
Fine-tuning
Reinforcement Learning
AI Agents
LLM
Explainable AI
Machine Learning
Apply
$180k – $400k per year • Equity • In office • Full-Time • 2+ years exp • San Francisco
Python
JavaScript
TypeScript
AI/ML
LangGraph
LangChain
Embeddings
AI Agents
CrewAI
LLM
RAG
Multi-Agent Systems
Tool Use
Frontend
React.js
DevOps
GCP
CI/CD
AWS
Apply
In office • Seoul
Python
Rust
C++
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
SGLang
TensorRT
TensorRT-LLM
PyTorch
LLM
Mixture of Experts
CUDA
Triton
TPU
KV Cache
Apply
Hybrid • 3+ years exp • Bachelor's Degree • Seoul
Python
Rust
C++
Bash
Cython
Rust
PyO3
C++
CMake
Cython
Manylinux
PyBind11
DevOps
GitHub Actions
CI/CD
Git
Docker
Ubuntu
Bazel
CentOS Stream
Apply
Hybrid • Hwaseong
Rust
C++
DevOps
Linux
Apply
In office • Seoul
C
C++
C
Embedded C
DevOps
RTOS
HPC
Apply
In office • Bachelor's Degree • Seoul
Go
Rust
C++
C++
PyTorch C++
LLVM
AI/ML
PyTorch
MLIR
Apply
In office • Full-Time • Seoul
Apply
In office • Full-Time • Seoul
Analytics
Microsoft Excel
Management
Microsoft Office
Apply
Service Engineer 9 hours ago
≈ $43k – $92k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Seoul
Apply
≈ $85k – $202k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Seoul
Python
SQL
AI/ML
AI Agents
LLM
DevOps
Splunk
GCP
OpenTelemetry
Azure
AWS
Kubernetes
Cybersecurity
MITRE ATT&CK
SIEM
Apply
In office • Full-Time • 5+ years exp • Seoul
JavaScript
Frontend
Three.JS
React.js
React Three Fiber
DevOps
AWS
Apply
See all jobs
This is one of many
904,400 more open roles from verified company boards, updated every day.