368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$158k – $288k per year (Estimated)
Location
In office (Santa Clara)
Seniority
Architect
Overview
Company
Impact
Profile match
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.

About FuriosaAI

FuriosaAI builds high-performance, high-efficiency AI compute for the Inference Era. Founded in 2017 by veteran semiconductor and AI algorithm engineers, Furiosa operates globally with offices in Korea and Silicon Valley, along with a compiler-focused R&D lab in Lisbon. 

Our vision is to make AI computing sustainable, enabling access to powerful AI for everyone on Earth. We solve the AI hardware energy and operational cost crisis at the architectural level, rather than through brute force, building the world's first truly AI-native compute platform to unlock the full potential of artificial intelligence  for every enterprise.

About the Role

FuriosaAI is looking for a Solutions Architect to bring the full potential of our powerful RNGD chips/servers to our customers by acting as the primary technical authority in AI/LLM model deployments. From running POCs to benchmarking and debugging, you will translate RNGD’s powerful system to real-world deployments of customers’ models, empowering customers with FuriosaAI’s powerful solutions.

If you are interested in providing the technical expertise in challenging the current status-quo of AI infrastructure in real-world environments, join us in our path to a sustainable future of AI.

Key Responsibilities

  • Own end-to-end technical enablement for US customers deploying AI models on FuriosaAI's RNGD NPU using the Furiosa SDK

  • Develop POCs, benchmarking studies, and live debugging sessions directly in customer environments

  • Act as the technical authority to the US BD/Sales team during pre-sales and enterprise evaluations; translate deep technical capability into business value for engineering and C-suite audiences

  • Develop deep, current expertise in FuriosaAI's hardware and software stack and demonstrate it at US technical forums, AI conferences, and customer workshops

  • Onboard and train customers on integration patterns, optimization workflows, and best practices post-purchase

  • Serve as a technical feedback loop from US customers back to Seoul HQ product and engineering teams

Minimum Qualifications

  • 2-5 years in a US customer-facing technical role: Solutions Architect, Sales Engineer, Forward Deployed Engineer, or equivalent at an AI infra, cloud, or semiconductor company

  • Actively current on the AI/LLM landscape - tracking model releases, inference frameworks, and serving stack evolution in real time

  • Hands-on experience with modern inference stacks: vLLM, SGLang, TensorRT-LLM, Triton Inference Server, or similar

  • Hands-on experience with agent and orchestration frameworks: LangChain, LlamaIndex, LangGraph, AutoGen, or MCP-based tooling

  • Proficiency in Python; comfortable with DNN frameworks (PyTorch, TensorFlow)

  • Strong written and verbal communication - able to engage credibly with ML engineers at frontier labs and VP/C-suite executives

  • Authorized to work in the US; able to travel to customer sites and to Seoul HQ periodically

Preferred Qualifications

  • Prior experience at a US AI chip company, cloud silicon team, or AI infrastructure startup

  • Familiarity with NPU/GPU accelerator ecosystems, PCIe integration, and data center hardware deployment

  • Experience with inference optimization: quantization, kernel tuning, batching strategies, memory bandwidth optimization

  • Proficiency in C, C++, or Rust

  • Experience working with distributed or cross-timezone engineering teams

Contact

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
$152k – $239k per year • Remote • Full-Time • 8+ years exp
SQL
Databases
Snowflake
AI/ML
LLM
Model Context Protocol
Analytics
A/B Testing
Apply
$69k – $171k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Bogotá
SQL
Databases
Databricks
AI/ML
dbt
LLM
NLP
Context Engineering
Analytics
Power BI
Tableau
Apply
UI/UX Designer 1 day ago
$12k – $55k per year (Estimated) • Remote • Full-Time
AI/ML
Claude
LLM
LLM Guardrails
Lovable
Design
Figma
Apply
$42k – $107k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Mexico City
C#
Go
Java
Rust
AI/ML
AI Agents
LLM
Model Context Protocol
Cybersecurity
Delinea
Zero Trust
Apply
In office • Contractor • Bachelor's Degree • Seoul
AI/ML
Mamba
Multimodal AI
TPU
Apply
In office • Bachelor's Degree • Seoul
AI/ML
Multimodal AI
RAG
AI Agents
DevOps
GitHub
Analytics
A/B Testing
Apply
In office • 3+ years exp • Bachelor's Degree • Seoul
Apply
In office • Seoul
C++
Rust
Apply
Remote/Hybrid • Hwaseong
C++
Rust
Apply
$72k – $99k per year • Equity • In office • Full-Time • Santa Clara
Apply
$166k – $290k per year • Equity • In office • Full-Time • 8+ years exp • Santa Clara
Management
ServiceNow
Apply
$133k – $272k per year (Estimated) • In office • Santa Clara
Go
Python
AI/ML
Edge AI
LLM
RAG
Apply
$80k – $110k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
MATLAB
Python
Apply
$142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Toronto
C++
Go
IoT
MQTT
OPC UA
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.