1,285,480open jobs
74,416companies
214,296added this week
Browse all
Salary
≈ $76k – $167k per year (Estimated)
Location
In office (Munich)

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Oct 5, 2026. Tensordyne scores C on the Alion truth index.

Overview
Company
Impact
Profile match
Tensordyne is an artificial intelligence infrastructure company headquartered in Sunnyvale, California, and founded in 2017. The company develops high-performance AI inference systems, including the Napier processor and Scorpio silicon, which utilize a proprietary logarithmic number system to reduce power consumption and increase processing speed. It operates as a fabless semiconductor firm serving data centers, hyperscalers, and neocloud providers with hardware and software solutions optimized for large-scale generative AI workloads.

About Tensordyne

Tensordyne is building a new class of AI inference system designed for high-performance, power-efficient deployment of the world’s most demanding generative AI workloads.

Our platform combines purpose-built silicon, new AI math, optimized scale-up networking, and memory architecture into a tightly integrated system purpose built for large-scale AI inference. We work with hyperscalers, Neoclouds, frontier model developers, enterprises, and infrastructure partners operating at the leading edge of AI.

As Tensordyne moves from system development into silicon bring-up, customer validation, beta deployments, and production rollout, we are building the technical customer organization that will sit directly between our engineering teams and the companies deploying the platform.

Role summary

We are looking for a Forward Deployed Inference Engineer who combines deep AI systems expertise with strong customer instincts. This person will own the path from a customer workload or model request to a technical result and, where needed, to an optimized model running successfully on Tensordyne hardware and software.

The role sits at the intersection of model architecture, inference performance, systems optimization, developer tooling, and customer deployment. You will work hands-on with engineering while also acting as a technical bridge to Product, BizDev, Sales, and customers.

What you will do

  • Turn customer workloads into fast, credible performance answers through profiling & benchmarking. Define relevant KPIs, compare against competitive baselines, and keep our evaluation methodology current with external benchmarks.
  • Model enablement & optimization: convert and bring up customer models on the Tensordyne stack, validate numerical quality, identify performance bottlenecks, and work with compiler, runtime, kernel, and system teams to improve results.
  • Deployment / forward engineering: work directly with customers and partners on technical PoCs, integration, deployment, and debugging; translate requirements into measurable acceptance criteria for quality, latency, throughput, and other relevant KPIs.
  • Track profiling-to-hardware accuracy by continuously comparing profiling/simulation results with actual hardware deployments, explain material gaps, and flag missing capabilities in the compiler, SDK, inference server, KV-cache management, or adjacent systems to the owning teams.
  • Turn repeated customer-specific learnings into reusable tooling, documentation, benchmarks, or product improvements.

Core qualifications

  • Strong hands-on experience with AI models and inference systems, especially dense and MoE LLMs (Llama, DeepSeek, Qwen, GPT-OSS, Kimi, GLM), and VLM, speech and diffusion models.
  • Strong Python and PyTorch skills and the ability to understand and modify model code.
  • Experience profiling, benchmarking, or optimizing model inference and reasoning about latency, throughput, memory, and utilization.
  • Strong problem-solving and communication skills, with the ability to drive ambiguous technical problems across team boundaries.
  • Proficiency in using AI-powered developer tools (e.g., Claude Code, Cursor).

Strong pluses

  • Experience with LLM serving and deployment stacks such as vLLM, SGLang, or similar systems.
  • Experience working directly with customers or external technical partners.
  • Experience bringing models up on new accelerators or non-standard hardware, including performance debugging across framework/runtime/hardware boundaries.
  • Practical experience with production inference techniques or environments such as quantization, distributed inference, or Kubernetes.
  • Experience navigating and contributing to Rust codebases.

Tensordyne Values

  • Think big. Pursue ambitious technical and business goals.
  • Aim for excellence. Quality matters in everything we build and deliver.
  • Own it and get it done. Take responsibility and drive results.
  • Operate with integrity. Be direct, transparent, and respectful.
  • Win as a team. Make the people around you more effective.
  • Value different perspectives. The strongest teams challenge assumptions and bring diverse experience to difficult problems.

Tensordyne is an equal opportunity employer. We believe diverse teams are better equipped to solve complex problems and build exceptional technology. All qualified applicants will receive consideration for employment without regard to age, color, gender identity or expression, marital status, national origin, disability, protected veteran status, race, religion, pregnancy, sexual orientation, or any other characteristic protected by applicable laws, regulations, and ordinances.

A note to recruitment agencies: Please do not contact Tensordyne employees or leaders regarding this role. We do not accept unsolicited agency resumes and are not responsible for fees associated with unsolicited submissions.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,285,480 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Solutions
Similar stack
Same company
Munich
≈ $82k – $170k per year (Estimated) • Hybrid • Full-Time • 6+ years exp • Munich
ABAP
ABAP
SAP BTP
Databases
Snowflake
Databricks
AI/ML
AI Agents
OpenAI
DevOps
Rest API
Design
Canva
Management
Notion
Apply
≈ $70k – $131k per year (Estimated) • Equity • In office • 2+ years exp • Berlin
Design
Canva
Management
Asana
Apply
≈ $66k – $144k per year (Estimated) • In office • Full-Time • Cologne
JavaScript
Apply
≈ $73k – $150k per year (Estimated) • Hybrid • 8+ years exp • Berlin
AI/ML
Claude
DevOps
Rest API
Cybersecurity
GDPR
Least Privilege
Management
Jira
n8n
Agile
Scrum
Apply
≈ $77k – $159k per year (Estimated) • Hybrid • 8+ years exp • Munich
AI/ML
Claude
DevOps
Rest API
Cybersecurity
GDPR
Least Privilege
Management
Jira
n8n
Agile
Scrum
Apply
$167k – $226k per year • Equity • In office • Full-Time • 6+ years exp • Master's Degree • Seattle
Python
Java
C++
C++
TensorFlow C++
AI/ML
Spark
Fine-tuning
RLHF
Scikit-learn
SciPy
AI Agents
TensorFlow
NumPy
LLM
LLM Evaluation
DevOps
AWS
Apply
$144k – $194k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Sunnyvale
Java
Rust
C#
C++
Perl
Apply
≈ $21k – $52k per year (Estimated) • In office • Full-Time • 10+ years exp • Manila
Python
SQL
Databases
PostgreSQL
Snowflake
AI/ML
Hadoop
Spark
dbt
AI Agents
LLM
Anomaly Detection
Tool Use
DevOps
GitHub Actions
CI/CD
Jenkins
AWS
Incident Management
Amazon S3
Analytics
Tableau
ETL/ELT
Management
Agile
Scrum
Apply
≈ $58k – $110k per year (Estimated) • In office • Part-Time • Australia
Python
Apply
≈ $71k – $174k per year (Estimated) • In office • Bachelor's Degree • Zachary
Python
Analytics
Tableau
Power BI
Apply
≈ $170k – $308k per year (Estimated) • In office • 10+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
Multimodal AI
DevOps
HPC
Chips/EDA
Synopsys PrimeTime
Management
Agile
Apply
≈ $64k – $151k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Munich
Python
C++
SystemC
AI/ML
Multimodal AI
Apply
≈ $204k – $382k per year (Estimated) • In office • Sunnyvale
AI/ML
vLLM
Quantization
Multimodal AI
SGLang
PyTorch
LLM
Triton
KV Cache
Apply
≈ $106k – $217k per year (Estimated) • In office • 7+ years exp • Sunnyvale
AI/ML
Multimodal AI
Apply
≈ $86k – $209k per year (Estimated) • In office • Master's Degree • Sunnyvale
Python
C++
AI/ML
Multimodal AI
LLM
TPU
NCCL
KV Cache
DevOps
HPC
Apply
CRM Manager (m/f/d) 7 hours ago
≈ $54k – $94k per year (Estimated) • In office • 4+ years exp • Munich
Apply
Remote (Germany) • Full-Time • Munich
DevOps
Windows
Cybersecurity
Okta
Management
Slack
Google Workspace
SharePoint
Service Desk
Apply
In office • Internship • San Francisco
Python
JavaScript
TypeScript
Python
FastAPI
Pydantic
Databases
PostgreSQL
pgvector
AI/ML
Langfuse
Pydantic AI
LLM
LLM Guardrails
Frontend
React.js
Vite
TanStack Router
Chakra UI
Apply
≈ $73k – $168k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Munich • Berlin
SQL
AI/ML
Cursor
Claude Code
Devin
Analytics
A/B Testing
Apply
≈ $84k – $220k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Basel • Munich
C
C
SDL
Apply
See all jobs
This is one of many
1,285,480 more open roles from verified company boards, updated every day.