1,337,370open jobs
78,389companies
205,870added this week
Browse all
Salary
≈ $196k – $390k per year (Estimated)
Location
In office (Germany)
Seniority
Middle · 3+ years exp
Visa
H-1B filings in 12 months: 7 · for this role: 3

Confirmed on the employer's own hiring board on Oct 8, 2026. First seen by Alion on Jul 10, 2025. EnCharge AI scores C on the Alion truth index.

Overview
Company
Impact
Profile match
AI Compute from the Edge-to-Cloud for Every Business. Transformative technology for AI computation, breaking records in efficiency and sustainability to enable state-of-the-art models uninhibited by power, space, and cost constraints.

EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. EnCharge’s robust and scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute efficiency and density compared to today’s best-in-class solutions. The high-performance architecture is coupled with seamless software integration and will enable the immense potential of AI to be accessible in power, energy, and space constrained applications. EnCharge AI launched in 2022 and is led by veteran technologists with backgrounds in semiconductor design and AI systems.

About the Role

EnCharge AI is seeking an AI Runtime Engineer to develop and optimize the execution stack for our next-generation AI accelerator. In this role, you will work on low-latency, high-performance runtime software that enables efficient execution of deep learning models on specialized hardware. You will collaborate with hardware, compiler, and AI framework teams to deliver optimized AI inference and training performance across cloud and edge environments. 

Responsibilities

  • Develop and optimize the AI runtime software stack for executing deep learning workloads on AI accelerators.
  • Implement task scheduling, memory management, and kernel execution strategies for efficient computation.
  • Optimize data movement between host and device using PCIe, DMA, shared memory.
  • Design and implement high-performance APIs for AI Inference frameworks such as OpenVino, ONNX Runtime, vLLM
  • Work on graph execution optimizations, including kernel fusion, pipelining, tensor tiling, and caching.
  • Integrate runtime components with AI compilers (LLVM, MLIR, XLA, TVM) for optimized execution.
  • Ensure scalability and reliability of the AI runtime for cloud-based and edge AI deployments. 

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, or a related field.
  • 3+ years of experience in developing low-level runtime software for AI accelerators, GPUs, or HPC systems.
  • Strong proficiency in C/C++ and low-level systems programming.
  • Deep understanding of task scheduling, concurrency, and memory hierarchy.
  • Experience with hardware-aware optimizations and dataflow architectures.
  • Familiarity with deep learning execution frameworks (ONNX Runtime, TensorRT, TVM, OpenVINO).
  • Experience with low-latency, high-throughput workload execution for AI models.
  • Strong debugging and profiling skills for optimizing AI execution performance.
  • Exposure to AI model deployment pipelines (Triton, TensorFlow Serving). 

EnchargeAI is an equal employment opportunity employer in the United States.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,337,370 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Germany
Applied ML Engineer 3 hours ago
$102k – $202k per year • In office • 5+ years exp • Bachelor's Degree • Redmond
AI/ML
Multimodal AI
Edge AI
Machine Learning
Apply
≈ $139k – $277k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Mountain View
Python
Java
C++
AI/ML
Gemini
LLM
Machine Learning
Apply
$315k – $380k per year • Hybrid • 8+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
Claude
Multimodal AI
AI Agents
LLM
Anthropic
Context Engineering
Interpretability
Apply
≈ $166k – $379k per year (Estimated) • In office • 2+ years exp • Seattle
Python
SQL
AI/ML
Spark
PyTorch
Management
Stripe
Apply
Applied AI Engineer 2 hours ago
≈ $143k – $273k per year (Estimated) • In office • Secret • 5+ years exp • High School Diploma • Doral
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Fine-tuning
Computer Vision
NLP
TensorFlow
PyTorch
LLM
Machine Learning
DevOps
Azure
AWS
Docker
Kubernetes
Apply
$82k – $108k per year • In office • 2+ years exp • Bachelor's Degree • Princeton
Python
C++
AI/ML
OpenCV
NumPy
DevOps
Linux
Apply
≈ $122k – $232k per year (Estimated) • Hybrid • 5+ years exp • Bachelor's Degree • Sterling
Python
Java
SQL
C#
C#
.NET
Databases
PostgreSQL
Chroma
pgvector
Pinecone
FAISS
OpenSearch
AI/ML
LangChain
LlamaIndex
Embeddings
Scikit-learn
Prompt Engineering
Function Calling
NLP
Semantic Kernel
AWS Bedrock
TensorFlow
PyTorch
LLM
RAG
Reranking
Semantic Search
OpenAI
Hugging Face
Amazon SageMaker
LLMOps
Human-in-the-Loop
Semantic Search
LLM Guardrails
Machine Learning
DevOps
Rest API
Azure
CI/CD
Git
AWS
Docker
Kubernetes
AWS Lambda
Amazon S3
Management
Agile
Scrum
Apply
≈ $132k – $279k per year (Estimated) • In office • 6+ years exp • Master's Degree • Cambridge
Python
AI/ML
Scikit-learn
TensorFlow
PyTorch
Edge AI
Machine Learning
Apply
≈ $102k – $199k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • New Castle
C++
DevOps
RTOS
Linux
Apply
$150k – $200k per year • Remote (United States) • 3+ years exp • Bachelor's Degree
C++
Objective-C
Apply
≈ $53k – $120k per year (Estimated) • Remote (India) • 10+ years exp
Python
Go
Rust
C++
AI/ML
Model Context Protocol
Quantization
Function Calling
LLM
RAG
Mixture of Experts
LLM Evaluation
Tool Use
Apply
$130k – $172k per year • In office • 5+ years exp • Germany
Python
AI/ML
LoRA
Fine-tuning
Quantization
Multimodal AI
Knowledge Distillation
Diffusion Models
PEFT
PyTorch
Mixture of Experts
Post-training
Speculative Decoding
KV Cache
Model Distillation
Apply
Remote (India) • 5+ years exp
Python
AI/ML
LoRA
Fine-tuning
Quantization
Multimodal AI
Knowledge Distillation
Diffusion Models
PEFT
PyTorch
Mixture of Experts
Post-training
Speculative Decoding
KV Cache
Model Distillation
Apply
AI Compiler Engineer 16 days ago
$190k – $255k per year • Remote (United States, Germany, Canada, Norway) • 3+ years exp • Bachelor's Degree
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Quantization
TensorFlow
PyTorch
TPU
Edge AI
MLIR
Apply
AI Compiler Engineer 3 months ago
≈ $32k – $80k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • India
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Quantization
TensorFlow
PyTorch
TPU
Edge AI
MLIR
Apply
≈ $59k – $140k per year (Estimated) • In office • 8+ years exp • Germany
Apply
≈ $79k – $174k per year (Estimated) • In office • 10+ years exp • Bachelor's Degree • Germany
Apply
≈ $31k – $53k per year (Estimated) • In office • Part-Time • Germany
Management
Microsoft Office
Apply
≈ $31k – $53k per year (Estimated) • In office • Part-Time • Germany
Management
Microsoft Office
Apply
In office • Full-Time • Erfurt • Jena • Dresden • Leipzig
Management
Microsoft Office
Apply
See all jobs
This is one of many
1,337,370 more open roles from verified company boards, updated every day.