805,292open jobs
51,725companies
127,236added this week
Browse all
Salary
≈ $165k – $300k per year (Estimated)
Location
In office (San Mateo)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Dec 24, 2024.

Overview
Company
Impact
Profile match
OpenInfer is an innovation lab building one unified inference stack, from kernel to cloud — reimagining inference around the economics of AI demand: cost, reliability, and sovereignty.

Position Overview

We are looking for highly skilled engineers with a focus on C/C++, low level systems, performance and power optimization to join our team full-time. In this role, you will apply your expertise in performance optimization and contribute to building a scalable, efficient inference engine to power the future of on-device AI use cases, such as real-time agents and assistants. You will work with seasoned engineers to enhance our end to end inference stack, power management, and overall system efficiency.

Key Responsibilities

  • Innovate on the inference optimization pipeline through algorithmic and system optimization
  • Own end to end system characterization across a range of hardware
  • Own the design and implementation of inference optimizations for AI workloads, targeting peak efficiency on diverse hardware.
  • Engage in performance benchmarking, profiling, and troubleshooting to improve execution across various hardware.
  • Work on cross-functional teams to design, implement, and test new features.

Qualifications

  • 5+ years of hands-on experience in C/C++ development with a focus on performance optimization.
  • Strong understanding of low level systems and efficiency in the context of high-performance computing.
  • Familiarity with GPU-based computing and CUDA or similar GPU programming environments.
  • Solid knowledge of system design and performance optimization techniques.
  • Experience with open-source contributions and community-driven projects is a plus.

What You’ll Gain

  • Opportunity to work alongside industry experts in AI optimization, high-performance computing, and hardware acceleration.
  • Hands-on experience with cutting-edge technologies at the intersection of AI and hardware acceleration.
  • Exposure to open-source development and collaboration with a vibrant community.

Benefits We Offer:

At OpenInfer we offer comprehensive benefits, some include:

  • Medical, Dental, and Vision benefits
  • Flexible Paid Time Off, 10 days
  • Parental Leave
  • 401(k) Plan with company matching
  • Snacks and coffee to keep you energized

These benefits are further detailed in OpenInfer policies and are subject to change at any time, consistent with the terms of any applicable compensation or benefits plans.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
805,292 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Mateo
$130k – $150k per year • Remote (United States) • Full-Time • Master's Degree
Apply
$130k – $140k per year • In office • Contractor • Palo Alto
Apply
$175k – $350k per year • In office • Full-Time • New York
Python
AI/ML
AI Agents
NLP
Apply
$175k – $350k per year • In office • Full-Time • New York
Python
JavaScript
TypeScript
AI/ML
Jupyter Notebook
AI Agents
NLP
Frontend
React.js
Apply
$175k – $350k per year • In office • Full-Time • New York
Python
AI/ML
AI Agents
NLP
Management
Slack
Apply
In office • PhD
Python
C++
AI/ML
CUDA Toolkit
Quantization
Knowledge Distillation
TensorRT
OpenCL
OpenVINO
CUDA
ONNX Runtime
Model Distillation
Apply
≈ $106k – $228k per year (Estimated) • Hybrid • Full-Time • Singapore
Python
C#
C++
AI/ML
Edge AI
Machine Learning
Apply
In office • 5+ years exp
C#
C++
DevOps
RTOS
CI/CD
Linux
TCP/IP
Management
Jira
Agile
Apply
In office • 5+ years exp • Bachelor's Degree
Python
C++
DevOps
CI/CD
Jenkins
Git
Bitbucket
Linux
Windows
QA
Pytest
Apply
In office • 5+ years exp • Bachelor's Degree
Python
C++
MATLAB
Apply
≈ $100k – $226k per year (Estimated) • In office • Full-Time • 1+ year exp
Python
Java
Rust
Kotlin
C++
Swift
Rust
Axum
AI/ML
llama.cpp
LocalAI
Ollama
LLM
Mobile
Android NDK
DevOps
Rest API
GitHub Actions
CI/CD
Docker
Linux
Windows
Apply
≈ $155k – $335k per year (Estimated) • In office • Full-Time • San Mateo
Python
AI/ML
CUDA Toolkit
PyTorch
Tokenization
CUDA
KV Cache
Apply
Senior ML Engineer 1 day ago
$100k – $160k per year • Equity 0.3–0.3% • Remote (likely United States) • Full-Time • 11+ years exp • Bachelor's Degree • San Mateo
AI/ML
OpenCV
Fine-tuning
Computer Vision
Machine Learning
Apply
$100k – $160k per year • Equity 0–0.2% • In office • Full-Time • 11+ years exp • Bachelor's Degree • San Mateo
Python
C++
C++
TensorFlow C++
AI/ML
OpenCV
Fine-tuning
TensorFlow
Machine Learning
Apply
Sales Engineer 8 hours ago
$90k – $110k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • San Mateo
Python
Management
Confluence
Apply
$48k – $56k per year • In office • Internship • Bachelor's Degree • San Mateo
PHP
PHP
WordPress
Design
Adobe Photoshop
Adobe After Effects
Marketing
HubSpot
YouTube
Apply
$90k – $110k per year • Remote (United States) • Full-Time • 3+ years exp • San Mateo
Apply
See all jobs
This is one of many
805,292 more open roles from verified company boards, updated every day.