412,433open jobs
14,112companies
70,830added this week
Browse all
Salary
$75k – $141k per year (Estimated)
Location
In office (Paris)
Employment
Full-Time
Overview
Company
Impact
Profile match
Arago develops photonic processors that run artificial intelligence workloads using light instead of electrons. Its aim is to cut the energy consumption of large model inference by an order of magnitude. The company is headquartered in Paris.

Meet Arago and the Aragonians

Arago’s mission is to re-engineer the foundations of computing from first principles.

The explosive growth of AI is pushing the industry to rethink how processors are built. Arago is meeting that challenge with a proprietary technology that fuses optical and CMOS technologies to deliver an order-of-magnitude increase in performance.

Arago is the fastest, and currently the only, company to have built such a processor. It's backed by leading deep-tech investors and some of the most respected figures in semiconductors and computing, including the CEO of Arm, the founder of macOS who worked directly with Steve Jobs at Apple, an Nvidia Fellow, the Head of Optics at Google, and many other industry leaders.

Our work is guided by three clear values: do great things, move with high velocity, and operate as one unit. We work in a demanding environment where constant learning, ownership, and execution are expected, and where exceptional people have the opportunity to do their life’s work.

What you’ll do

Optimize the execution and serving of modern AI models on Arago's custom accelerator. Work across kernels, model execution, multi-device distribution, runtime, and inference serving, while helping shape the software stack around the capabilities of Arago's hardware.

Required Skills and Experience

  • Strong experience in high-performance ML inference, GPU/accelerator programming, or ML systems engineering.

  • Deep understanding of computer architecture, accelerator/GPU execution models, memory hierarchies, parallelism, and performance bottlenecks.

  • Experience developing and optimizing custom kernels using CUDA, Triton, ROCm/HIP, or equivalent low-level programming environments.

  • Experience with operator fusion, tiling, scheduling, data movement optimization, graph execution, and profiling of compute- and memory-bound workloads.

  • Strong understanding of distributed model execution, including tensor, pipeline, sequence, and/or expert parallelism and communication/computation overlap.

  • Hands-on experience with modern inference-serving systems such as vLLM, SGLang, TensorRT-LLM, or equivalent, including KV-cache management, continuous batching, paged attention, and prefill/decode scheduling.

  • Strong C++ and Python skills, and comfort working on a custom accelerator stack where compiler, runtime, kernels, and abstractions are actively being developed. Exposure to or experience with MLIR and MLIR dialects is a strong plus.

  • Language: English at a proficient level.

Responsibilities

  • Analyze modern AI workloads and identify kernel-, runtime-, memory-, and system-level bottlenecks on Arago's accelerator.

  • Develop and optimize custom kernels, fused operators, and execution strategies to maximize device utilization.

  • Design efficient mappings of models and operators across multiple Arago devices, including communication and synchronization strategies.

  • Develop inference-serving techniques such as continuous batching, paged KV caches, prefix/context caching, chunked prefill, and prefill/decode interleaving or disaggregation.

  • Build profiling, benchmarking, and performance-analysis infrastructure spanning kernels, full models, and serving workloads.

  • Work closely with Arago's hardware, compiler, and runtime teams to co-design software abstractions and influence future hardware features based on real model workloads.

Pay and benefits

  • Competitive cash compensation, with final package based on location, experience, and the pay of team members in similar positions.

  • Meaningful stock option plan offered (included in the majority of full time offers).

  • Healthcare coverage (including family-friendly options), pension contributions, professional development support, and 25 days of PTO, in addition to public holidays.

  • Ownership of a key technical domain, with significant vertical and/or horizontal growth opportunities, based on performance and individual drive.

We look forward to hearing how you can help shape the future of AI at Arago.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
412,433 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Paris
SMTS Logic Design 1 day ago
$80k – $189k per year (Estimated) • Remote/Hybrid • 8+ years exp • Bachelor's Degree • Montreal
Python
C++
SystemVerilog
Chips/EDA
Synopsys Design Compiler
Synopsys Fusion Compiler
Synopsys SpyGlass
Cadence Genus
Cadence Palladium
Siemens Veloce
Synopsys HAPS
Apply
In office • Bachelor's Degree
Python
C
C
FFmpeg
AI/ML
CUDA Toolkit
Triton Inference Server
Embeddings
Multimodal AI
Function Calling
Computer Vision
VLM
TensorRT
PyTorch
LLM
RAG
CUDA
Triton
NVIDIA NeMo
Structured Outputs
DevOps
AWS
Docker
Robotics
GStreamer
Apply
$37k – $85k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • United Kingdom
Python
Java
C++
Apply
$37k – $85k per year (Estimated) • Remote/Hybrid • Full-Time • United Kingdom
Python
Java
C++
Apply
$124k – $230k per year • Equity • Remote/Hybrid • Bachelor's Degree • San Jose
C++
DevOps
Red Hat
Apply
$57k – $132k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Paris
AI/ML
Hugging Face
Apply
$66k – $154k per year (Estimated) • Equity • In office • Full-Time • Paris • Tel Aviv
Python
Verilog
SystemVerilog
Chips/EDA
UVM
Apply
$66k – $153k per year (Estimated) • Equity • In office • Full-Time • Paris • Tel Aviv
Verilog
SystemVerilog
AI/ML
Hugging Face
Chips/EDA
UVM
Apply
Equity • In office • Full-Time • Paris • Tel Aviv
Verilog
SystemVerilog
AI/ML
Hugging Face
Apply
$46k – $68k per year (Estimated) • Equity • In office • Full-Time • Paris
Python
C++
Apply
$85k – $203k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Paris
Apply
$8.5k – $22k per year (Estimated) • In office • Internship • Paris
Apply
$86k – $130k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree • Madrid • Milan • Lisbon • Paris
Apply
In office • Internship • Paris
Analytics
Power BI
Management
Google Workspace
Google Sheets
Apply
$42k – $167k per year (Estimated) • In office • Full-Time • Paris
Apply
See all jobs
This is one of many
412,433 more open roles from verified company boards, updated every day.