368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$87k – $207k per year (Estimated)
Location
Remote/Hybrid (Berlin, Tübingen, Germany)
Employment
Full-Time
Overview
Company
Impact
Profile match
AI-powered sound and music generation for videos. We are passionate AI researchers, product experts and musicians.

Mirelo AI is building the next generation of creative tools by generating realistic sound, speech and music from video.

We develop cutting-edge foundational generative AI models that "unmute" silent video content and create custom, hyper-realistic audio for gaming, video platforms, and creators. Our technology empowers global storytellers to transform their content.

We recently closed a $41 million Seed round co-led by Andreessen Horowitz and Index Ventures with participation from Atlantic, and are rapidly expanding across Product, Engineering, Go-to-Market, and Growth.

About the Role

At Mirelo, you’ll work at the centre of how we build the next generation of multimodal video-to-audio models. This role is deeply hands-on and research-heavy: with a great H100/200-per-engineer ratio you explore and build new multimodal models and push the boundaries of what’s possible in music, sound, and speech generation. You’ll collaborate closely across research and engineering, run focused ablations, and translate experimental results into clear next steps for the team. From data curation to deployment, you’ll help shape the full lifecycle of the models that power our products and partnerships.

Key Responsibilities

  • Design, implement and train large-scale multimodal generative models for audio generation (diffusion and/or autoregressive models).

  • Explore new modeling ideas for audio generation (music, sound, speech) while taking inspiration from the language and image domains.

  • Develop and experiment with post-training for new capabilities (fine-grained control, in/out-painting, editing, …)

  • Conduct rigorous ablation studies, get actionable insights and communicate results to the team to discuss new research directions.

  • Contribute hands-on to all stages of model development including data curation, experimentation, evaluation, and deployment.

Ideal Candidate Profile

  • Hands-on experience in training large-scale generative models in a fast-paced research environment.

  • Deep understanding of cutting-edge methods and ML research in at least one of the domains: image, language, video or audio (specific audio experience not necessary, but nice to have).

  • Strong proficiency in PyTorch, transformer architectures, and the full ecosystem of modern deep learning.

  • Solid understanding of distributed training techniques-FSDP, low precision training, model parallelism

  • Strong track-record in working on generative models (publications in top-tier venues, open-source contributions or applied ML projects).

Nice to Have

  • Proficiency with profiling, debugging, and optimizing single and multi-GPU operations using tools like Nsight or stack trace viewers.

  • Strong software engineering skills/experience in collaborating on large codebases that go beyond PhD research code.

  • Experience with generative models for audio (sound, music or speech) and audio codec design.

Why Join?

  • Join at a pivotal moment. We've secured fresh funding and are gaining traction - now is when your contributions can make a real difference to our success.

  • True ownership from day one. You'll have genuine autonomy and responsibility. Your ideas and work will directly shape our product and company direction.

  • Competitive compensation and equity. We offer strong packages that ensure you share in the success you help create.

  • Build for the next generation of creators. Be part of the innovation that will transform how creators work and thrive.

We welcome applications from all individuals, regardless of ethnic origin, gender, disability, religion or belief, age, or sexual orientation and identity.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Berlin
Founding Engineer 1 day ago
$180k – $250k per year • In office • Full-Time • 3+ years exp • New York
Node JS
Python
TypeScript
JavaScript
Node JS
BullMQ
Databases
Redis
AI/ML
AI Agents
LangGraph
LLM
Multimodal AI
LangChain
Anthropic
Deepgram
LiveKit
LLM Guardrails
OpenAI
Text-to-Speech
Frontend
Next.js
React.js
DevOps
Vercel
Apply
$67k – $160k per year (Estimated) • In office • Full-Time • France
C++
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
Multimodal AI
OpenCV
PyTorch
TensorFlow
Edge AI
Apply
In office • Contractor • Bachelor's Degree • Seoul
AI/ML
Mamba
Multimodal AI
TPU
Apply
In office • Bachelor's Degree • Seoul
AI/ML
Multimodal AI
RAG
AI Agents
DevOps
GitHub
Analytics
A/B Testing
Apply
System Designer 2 days ago
$170k – $200k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
AI Agents
Multimodal AI
Apply
$82k – $177k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp
AI/ML
MusicGen
AudioCraft
Apply
$87k – $163k per year (Estimated) • Remote/Hybrid • Full-Time
AI/ML
MusicGen
AudioCraft
Apply
$76k – $152k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Berlin • Tübingen
Go
Python
TypeScript
JavaScript
AI/ML
Claude
Claude Code
Cursor
LLM
Frontend
React.js
shadcn/ui
Tailwind CSS
Radix UI
DevOps
AWS
AWS Lambda
Amazon S3
IAM
Apply
$86k – $204k per year (Estimated) • Remote/Hybrid • Full-Time • Berlin • Tübingen
AI/ML
Diffusion Models
Apply
$76k – $165k per year (Estimated) • Remote/Hybrid • Full-Time • Berlin • Tübingen
AI/ML
PyTorch
DevOps
SLURM
Apply
Product Engineer 4 hours ago
$81k – $128k per year • In office • Full-Time • 1+ year exp • Berlin
TypeScript
JavaScript
AI/ML
AI Agents
Human-in-the-Loop
Frontend
React.js
Tailwind CSS
Apply
Backend Engineer 4 hours ago
$81k – $128k per year • In office • Full-Time • 1+ year exp • Berlin
TypeScript
Node JS
JavaScript
Node JS
Nest.JS
Prisma
Databases
PostgreSQL
DevOps
Pulumi
Terraform
Apply
$81k – $128k per year • In office • Full-Time • 1+ year exp • Berlin
Node JS
TypeScript
JavaScript
Node JS
Nest.JS
Prisma
Databases
PostgreSQL
Frontend
React.js
Apply
$81k – $178k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Berlin
Design
Adobe Photoshop
Figma
Apply
$73k – $146k per year (Estimated) • Equity • In office • Full-Time • Berlin
Java
Kotlin
Java
Hibernate
Spring Boot
Mobile
JUnit
DevOps
CI/CD
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.