- Location
- Berlin, Berlin, Germany
- Workplace
- Hybrid
- Employment
- Full time
- Level
- Mid level
- Posted
- 9 months ago
About this role
Mirelo AI is building the next generation of creative tools by generating realistic sound, speech and music from video.
We develop cutting-edge foundational generative AI models that "unmute" silent video content and create custom, hyper-realistic audio for gaming, video platforms, and creators. Our technology empowers global storytellers to transform their content.
We recently closed a $41 million Seed round co-led by Andreessen Horowitz and Index Ventures with participation from Atlantic, and are rapidly expanding across Product, Engineering, Go-to-Market, and Growth.
About the Role
In this role, you’ll focus on the full training stack - profiling GPU behavior, debugging training pipelines, improving throughput, choosing the right parallelism strategies, and designing the infrastructure that lets us train models efficiently at scale. You’ll work across cluster management, model training, efficient data pipelines for video and audio, inference and optimizing pytorch code. Your work will shape the foundation on which all of our generative models are built and iterated.
Key Responsibilities
Find ideal training strategies (parallelism approaches, precision trade-offs) for a variety of model sizes and compute loads
Profile, debug, and optimize single and multi-GPU operations using tools like Nsight and stack trace viewers to understand what's actually happening at the hardware level
Analyze and improve the whole training pipeline from start to end (efficient data storage, data loading, distributed training, checkpoint/artifact saving, logging, …)
Set up scalable systems for experiment tracking, data/model versioning, experiment insights.
Design, deploy and maintain large-scale ML training clusters running SLURM for distributed workload orchestration
Ideal Candidate Profile
Familiarity with the latest and most effective techniques in optimizing training and inference workloads - not from reading papers, but from implementing them
Deep understanding of GPU memory hierarchy and computation capabilities - knowing what the hardware can do theoretically and what prevents us from achieving it
Experience optimizing for both memory-bound and compute-bound operations and understanding when each constraint matters
Expertise with efficient attention algorithms and their performance characteristics at different scales
Nice to Have
Experience in implementing custom GPU kernels and integrating them into PyTorch.
Experience with diffusion and autoregressive models and understanding of their specific optimization challenges
Familiarity with high-performance storage solutions (VAST, blob storage) and understanding of their performance characteristics for ML workloads
Experience with managing SLURM clusters at scale
Why Join?
Join at a pivotal moment. We've secured fresh funding and are gaining traction - now is when your contributions can make a real difference to our success.
True ownership from day one. You'll have genuine autonomy and responsibility. Your ideas and work will directly shape our product and company direction.
Competitive compensation and equity. We offer strong packages that ensure you share in the success you help create.
Build for the next generation of creators. Be part of the innovation that will transform how creators work and thrive.
We welcome applications from all individuals, regardless of ethnic origin, gender, disability, religion or belief, age, or sexual orientation and identity.
As published by Mirelo AI. Applications are handled on their site.
Skills this posting mentions
About Mirelo AI
Mirelo is a research lab building frontier AI audio models for videos. Our goal is to make audio a fun and central part of the creative process, so every frame lands with the emotion, depth, and impact it deserves. Mirelo. Sound on.
All 8 openings at Mirelo AIOne click, then it is written
Apply to Mirelo AI with a resume written for this role.
Queue Training Infrastructure Engineer and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Mirelo AI’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at Mirelo AI
See all- 7 months ago
- 9 months ago
- 9 months ago
- 9 months ago
- 9 months ago
- 9 months ago
Similar roles elsewhere
See morePut this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board Mirelo AI publishes. Refolk is not the employer and does not handle their hiring. Applications go to Mirelo AI directly.