RefolkCandidates
Open nowScience and researchResearch

Research Scientist - Audio Codec

Mirelo AI · Berlin, Berlin

Location
Berlin, Berlin, Germany
Workplace
Hybrid
Employment
Full time
Level
Mid level
Posted
9 months ago

About this role

Mirelo AI is building the next generation of creative tools by generating realistic sound, speech and music from video.

We develop cutting-edge foundational generative AI models that "unmute" silent video content and create custom, hyper-realistic audio for gaming, video platforms, and creators. Our technology empowers global storytellers to transform their content.

We recently closed a $41 million Seed round co-led by Andreessen Horowitz and Index Ventures with participation from Atlantic, and are rapidly expanding across Product, Engineering, Go-to-Market, and Growth.

About the Role

At Mirelo, we’re pushing the limits of what generative audio can do, and our ability to innovate depends heavily on the quality of our underlying audio representations. In this role, you’ll work at the core of our modeling stack - designing, training, and evaluating neural audio codecs that directly shape the performance of our next-generation music and sound models. You’ll collaborate closely with the model team, experiment with both continuous and discrete representations, and build the evaluation tools that help us understand what actually moves the needle. Your work will sit at the foundation of building the best-sounding generative models in the world.

Key Responsibilities

  • Develop and implement new neural audio codecs for sound, music and speech that push the state-of-the art in sound quality and are optimized for the the use case of generative models.

  • Think about the specific challenges that arise when the codec is primarily used as a latent representation in the context of generative audio models (in the end, the ultimate goal is to build the best audio generative models)

  • Explore the trade-offs of continuous (as typically used for diffusion models) vs. discrete audio representations (as typically used for autoregressive models).

  • Develop benchmarking pipelines for codec evaluation

  • Conduct initial experiments with generative models to verify that a new candidate codec is actually useful for our downstream tasks.

Ideal Candidate Profile

  • Strong background in deep learning for audio: neural codecs, source separation, speech models, or generative audio systems

  • Specific hands-on experience in designing and training neural audio codecs

  • Solid understanding of audio signal processing fundamentals

  • Strong track record (research and/or open-source) in the field of audio ML

Nice to Have

  • Hands-on experience with generative audio models and good intuition of how the choice of the codec influences the training and performance of the generative model

  • Strong publication record (e.g., NeurIPS, ICML, ICLR, Interspeech, ICASSP, WASPAA)

Why Join?

  • Join at a pivotal moment. We've secured fresh funding and are gaining traction - now is when your contributions can make a real difference to our success.

  • True ownership from day one. You'll have genuine autonomy and responsibility. Your ideas and work will directly shape our product and company direction.

  • Competitive compensation and equity. We offer strong packages that ensure you share in the success you help create.

  • Build for the next generation of creators. Be part of the innovation that will transform how creators work and thrive.

We welcome applications from all individuals, regardless of ethnic origin, gender, disability, religion or belief, age, or sexual orientation and identity.

As published by Mirelo AI. Applications are handled on their site.

Skills this posting mentions

Product EngineeringMarketing StrategyMusic

About Mirelo AI

Mirelo is a research lab building frontier AI audio models for videos. Our goal is to make audio a fun and central part of the creative process, so every frame lands with the emotion, depth, and impact it deserves. Mirelo. Sound on.

All 8 openings at Mirelo AI

One click, then it is written

Apply to Mirelo AI with a resume written for this role.

Queue Research Scientist - Audio Codec and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Mirelo AI’s form for you.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • 25 sent a week, free
  • No card
  • Nothing sent until you say so

More roles at Mirelo AI

See all

Similar roles elsewhere

See more

Put this to work

Paste your career in once. Every application after that is written for you.

Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • New matches ranked and written before you are up.
  • Every bullet stays inside what your history supports. Nothing invented.
  • Queued, submitted, interviewing, offer: one screen, not a spreadsheet.

500 free credits on sign-up. No card. Nothing is sent until you say so.

Listed from the job board Mirelo AI publishes. Refolk is not the employer and does not handle their hiring. Applications go to Mirelo AI directly.