Logo-of-Institute-Of-Foundation-Models-hiring-for-jobs-in-France-on-GrabJobs

Research Scientist - Speech/Audio Machine Learning

icon briefcase Type d'emploi : À plein temps

Nombre de candidats

 : 

000+

Click to reveal the number of candidates who applied for this job.
icon loader
icon loader

Let AI Supercharge Your Job Hunt!

JobCopilot scans 500,000+ company career sites daily to find jobs for you

Never miss an opportunity Save hours by auto-filling applications forms Land more interviews with tailored applications
happy man
thunder iconActivate JobCopilot

Description de l'emploi - Research Scientist - Speech/Audio Machine Learning

About the Institute of Foundation Models 

We are a dedicated research lab for building, understanding, using, and risk-managing foundation models. Our mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-driven economy.

As part of our team, you’ll have the opportunity to work on the core of cutting-edge foundation model training, alongside world-class researchers, data scientists, and engineers, tackling the most fundamental and impactful challenges in AI development. You will participate in the development of groundbreaking AI solutions that have the potential to reshape entire industries. Strategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers.

The Role

As a Research Scientist specializing in speech/audio machine learning, you will contribute to the design and training of SoTA end-to-end neural speech models. You will be responsible for developing the core intellectual property, moving beyond cascaded ASR → TTS systems toward native audio-to-audio multimodal architectures.

Key responsibilities

    • Architectural Design: Develop novel neural architectures for low-latency speech-to-speech translation and generation (e.g., Diffusion, Flow-matching, Transformer-based audio LLMs). 
    • Loss Function Engineering: Design and implement custom objective functions to optimize prosody (emotions, intelligibility, naturalness). 
    • Experimental Iteration: Conduct large-scale training runs, performing ablation studies on model architecture and tokenization strategies. 
    • Evaluation Frameworks: Establish rigorous internal benchmarks using both objective metrics (WER, MCD) and subjective human-in-the-loop (MOS) testing. 

Qualifications

    • PhD or MSc in Computer Science with a focus on Deep Learning, Signal Processing, or Computational Linguistics. 
    • Record of Research: Published work in top-tier venues (NeurIPS, ICLR, ICASSP, Interspeech). 
Original job Research Scientist - Speech/Audio Machine Learning posted on GrabJobs ©. To flag any issues with this job please use the Report Job button on GrabJobs.
Share Job
Share Job

Auto-Apply to Research Scientist Jobs with your AI JobCopilot

thunder icon Auto-Apply with AI

Similar Research Scientist Jobs in France

GrabJobs est le portail d'emploi n°1 en :country, te connectant rapidement à des milliers d'emplois ! Trouve les meilleurs emplois de dans France, postule en 1 clic et obtiens un emploi dès aujourd'hui !

Applications mobile

Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.