Skip to content

Machine Learning Scientist (L4/L5) - Audio & Speech for Games

Is this job for you?

Build your CV and see how well you match this role — and every other one.

Build my CV

The role

You will redefine the auditory experience in interactive games by combining cutting-edge speech technology with production-focused engineering.
Evaluate and integrate open-source and commercial models (ASR, TTS, diarization, VAD), balancing quality, latency, and cost for scalable gameplay.
Lead R&D and adaptation of speech models (TTS, voice cloning, ASR) using parameter-efficient techniques like LoRA/PEFT to address production gaps.
Drive data and alignment efforts—audio and multimodal data curation, labeling, segmentation, and synthetic data generation.
Scale speech model training through distributed data pipelines and training strategies for large datasets.
Collaborate across teams to translate research into production-ready audio features for games, balancing practicality and innovation.

See the full job post

Responsibilities, requirements, skills and benefits — create your free account.

At least 8 characters, one uppercase letter and one digit.

Already have an account? Log in

Similar openings

Other roles that could suit you.

See all →

Your location

Jobs and companies will be filtered on this country.

Suggested

All countries 65