Machine Learning Scientist (L4/L5) - Audio & Speech for Games
职位介绍
You will redefine the auditory experience in interactive games by combining cutting-edge speech technology with production-focused engineering.
Evaluate and integrate open-source and commercial models (ASR, TTS, diarization, VAD), balancing quality, latency, and cost for scalable gameplay.
Lead R&D and adaptation of speech models (TTS, voice cloning, ASR) using parameter-efficient techniques like LoRA/PEFT to address production gaps.
Drive data and alignment efforts—audio and multimodal data curation, labeling, segmentation, and synthetic data generation.
Scale speech model training through distributed data pipelines and training strategies for large datasets.
Collaborate across teams to translate research into production-ready audio features for games, balancing practicality and innovation.
Evaluate and integrate open-source and commercial models (ASR, TTS, diarization, VAD), balancing quality, latency, and cost for scalable gameplay.
Lead R&D and adaptation of speech models (TTS, voice cloning, ASR) using parameter-efficient techniques like LoRA/PEFT to address production gaps.
Drive data and alignment efforts—audio and multimodal data curation, labeling, segmentation, and synthetic data generation.
Scale speech model training through distributed data pipelines and training strategies for large datasets.
Collaborate across teams to translate research into production-ready audio features for games, balancing practicality and innovation.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
? Machine Learning Intern Fall 2026 (Toronto) ? Machine Learning Engineer, Behavior Internship ? Machine Learning Engineer (L3) ? Senior C++ Programmer - Machine Learning Content Creation Technology Group ? Senior C++ Programmer - Machine Learning - Content Creation Technology Group ? Team Lead - Machine Learning Engineer
远程办公No
城市Los Gatos, USA