Machine Learning Scientist (L4/L5) - Audio & Speech for Games
仕事内容
You will redefine the auditory experience in interactive games by combining cutting-edge speech technology with production-focused engineering.
Evaluate and integrate open-source and commercial models (ASR, TTS, diarization, VAD), balancing quality, latency, and cost for scalable gameplay.
Lead R&D and adaptation of speech models (TTS, voice cloning, ASR) using parameter-efficient techniques like LoRA/PEFT to address production gaps.
Drive data and alignment efforts—audio and multimodal data curation, labeling, segmentation, and synthetic data generation.
Scale speech model training through distributed data pipelines and training strategies for large datasets.
Collaborate across teams to translate research into production-ready audio features for games, balancing practicality and innovation.
Evaluate and integrate open-source and commercial models (ASR, TTS, diarization, VAD), balancing quality, latency, and cost for scalable gameplay.
Lead R&D and adaptation of speech models (TTS, voice cloning, ASR) using parameter-efficient techniques like LoRA/PEFT to address production gaps.
Drive data and alignment efforts—audio and multimodal data curation, labeling, segmentation, and synthetic data generation.
Scale speech model training through distributed data pipelines and training strategies for large datasets.
Collaborate across teams to translate research into production-ready audio features for games, balancing practicality and innovation.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログイン類似の求人
あなたに合いそうな他の職種。
? Machine Learning Intern Fall 2026 (Toronto) ? Machine Learning Engineer, Behavior Internship ? Machine Learning Engineer (L3) ? Senior C++ Programmer - Machine Learning Content Creation Technology Group ? Senior C++ Programmer - Machine Learning - Content Creation Technology Group ? Team Lead - Machine Learning Engineer
リモートワークNo
勤務地Los Gatos, USA