Research Engineer, Artificial General Intelligence
応募状況をスマホで確認 Whileresume の無料アプリ(iPhone・Android)。
仕事内容
Design and implement scalable multimodal LLM pipelines for pre- and post-training stages.
Scale model training on large GPU clusters and AWS Trainium, optimizing distributed training and parallelism.
Tune low-level training components, including CUDA kernels, collectives, and network I/O to improve efficiency.
Prototype and evaluate novel algorithms using industry-leading frameworks (NeMo, Megatron Core, PyTorch, Jax, vLLM, TRT).
Collaborate across teams in an Agile environment, delivering robust features while adapting to new scientific advances.
Contribute to system architecture decisions and establish best practices for scalable AI infrastructure.
Scale model training on large GPU clusters and AWS Trainium, optimizing distributed training and parallelism.
Tune low-level training components, including CUDA kernels, collectives, and network I/O to improve efficiency.
Prototype and evaluate novel algorithms using industry-leading frameworks (NeMo, Megatron Core, PyTorch, Jax, vLLM, TRT).
Collaborate across teams in an Agile environment, delivering robust features while adapting to new scientific advances.
Contribute to system architecture decisions and establish best practices for scalable AI infrastructure.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログインこちらの求人もおすすめです
近い求人はまだありません — 最新の求人をご紹介します。