Principal Machine Learning Engineer, AI Platform (Foundation Model Post-Training)
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
You will lead the post-training phase of foundation models, aligning models with business intent and domain requirements.
Design and implement scalable, distributed pipelines for SFT, RLHF, and instruction tuning on multi-node GPU clusters.
Oversee data strategy, annotation workflows, and automated data filtering to curate high-quality instruction data.
Build comprehensive evaluation suites, including automated benchmarks and human-in-the-loop protocols, to assess performance and safety.
Optimize training efficiency with quantization, distillation, LoRA/Q-LoRA, and memory optimizations.
Provide technical leadership and mentorship across engineering, research, and product teams, translating research into production-ready systems.
Design and implement scalable, distributed pipelines for SFT, RLHF, and instruction tuning on multi-node GPU clusters.
Oversee data strategy, annotation workflows, and automated data filtering to curate high-quality instruction data.
Build comprehensive evaluation suites, including automated benchmarks and human-in-the-loop protocols, to assess performance and safety.
Optimize training efficiency with quantization, distillation, LoRA/Q-LoRA, and memory optimizations.
Provide technical leadership and mentorship across engineering, research, and product teams, translating research into production-ready systems.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.