Senior Machine Learning Engineer, Model Customization, Generative AI Innovation Center
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
Track your applications on mobile The free Whileresume app, on iPhone and Android.
The role
Design and implement distributed training pipelines for LLMs using Fully Sharded Data Parallel (FSDP) and DeepSpeed to ensure scalable, efficient training.
Adapt LLMs for new languages, domains, and vision tasks through continued pre-training, fine-tuning, and RLHF.
Optimize models for deployment on custom accelerators by leveraging accelerator SDKs and developing kernel optimizations for performance.
Collaborate with enterprise customers and foundation model providers to understand their business and technical challenges and co-create tailored generative AI solutions.
Own end-to-end training pipelines at massive scale and drive improvements across multilingual and multimodal use cases.
Work with cross-functional teams of scientists, engineers, and architects to deliver production-ready models and scalable AI capabilities.
Adapt LLMs for new languages, domains, and vision tasks through continued pre-training, fine-tuning, and RLHF.
Optimize models for deployment on custom accelerators by leveraging accelerator SDKs and developing kernel optimizations for performance.
Collaborate with enterprise customers and foundation model providers to understand their business and technical challenges and co-create tailored generative AI solutions.
Own end-to-end training pipelines at massive scale and drive improvements across multilingual and multimodal use cases.
Work with cross-functional teams of scientists, engineers, and architects to deliver production-ready models and scalable AI capabilities.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inYou might also like these jobs
No closely matching jobs yet — here are the most recent ones.