Applied AI Engineer - Multimodal Transformers
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Design and develop multimodal transformer architectures that fuse camera, LiDAR, and radar into unified representations.
Advance cross-modal attention, token fusion strategies, and efficient multi-stream tokenization for robust sensor fusion.
Build scalable training pipelines for large-scale multimodal transformers across real-world datasets.
Explore self-supervised and contrastive pretraining objectives to learn transferable multimodal representations.
Optimize transformer models for real-time inference under latency and compute constraints while prioritizing safety.
Collaborate with cross-functional teams and contribute to cutting-edge research and scalable deployment.
Advance cross-modal attention, token fusion strategies, and efficient multi-stream tokenization for robust sensor fusion.
Build scalable training pipelines for large-scale multimodal transformers across real-world datasets.
Explore self-supervised and contrastive pretraining objectives to learn transferable multimodal representations.
Optimize transformer models for real-time inference under latency and compute constraints while prioritizing safety.
Collaborate with cross-functional teams and contribute to cutting-edge research and scalable deployment.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
or
Already have an account?
Log inSimilar openings
Other roles that could suit you.
CityMountain View, United States