Applied AI Engineer - Multimodal Transformers
职位介绍
Design and develop multimodal transformer architectures that fuse camera, LiDAR, and radar into unified representations.
Advance cross-modal attention, token fusion strategies, and efficient multi-stream tokenization for robust sensor fusion.
Build scalable training pipelines for large-scale multimodal transformers across real-world datasets.
Explore self-supervised and contrastive pretraining objectives to learn transferable multimodal representations.
Optimize transformer models for real-time inference under latency and compute constraints while prioritizing safety.
Collaborate with cross-functional teams and contribute to cutting-edge research and scalable deployment.
Advance cross-modal attention, token fusion strategies, and efficient multi-stream tokenization for robust sensor fusion.
Build scalable training pipelines for large-scale multimodal transformers across real-world datasets.
Explore self-supervised and contrastive pretraining objectives to learn transferable multimodal representations.
Optimize transformer models for real-time inference under latency and compute constraints while prioritizing safety.
Collaborate with cross-functional teams and contribute to cutting-edge research and scalable deployment.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录相似职位
其他可能适合您的职位。
城市Mountain View, United States