本文へスキップ

Research Scientist, Frontier, Zurich

この求人はあなたに合っていますか?

履歴書を作成すると、この求人との適合度が表示されます。ほかのすべての求人についても同様です。

履歴書を作る

仕事内容

Design and validate novel post-training pipelines (SFT, RLHF, RLAIF) for frontier-class models where no teacher model exists.
Lead research into next-generation Reward Models, exploring architectures, reducing reward hacking, and improving data signal-to-noise ratios.
Develop methods to enhance internal reasoning and Chain-of-Thought capabilities, emphasizing correctness and multi-step problem solving.
Revamp RL paradigms and prompts to maximize performance while maintaining alignment and safety.
Create robust mechanisms to transform user signals into training data, building a flywheel that scales without regression or bias.
Collaborate across teams to apply these recipes to various sizes and modalities, including audio and multimodal tasks.

求人の全文を見る

業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。

8文字以上、大文字と数字をそれぞれ1つ以上含めてください。

すでにアカウントをお持ちですか? ログイン

類似の求人

あなたに合いそうな他の職種。

すべて見る →

あなたの地域

求人と企業がこの国で絞り込まれます。

おすすめ

すべての国 69