Research Scientist, Interpretability
Is this job for you?
Build my CV Build your CV and see how well you match this role — and every other one.
The role
Develop methods for understanding LLMs by reverse engineering the algorithms learned in their weights.
Design and run robust experiments, both quickly in toy scenarios and at scale in large models.
Create and analyze new interpretability features and circuits to better understand how models work.
Build infrastructure for running experiments and visualizing results.
Work with colleagues to communicate results internally and publicly.
You will thrive if you have a strong research track record, enjoy team science, are comfortable with messy experimental science, and can bridge research and engineering.
Design and run robust experiments, both quickly in toy scenarios and at scale in large models.
Create and analyze new interpretability features and circuits to better understand how models work.
Build infrastructure for running experiments and visualizing results.
Work with colleagues to communicate results internally and publicly.
You will thrive if you have a strong research track record, enjoy team science, are comfortable with messy experimental science, and can bridge research and engineering.
See the full job post
Responsibilities, requirements, skills and benefits — create your free account.
Already have an account? Log in
Similar openings
Other roles that could suit you.
Remote workPartial
CitySan Francisco, United States