Research Scientist, Web Data
在手机上跟进您的申请 Whileresume 免费应用,支持 iPhone 和 Android。
职位介绍
Own and lead improvements to meaningful chunks of the web data pipeline, including scraping raw HTML into clean text data and integrating relevant image data.
Develop and apply measurements of weakness, via model evals or data pipeline statistics, to drive progress.
Define a medium-term agenda to improve data quality and pipeline efficiency, and build consensus with peers and stakeholders.
Collaborate with partner teams to leverage existing solutions and communicate necessary infrastructure improvements.
Execute day-to-day work through coding, running experiments, and reviewing contributions.
Qualifications include at least 3 years of self-directed work, building large-scale data pipelines (>=100M examples) in Python and/or C++, and evaluating pretrained LLMs.
Develop and apply measurements of weakness, via model evals or data pipeline statistics, to drive progress.
Define a medium-term agenda to improve data quality and pipeline efficiency, and build consensus with peers and stakeholders.
Collaborate with partner teams to leverage existing solutions and communicate necessary infrastructure improvements.
Execute day-to-day work through coding, running experiments, and reviewing contributions.
Qualifications include at least 3 years of self-directed work, building large-scale data pipelines (>=100M examples) in Python and/or C++, and evaluating pretrained LLMs.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录您可能也感兴趣的职位
暂无高度相似的职位 — 以下是最新职位。