Senior Research Scientist, Model Evaluation
Cohere UK Ltd · Toronto · Remote · posted 323 days ago
Going rate £43,600UK median £44,580
Home Office going rates from
Occupation
2162Other researchers, unspecified discipline
Going rate for this occupation: £43,600 · UK median pay £44,580
Home Office going rates from
Where this salary sits
- UK pay for this occupation
- This role£42,283 to £49,174estimated · above the range ONS published
- Going rate£43,600
- UK median£44,580
View these figures as a table
| Percentile | Pay |
|---|---|
| 25th | £38,838 |
| 50th | £44,580 |
| Going rate | £43,600 |
| UK median | £44,580 |
Sponsorship
Sponsorship chance
High
- Licensed for Skilled Worker
- Occupation is eligible for Skilled Worker
- Estimated salary is below the going rate
On the public records we hold, sponsorship for this role looks likely: licensed for Skilled Worker, and occupation is eligible for Skilled Worker.
Sign in to see how you fit and draft a cover letter
We read this advert for what it asks, check each line against your CV and show you the evidence for every judgement.
Sign inFull advert
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Role Overview: Evaluation is critical to making progress in scaling intelligence. As models continue to become superhuman in many real-world use cases, we must continue to develop new evaluation techniques that accurately reflect what models are already capable of, as well as set the agenda for what future models should be capable of. In this role, you are responsible for creating these next-generation evaluation methods and infrastructure to measure LLM progress. Key Responsibilities: Create ambitious new evaluation benchmarks that push the limits of what our models can accomplish. Work on highly cross-functional teams to translate model feedback into trustworthy, repeatable evaluations. Conduct research to advance the state-of-the-art in LLM evaluation methods, including training LLM judges; refining LLM-based data synthesis pipelines; and improving evaluation efficiency. Build scalable and reusable tools for digging into model performance. Qualifications: You enjoy rapidly building prototypes that demonstrate the boundaries of what LLMs are capable of, and you have developed resources to measure those capabilities. You have spent dozens of hours reviewing complex data and LLM outputs to ensure high data quality. You are obsessive about rigorously measuring AI capabilities, and also about making sure your measurements actually align with the capabilities you care about. You have strong software engineering skills. Working Location: This role can be based remotely or from one of our office locations listed on the job description - there is no minimum in-office qualification requirement. We care most about hiring exceptional people regardless of locations, though please check the location listed on the posting for guidance around the core time zone or working hours alignment expected for the role.
Information from public records, not immigration advice.