1wk ago - Alice is hiring a remote Research Lead, Evaluations and Benchmarks. ๐ธ Salary: unspecified ๐Location: USA
Role Description
You ship a benchmark every two to three weeks. Each one measures a frontier risk that nobody has measured yet. Some go public, while others go only to the labs. Some benchmarks and papers are done in collaboration with leading AI Labs and universities.
You will not write every eval yourself. Each benchmark pairs you with an in-house researcher who owns that harm area, and you get a budget for freelancers you direct. You own the taxonomy, the harness, the quality bar, and the release.
The seat sits in the CTO office alongside the research lead who sets our public research agenda. Around 150 researchers here work on harms directly, and you can pull any of them onto a subject.
Qualifications
Requirements
Benefits
Company Description
Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interactโwhether with each other or with machines.
If you're creative and driven to secure the future of AI, we want to hear from you!
| ๐บ๐ธ | Be aware of the location restriction for this remote position: USA Only |
| โผ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | ๏ธ
| ๐บ๐ธ | Be aware of the location restriction for this remote position: USA Only |
| โผ | Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more. | ๏ธ
Access 125,000+ vetted remote jobs and get daily alerts.