AI Safety Expert @Weekday AI
Artificial Intelligence
Salary $48 - $62 per h..
Remote Location
๐Ÿ‡บ๐Ÿ‡ธ USA Only
Employment Type contract
Posted 1wk ago

[Hiring] AI Safety Expert @Weekday AI

1wk ago - Weekday AI is hiring a remote AI Safety Expert. ๐Ÿ’ธ Salary: $48 - $62 per hour ๐Ÿ“Location: USA

Role Description

This role is for one of our clients.

Compensation: $48 - $62 per hour

Fluent Language Skills Required: English & Finnish. Native fluency in English and Finnish is required for this position.

Why This Role Exists:

  • We are assembling a red team for this project - human data experts who probe AI models with adversarial inputs, surface vulnerabilities, and generate the red team data that makes AI safer for our customers.
  • This project involves reviewing AI outputs that touch on sensitive topics such as bias, misinformation, or harmful behaviors.
  • All work is text-based, and participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources.
  • Before being exposed to any content, the topics will be clearly communicated.

What Youโ€™ll Do:

  • Red team conversational AI models and agents: jailbreaks, prompt injections, misuse cases, bias exploitation, multi-turn manipulation.
  • Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure: follow taxonomies, benchmarks, and playbooks to keep testing consistent.
  • Document reproducibly: produce reports, datasets, and attack cases customers can act on.

Qualifications

  • You bring prior red teaming experience (AI adversarial work, cybersecurity, socio-technical probing).
  • Youโ€™re curious and adversarial: you instinctively push systems to breaking points.
  • Youโ€™re structured: you use frameworks or benchmarks, not just random hacks.
  • Youโ€™re communicative: you explain risks clearly to technical and non-technical stakeholders.
  • Youโ€™re adaptable: thrive on moving across projects and customers.

Nice-to-Have Specialties

  • Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
  • Cybersecurity: penetration testing, exploit development, reverse engineering.
  • Socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
  • Creative probing: psychology, acting, writing for unconventional adversarial thinking.

What Success Looks Like

  • You uncover vulnerabilities automated tests miss.
  • You deliver reproducible artifacts that strengthen customer AI systems.
  • Evaluation coverage expands: more scenarios tested, fewer surprises in production.
Before You Apply
๏ธ
๐Ÿ‡บ๐Ÿ‡ธ Be aware of the location restriction for this remote position: USA Only
โ€ผ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
AI Safety Expert @Weekday AI
Artificial Intelligence
Salary $48 - $62 per h..
Remote Location
๐Ÿ‡บ๐Ÿ‡ธ USA Only
Employment Type contract
Posted 1wk ago
Apply for this position
Did not apply โœ“
Applied โœ“
Sent Follow-Up โœ“
Interview Scheduled โœ“
Interview Completed โœ“
Offer Accepted โœ“
Offer Declined โœ“
Application Denied โœ“
Unlock 125,000+ Remote Jobs
๏ธ
๐Ÿ‡บ๐Ÿ‡ธ Be aware of the location restriction for this remote position: USA Only
โ€ผ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply โœ“
Applied โœ“
Sent Follow-Up โœ“
Interview Scheduled โœ“
Interview Completed โœ“
Offer Accepted โœ“
Offer Declined โœ“
Application Denied โœ“
Unlock 125,000+ Remote Jobs
ร—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 โ˜…โ˜…โ˜…โ˜…โ˜… from 500+ reviews

โšก 129,706+ remote jobs, refreshed hourly

๐Ÿ”” Real-time alerts: Apply first

๐Ÿ›ก๏ธ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later