Open bounties

AI Safety & Red Team Evaluator — Short-Term Sprint

Paid remote expert bounty in Global (Remote) Experts (Global (Remote)). Work in the software you already know; rates and ladder published up front.

$30-$250
per hr, published up front
1
slots left of 1
Remote
no AI experience required
Cybersecurity

What you'll be paid to show

About the role

We're assembling a focused cohort of domain experts to support a fast-paced AI safety evaluation sprint. This is a short-term, high-intensity engagement designed for professionals who can quickly assess model behavior, identify risks, and apply structured judgment under time pressure. The work centers on stress-testing AI systems for safety, robustness, and policy alignment through hands-on red teaming and adversarial evaluation.

This is an ideal opportunity for specialists in trust & safety, cybersecurity, content moderation, or AI policy who want meaningful, well-compensated project work without a long-term commitment.

What you'll do

  • Conduct structured red-teaming and adversarial testing sessions against AI model outputs
  • Evaluate responses for safety violations, policy breaches, and harmful edge cases
  • Document findings with precision, using strong written English and clear reasoning
  • Apply content moderation and trust & safety judgment to ambiguous or adversarial scenarios
  • Collaborate closely with a coordination team during a compressed, fast-moving sprint window
  • Maintain reliable, responsive availability across consecutive working days to keep pace with the sprint schedule

Requirements

  • Demonstrated experience in AI safety, red teaming, cybersecurity, trust & safety, content moderation, policy evaluation, or adversarial testing
  • Strong written English communication and meticulous attention to detail
  • Ability to work independently while staying responsive and coordinated with the broader team
  • Comfort operating in a fast-paced, deadline-driven sprint environment
  • Reliable internet connection and availability for concentrated blocks of time over consecutive days

Compensation

This engagement pays a flat hourly rate of **$45/hour**, with total hours determined by sprint duration and scope. Payment is processed through the platform upon verified completion of assigned sessions.

What we're looking for

  • AI Safety
  • Red Teaming
  • Adversarial Testing
  • Trust & Safety
  • Content Moderation
  • Policy Evaluation
  • Cybersecurity