Position: AI Safety Experts English & Bengali
Type: Contract
Compensation: $20-$22 per hour
Location: Remote
Commitment: 10-40 hrs/week
Role Responsibilities- Red team conversational AI models and agents by conducting jailbreaks, prompt injections, and bias exploitation.
- Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structured methodologies by following taxonomies, benchmarks, and playbooks to ensure consistent testing.
- Document findings reproducibly by producing reports, datasets, and attack cases for customer action.
Requirements- Have prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
- Demonstrate curiosity and an adversarial mindset, pushing systems to their breaking points.
- Be structured in their approach, utilizing frameworks or benchmarks rather than random hacks.
- Possess strong communication skills to explain risks clearly to both technical and non-technical stakeholders.
- Be adaptable and thrive on moving across various projects and customers.