AI Safety Experts — English & Punjabi | $20-$22/hr Remote
Overview
We're looking for bilingual experts fluent in English and Punjabi to help stress-test and improve the safety of AI conversational models. In this remote, project-based role, you'll probe AI systems for hidden vulnerabilities—from jailbreaks to bias exploitation—and deliver actionable data that makes these models more trustworthy. This is hands-on adversarial testing with a direct impact on real-world AI safety.
What You'll Do4
- 1Attack conversational AI models and agents using techniques like prompt injection, jailbreaks, misuse cases, bias exploitation, and multi-turn manipulation to uncover weaknesses.
- 2Create high-quality human-generated data by annotating failure modes, classifying vulnerability types, and flagging systemic risks across model outputs.
- 3Follow structured playbooks, taxonomies, and benchmarks to ensure consistent testing coverage and reproducible results.
- 4Document findings with clear reports, datasets, and attack case studies that engineering and product teams can act on immediately.
Requirements4
- 1Native-level fluency in both English and Punjabi (written and verbal) is mandatory for this position.
- 2Proven experience in red teaming, adversarial AI testing, cybersecurity probing, or socio-technical vulnerability assessment.
- 3Ability to think like an adversary—curious, methodical, and comfortable pushing systems to their breaking points while following structured evaluation frameworks.
- 4Strong communication skills to explain technical risks clearly to both technical and non-technical stakeholders.
Who Should Apply
This role is ideal for someone who enjoys breaking things to make them better—whether your background is in adversarial ML, penetration testing, or creative probing (e.g., psychology, writing). You're bilingual (English & Punjabi), self-directed, and thrive on uncovering edge cases that automated tests miss. You care about AI safety and want to directly contribute to more robust, trustworthy systems.
Salary Insight
Pay ranges from $20.00 to $22.00 per hour, depending on experience and project scope.
Application Tip
In your application, explicitly list how your bilingual skills (English & Punjabi) enhance your red teaming work, and include a brief example of a vulnerability you've uncovered in an AI or conversational system—even if from a personal project.
Similar open positions
Explore active roles that match your skills and interests.
Mercor
VerifiedAI Safety Experts — English & Bengali | $20-$22/hr Remote
Neon is seeking bilingual (English & Bengali) AI safety experts to join a remote red teaming project. In this role, you'll proactively probe conversational AI models using adversarial techniques like jailbreaks and prompt injections to uncover vulnerabilities before they reach production. Your work directly contributes to making AI systems more robust and trustworthy.
Mercor
VerifiedAI Safety Experts — English & Assamese | $20-$22/hr Remote
We're assembling a remote team of AI red teamers to stress-test conversational models and uncover hidden vulnerabilities. In this role, you'll probe AI systems with adversarial inputs—jailbreaks, prompt injections, bias exploitation—and generate the high-quality human data that makes AI safer for real-world use. Fluency in both English and Assamese is essential, as you'll review sensitive outputs and document reproducible attack cases. This project is ideal for structured, curious thinkers who enjoy pushing systems to their limits.
Mercor
VerifiedAI Safety Experts — English & Odia | $20-$22/hr Remote
Mercor is looking for native English and Odia speakers to join a remote red team that stress-tests AI conversational models. You'll probe for vulnerabilities using adversarial techniques like jailbreaks and prompt injections, generating critical data that makes AI systems safer. This text-based project covers sensitive topics with clear guidelines and wellness support, allowing you to work on high-impact safety work from anywhere.
Mercor
VerifiedGeneralist - English & Punjabi | $15-$20/hr Remote
Join a team dedicated to refining AI communication by evaluating Punjabi-language model responses. In this remote contract role, you'll assess the quality, accuracy, and tone of AI-generated text, providing detailed feedback in English. Your analytical eye will help shape the next generation of conversational AI.
Mercor
VerifiedAI Safety Red Teamer | $70-$84/hr Remote
Join a team focused on strengthening the safety of advanced AI systems by uncovering their hidden weaknesses. As an AI Safety Red Teamer, you'll design clever prompts to stress-test models, spot dangerous behaviors, and help improve how these systems handle tricky real-world scenarios. This remote contractor role lets you work alongside top researchers while earning $70–$84 per hour.
Mercor
VerifiedCybersecurity SWE — AI Safety | $60-$90/hr Remote
This part-time remote role invites a cybersecurity engineer to apply their skills in AI safety — no prior machine learning background needed. You'll work hands-on evaluating how advanced AI models handle sensitive security topics, with full training on the workflow. It's a unique chance to blend deep security knowledge with the frontier of AI development.