Mercor
MercorVerified listing
Remote

AI Safety Experts — English & Bengali | $20-$22/hr Remote

20–22/hr
Remote
Posted August 5, 2026
hourly
10 openings

Overview

This role brings fluent English and Bengali speakers into a remote AI red team. You'll probe conversational AI systems with adversarial inputs to uncover biases, misinformation, and manipulation risks. All work is text-based, and you'll produce structured red team data that improves AI safety. You can opt out of higher-sensitivity topics, and clear guidelines plus wellness resources are provided.

What You'll Do4

  • 1Pressure-test conversational AI models using jailbreak techniques, prompt injection, and bias exploitation to surface failure modes.
  • 2Create high-quality human data by labeling model failures, categorizing vulnerability types, and flagging systemic risks.
  • 3Apply established taxonomies, benchmarks, and playbooks to keep your adversarial testing structured and repeatable.
  • 4Document findings in clear reports and prepare attack-case datasets that customers can act on.

Requirements5

  • 1Native-level fluency in English and Bengali for written and spoken communication.
  • 2Hands-on experience with adversarial AI testing, cybersecurity, or socio-technical risk probing.
  • 3A systematic mindset: you use frameworks, benchmarks, or playbooks rather than random testing.
  • 4Strong ability to explain technical risks to both technical and non-technical audiences.
  • 5Adaptability to work across diverse projects and customer environments.

Who Should Apply

You have a natural curiosity for how systems fail and a structured approach to probing them. Your background may include adversarial machine learning, penetration testing, abuse analysis, or creative fields like psychology and writing. You're comfortable handling sensitive content and following established guidelines. You want a remote, hourly role where your adversarial thinking shapes safer AI.

Salary Insight

Pay is $20.00 - $22.00 per hour for this remote, hourly engagement.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

englishbengaliconversational aired teamingprompt injectionjailbreakingbias exploitationadversarial machine learningrlhfdpomodel extractionpenetration testingexploit developmentreverse engineeringsocio-technical risk analysisabuse analysiscybersecurity

Application Tip

In your application, describe one instance where you successfully jailbroke or exposed a flaw in an AI system. Explain the approach you took and what you learned.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Odia

We're assembling a red team of human experts to stress-test conversational AI models for safety vulnerabilities. This remote, text-based position requires native fluency in English and Odia. You'll attack AI systems with adversarial inputs, uncover weaknesses, and generate the high-quality data needed to make AI safer. The work covers sensitive topics like bias and misinformation, with clear guidelines and wellness support available.

20–22/hr
· 10 openings
Red TeamingPrompt InjectionJailbreaking+13 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Punjabi

Join a remote red team that stress-tests AI models for safety vulnerabilities. You will probe conversational AI with adversarial inputs, uncover weak spots like jailbreaks or bias, and turn those findings into data that makes AI safer for customers. Native-level fluency in English and Punjabi is a must. The work is fully text-based and paid hourly, with optional exposure to sensitive topics supported by clear guidelines.

20–22/hr
· 10 openings
EnglishPunjabiRed Teaming+16 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Malay

This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

17–25/hr
· 20 openings
Red TeamingAdversarial Machine LearningPrompt Injection+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Assamese

Mercor is assembling a remote red team of human experts to attack AI models with adversarial inputs. This role focuses on conversational AI, using native English and Assamese to probe for jailbreaks, prompt injections, and bias exploitation. You'll generate red team data and vulnerability reports that make AI systems safer for customers. The work is entirely text-based, with optional participation in higher-sensitivity projects supported by clear guidelines and wellness resources. Hourly compensation is $20-$22.

20–22/hr
· 10 openings
Red TeamingAdversarial Machine LearningJailbreak+16 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Norwegian

This remote role invites fluent English and Norwegian speakers to join an elite red team probing AI models for vulnerabilities. You'll simulate adversarial attacks, document exploits, and generate high-quality human data that directly strengthens AI safety. The work is text-based and focuses on sensitive topics like bias and misinformation, with optional high-sensitivity projects supported by clear guidelines.

48–62/hr
· 5 openings
EnglishNorwegianRed Teaming+14 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Swedish

This role involves stress-testing AI systems by simulating adversarial attacks in English and Swedish. You'll generate critical safety data by probing models for vulnerabilities like bias, misinformation, and security flaws. As part of a human red team, you'll help ensure AI behaves safely before it reaches users.

48–62/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+12 more