Mercor
MercorVerified listing
Remote

AI Safety Experts — English & Punjabi | $20-$22/hr Remote

20–22/hr
Remote
Posted August 3, 2026
hourly
10 openings

Overview

Join a remote red team that stress-tests AI models for safety vulnerabilities. You will probe conversational AI with adversarial inputs, uncover weak spots like jailbreaks or bias, and turn those findings into data that makes AI safer for customers. Native-level fluency in English and Punjabi is a must. The work is fully text-based and paid hourly, with optional exposure to sensitive topics supported by clear guidelines.

What You'll Do4

  • 1Probe conversational AI models and agents for adversarial failure modes, including jailbreaks, prompt injections, misuse cases, bias exploitation, and multi-turn manipulation
  • 2Create high-quality red team data by labeling model failures, classifying vulnerability types, and flagging systemic risks
  • 3Apply structured testing methods using taxonomies, benchmarks, and playbooks to keep efforts consistent across projects
  • 4Document findings in clear reports, attack cases, and datasets that customers can directly use to harden their systems

Requirements5

  • 1Native or bilingual fluency in both English and Punjabi
  • 2Prior experience in red teaming, AI adversarial testing, cybersecurity, or socio-technical probing
  • 3Familiarity with structured frameworks or benchmarks for vulnerability testing, not just ad hoc probing
  • 4Strong ability to communicate technical risks clearly to technical and non-technical stakeholders
  • 5Comfort adapting across different projects and customer environments

Who Should Apply

You are naturally curious and enjoy breaking things. You think like an attacker but work like a scientist, using structured methods instead of random attempts. You are comfortable with text-based work on sensitive topics and can switch between projects easily. Prior experience in adversarial machine learning, penetration testing, or abuse analysis is a strong plus.

Salary Insight

Hourly rate of $20.00 to $22.00, paid for remote work.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

englishpunjabired teamingadversarial machine learningjailbreakingprompt injectionbias testingmulti-turn manipulationtaxonomy developmentbenchmark testingcybersecuritypenetration testingexploit developmentreverse engineeringrlhfdpomodel extractionabuse analysisconversational ai

Application Tip

Show exactly how you have probed AI systems in the past. Include a specific jailbreak or prompt injection you uncovered, and explain how you documented the vulnerability so a team could fix it.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Bengali

This role brings fluent English and Bengali speakers into a remote AI red team. You'll probe conversational AI systems with adversarial inputs to uncover biases, misinformation, and manipulation risks. All work is text-based, and you'll produce structured red team data that improves AI safety. You can opt out of higher-sensitivity topics, and clear guidelines plus wellness resources are provided.

20–22/hr
· 10 openings
EnglishBengaliConversational AI+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Odia

We're assembling a red team of human experts to stress-test conversational AI models for safety vulnerabilities. This remote, text-based position requires native fluency in English and Odia. You'll attack AI systems with adversarial inputs, uncover weaknesses, and generate the high-quality data needed to make AI safer. The work covers sensitive topics like bias and misinformation, with clear guidelines and wellness support available.

20–22/hr
· 10 openings
Red TeamingPrompt InjectionJailbreaking+13 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Malay

This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

17–25/hr
· 20 openings
Red TeamingAdversarial Machine LearningPrompt Injection+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Finnish

We're assembling a specialized team to stress-test AI systems by simulating adversarial attacks. As an AI Safety Expert, you'll probe conversational models for vulnerabilities like bias, misinformation, and harmful behaviors — all while working remotely on an hourly basis. Native fluency in both English and Finnish is essential for this role.

48–62/hr
· 5 openings
Adversarial Machine LearningPrompt InjectionJailbreaking+8 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Norwegian

This remote role invites fluent English and Norwegian speakers to join an elite red team probing AI models for vulnerabilities. You'll simulate adversarial attacks, document exploits, and generate high-quality human data that directly strengthens AI safety. The work is text-based and focuses on sensitive topics like bias and misinformation, with optional high-sensitivity projects supported by clear guidelines.

48–62/hr
· 5 openings
EnglishNorwegianRed Teaming+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Danish

We're recruiting bilingual AI safety specialists (English & Danish) to join a red team that systematically probes conversational AI models for vulnerabilities. You'll generate adversarial data, surface hidden risks, and create actionable documentation to strengthen AI systems. This remote, hourly contract pays $48–$62/hr based on experience.

48–62/hr
· 5 openings
Red TeamingAdversarial MLPrompt Injection+12 more