Mercor
MercorVerified listing
RemoteAI & Machine LearningCybersecurity

AI Safety Experts — English & Vietnamese | $17-$25/hr Remote

17–25/hr
Remote
Posted August 3, 2026
hourly
5 openings

Overview

We're hiring bilingual (English & Vietnamese) red teamers to stress-test AI models by probing for vulnerabilities such as jailbreaks, bias, and misinformation. In this remote, contract role, you'll generate adversarial data and document findings to help make AI systems safer. It's a fit for anyone with a knack for pushing systems to their limits and a structured approach to testing.

What You'll Do4

  • 1Conduct adversarial testing on conversational AI models, including jailbreaks, prompt injections, and multi-turn manipulation to uncover weaknesses.
  • 2Produce high-quality human annotations by classifying failures, flagging systemic risks, and labeling vulnerabilities in AI outputs.
  • 3Follow established taxonomies, benchmarks, and playbooks to ensure consistent and repeatable testing across projects.
  • 4Create detailed reports, datasets, and attack cases that customers can directly use to strengthen their AI systems.

Requirements5

  • 1Prior experience in AI red teaming, adversarial machine learning, or cybersecurity penetration testing.
  • 2A naturally curious and adversarial mindset — you enjoy probing systems to find their breaking points.
  • 3Ability to structure your work using frameworks or benchmarks rather than ad-hoc testing.
  • 4Strong communication skills to explain complex risks clearly to both technical and non-technical audiences.
  • 5Fluency in both English and Vietnamese at a native level.

Who Should Apply

This role is ideal for security researchers, red teamers, or AI safety specialists who are fluent in English and Vietnamese. You should be comfortable working autonomously on adversarial tasks, have a systematic approach to testing, and enjoy collaborating across projects. If you've done work with jailbreaks, prompt injection, or socio-technical risk probing, you'll fit right in.

Salary Insight

The role offers $17.00 - $25.00 per hour, paid on an hourly basis. No other compensation details were provided.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

red teamingadversarial machine learningprompt injectionjailbreakrlhfdpomodel extractionpenetration testingexploit developmentreverse engineeringconversational aibias detectionmisinformation analysisdata annotationenglishvietnamese

Application Tip

When applying, highlight specific examples of adversarial testing you've done, especially with conversational AI models. Mention any experience with prompt injection or jailbreak techniques, and be ready to discuss how you document vulnerabilities in a reproducible way.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Thai

Neon is assembling a remote team of bilingual (English & Thai) AI safety specialists to stress-test conversational AI systems. You'll probe models for vulnerabilities like bias, misinformation, and prompt injection, then document and classify findings to help make AI safer. This text-based red teaming role offers flexible hourly work and the chance to shape AI safety at the frontier.

24–35/hr
· 5 openings
Red TeamingAdversarial MLPrompt Injection+10 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Malay

This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

17–25/hr
· 20 openings
Red TeamingAdversarial Machine LearningPrompt Injection+14 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Indonesian

Neon is looking for bilingual AI safety experts fluent in English and Indonesian to join a human-driven red team. You'll probe advanced conversational AI models for vulnerabilities like bias, misinformation, and harmful behavior — all through text-based work. This remote, hourly role pays $17–$25/hr and gives you a direct hand in making AI systems safer before they reach the public.

17–25/hr
· 5 openings
Adversarial Machine LearningRed TeamingPrompt Injection+13 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Bengali

This role brings fluent English and Bengali speakers into a remote AI red team. You'll probe conversational AI systems with adversarial inputs to uncover biases, misinformation, and manipulation risks. All work is text-based, and you'll produce structured red team data that improves AI safety. You can opt out of higher-sensitivity topics, and clear guidelines plus wellness resources are provided.

20–22/hr
· 10 openings
EnglishBengaliConversational AI+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Dutch

We're looking for bilingual English-Dutch AI red teamers to stress-test conversational AI models from every angle. In this remote, hourly role, you'll design and execute adversarial attacks—from jailbreaks to bias exploitation—and generate structured data that helps make AI systems safer. You'll work with clear guidelines and optional high-sensitivity projects, with wellness support available.

48–62/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+12 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Finnish

We're assembling a specialized team to stress-test AI systems by simulating adversarial attacks. As an AI Safety Expert, you'll probe conversational models for vulnerabilities like bias, misinformation, and harmful behaviors — all while working remotely on an hourly basis. Native fluency in both English and Finnish is essential for this role.

48–62/hr
· 5 openings
Adversarial Machine LearningPrompt InjectionJailbreaking+8 more