Mercor
MercorVerified listing
RemoteAI & Machine LearningCybersecurity

AI Safety Experts — English & Thai | $24-$35/hr Remote

24–35/hr
Remote
Posted August 5, 2026
hourly
5 openings

Overview

Neon is assembling a remote team of bilingual (English & Thai) AI safety specialists to stress-test conversational AI systems. You'll probe models for vulnerabilities like bias, misinformation, and prompt injection, then document and classify findings to help make AI safer. This text-based red teaming role offers flexible hourly work and the chance to shape AI safety at the frontier.

What You'll Do4

  • 1Run adversarial tests on conversational AI agents, including jailbreaks, prompt injections, misuse cases, and multi-turn manipulation attempts.
  • 2Label and categorize AI failures, classifying vulnerabilities and flagging systemic risks to improve model robustness.
  • 3Follow structured taxonomies, benchmarks, and testing playbooks to ensure consistent evaluation across different scenarios.
  • 4Produce detailed reports, datasets, and reproducible attack cases that customers can act on to strengthen their AI systems.

Requirements4

  • 1Native-level fluency in both English and Thai (spoken and written).
  • 2Prior hands-on experience in red teaming AI models, adversarial cybersecurity testing, or socio-technical probing.
  • 3Strong analytical thinking and a methodical approach to breaking systems, using frameworks rather than random tests.
  • 4Excellent communication skills to explain complex risks clearly to both technical and non-technical stakeholders.

Who Should Apply

This role is ideal for professionals with a curious, adversarial mindset who enjoy stress-testing AI systems to uncover hidden flaws. You should be comfortable working independently, able to follow structured testing protocols, and passionate about improving AI safety. Experience with prompt injection, bias probing, or penetration testing is a strong plus. Bilingual English-Thai speakers from cybersecurity, AI research, or creative adversarial backgrounds are especially welcome.

Salary Insight

Pay ranges from $24.00 to $35.00 per hour, depending on experience and project scope.

Location

Typeremote
LocationRemote
This is a remote position

Required Skills

red teamingadversarial mlprompt injectionjailbreakrlhfdpopenetration testingexploit developmentreverse engineeringsocio-technical risk analysisconversational aithaienglish

Application Tip

When applying, prepare a brief portfolio or description of past red teaming projects, especially those involving multilingual or conversational AI systems—real examples of jailbreaks or vulnerability reports will make your application stand out.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Indonesian

Neon is looking for bilingual AI safety experts fluent in English and Indonesian to join a human-driven red team. You'll probe advanced conversational AI models for vulnerabilities like bias, misinformation, and harmful behavior — all through text-based work. This remote, hourly role pays $17–$25/hr and gives you a direct hand in making AI systems safer before they reach the public.

17–25/hr
· 5 openings
Adversarial Machine LearningRed TeamingPrompt Injection+13 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Vietnamese

We're hiring bilingual (English & Vietnamese) red teamers to stress-test AI models by probing for vulnerabilities such as jailbreaks, bias, and misinformation. In this remote, contract role, you'll generate adversarial data and document findings to help make AI systems safer. It's a fit for anyone with a knack for pushing systems to their limits and a structured approach to testing.

17–25/hr
· 5 openings
Red TeamingAdversarial Machine LearningPrompt Injection+13 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Malay

This remote role asks bilingual experts in English and Malay to stress-test AI systems by simulating attacks and probing for weak points. You'll generate high-quality red team data that helps customers build safer, more trustworthy models. All work is text-based, and sensitive topics are clearly disclosed with wellness support.

17–25/hr
· 20 openings
Red TeamingAdversarial Machine LearningPrompt Injection+14 more
Mercor

Mercor

30d agoRemotehourly

AI Safety Experts — English & Bengali

This role brings fluent English and Bengali speakers into a remote AI red team. You'll probe conversational AI systems with adversarial inputs to uncover biases, misinformation, and manipulation risks. All work is text-based, and you'll produce structured red team data that improves AI safety. You can opt out of higher-sensitivity topics, and clear guidelines plus wellness resources are provided.

20–22/hr
· 10 openings
EnglishBengaliConversational AI+14 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Portuguese (global)

Neon is assembling a distributed red team to stress-test conversational AI models by probing for vulnerabilities like jailbreaks, prompt injections, and bias exploits. This remote, hourly role invites bilingual experts (fluent in English and global Portuguese, excluding Brazilian variants) to generate high-quality adversarial data that strengthens AI safety. Work is text-based, with optional exposure to sensitive topics supported by clear guidelines and wellness resources.

29–45/hr
· 5 openings
Red TeamingAdversarial Machine LearningCybersecurity+15 more
Mercor

Mercor

1mo agoRemotehourly

AI Safety Experts — English & Norwegian

This remote role invites fluent English and Norwegian speakers to join an elite red team probing AI models for vulnerabilities. You'll simulate adversarial attacks, document exploits, and generate high-quality human data that directly strengthens AI safety. The work is text-based and focuses on sensitive topics like bias and misinformation, with optional high-sensitivity projects supported by clear guidelines.

48–62/hr
· 5 openings
EnglishNorwegianRed Teaming+14 more